
at Capgemini
ConsultanciesPosted 13 days ago
No clicks
**Data Engineer with AI Skills @ Capgemini Engineering** Design, build, and maintain scalable data pipelines and products, empowering AI, GenAI, agent-based, and RAG applications. Key responsibilities include data curation, quality assurance, metadata management, and collaboration with data science and business teams. Required skills: Python, SQL, Azure Data Factory, Databricks, Synapse Analytics, Fabric, Generative AI, LLMs, RAG architectures, Data Lakehouse, ETL/ELT, Git, CI/CD, DataOps. Experience with Azure OpenAI, LangChain, Semantic Kernel, vector databases, and AI agents is a plus. Join a global engineering services leader driving innovation across industries.
- Compensation
- Not specified
- City
- Not specified
- Country
- Not specified
Currency: Not specified
Full Job Description
At Capgemini Engineering, the world leader in engineering services, we bring together a global team of engineers, scientists, and architects to help the worlds most innovative companies unleash their potential. From autonomous cars to life-saving robots, our digital and software technology experts think outside the box as they provide unique R&D and engineering services across all industries. Join us for a career full of opportunities. Where you can make a difference. Where no two days are the same.
Job Description
Design and build scalable data solutions that power analytics, AI applications, agents, and RAG ecosystems while ensuring data governance, quality, and security.
Key Responsibilities
- Develop and maintain scalable data pipelines and data products.
- Curate and prepare enterprise data for AI, GenAI, agent-based applications, and RAG solutions.
- Ensure data quality, lineage, metadata management, governance, and security standards.
- Build agentic solutions to automate data discovery, transformation, and enrichment processes.
- Collaborate with Data Science, AI, and business teams to enable production-ready AI use cases.
Technical Requirements
- Strong experience with Python and SQL.
- Hands-on experience with Azure Data Factory, Databricks, Synapse Analytics, and Fabric.
- Knowledge of Data Lakehouse architectures and ETL/ELT frameworks.
- Experience with Generative AI, LLMs, RAG architectures, vector databases, and AI agents.
- Familiarity with Azure OpenAI, LangChain, Semantic Kernel, or similar AI frameworks.
- Understanding of data governance, lineage, metadata management, and data security.
- Experience with Git, CI/CD, and DataOps practices.
#LI-DC10
#LI-Remote
Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.



