Data Engineer with AI skills
Capgemini Bogota, RAP (Especial) Central, Colombia
IT Services and IT Consulting · 10,001+ employees
About the role
Design and build scalable data solutions to power analytics, AI applications, and RAG ecosystems. Develop data pipelines and agentic solutions while ensuring data governance, quality, and security.
What they look for
Requirements
Requires strong proficiency in Python, SQL, and the Azure data stack including Databricks and Fabric. Candidates must have experience with Generative AI frameworks, LLMs, and DataOps practices.
Full description
At Capgemini Engineering, the world leader in engineering services, we bring together a global team of engineers, scientists, and architects to help the world’s most innovative companies unleash their potential. From autonomous cars to life-saving robots, our digital and software technology experts think outside the box as they provide unique R&D and engineering services across all industries. Join us for a career full of opportunities. Where you can make a difference. Where no two days are the same.
Job Description
Design and build scalable data solutions that power analytics, AI applications, agents, and RAG ecosystems while ensuring data governance, quality, and security.
Key Responsibilities
- Develop and maintain scalable data pipelines and data products.
- Curate and prepare enterprise data for AI, GenAI, agent-based applications, and RAG solutions.
- Ensure data quality, lineage, metadata management, governance, and security standards.
- Build agentic solutions to automate data discovery, transformation, and enrichment processes.
- Collaborate with Data Science, AI, and business teams to enable production-ready AI use cases.
Technical Requirements
- Strong experience with Python and SQL.
- Hands-on experience with Azure Data Factory, Databricks, Synapse Analytics, and Fabric.
- Knowledge of Data Lakehouse architectures and ETL/ELT frameworks.
- Experience with Generative AI, LLMs, RAG architectures, vector databases, and AI agents.
- Familiarity with Azure OpenAI, LangChain, Semantic Kernel, or similar AI frameworks.
- Understanding of data governance, lineage, metadata management, and data security.
- Experience with Git, CI/CD, and DataOps practices.
#LI-DC10
#LI-Remote
Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.
Similar roles
-
Staff Data Engineer
Overstory Canada
-
Lead Data Engineer
EXL United Kingdom
-
Data Engineer H/F
NEXTON Paris, Ile-de-France, France
-
Senior Data Engineer
EDF UK London, England, United Kingdom
-
Data Engineer
EDF UK London, England, United Kingdom
-
Senior Data Engineer ( M/F/D )
EVERIENCE Brussels, Brussels-Capital, Belgium