Capgemini

Data Engineer with AI skills

Capgemini Bogota, RAP (Especial) Central, Colombia

IT Services and IT Consulting · 10,001+ employees

20 h ago
Remote data-engineer Mid (2-5 yrs) Full-time Colombia
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Design and build scalable data solutions to power analytics, AI applications, and RAG ecosystems. Develop data pipelines and agentic solutions while ensuring data governance, quality, and security.

What they look for

Python SQL Azure Data Factory Databricks Synapse Analytics Microsoft Fabric Data Lakehouse ETL/ELT Generative AI LLMs RAG Architectures Vector Databases AI Agents Azure OpenAI LangChain Semantic Kernel

Requirements

Requires strong proficiency in Python, SQL, and the Azure data stack including Databricks and Fabric. Candidates must have experience with Generative AI frameworks, LLMs, and DataOps practices.

Full description

At Capgemini Engineering, the world leader in engineering services, we bring together a global team of engineers, scientists, and architects to help the world’s most innovative companies unleash their potential. From autonomous cars to life-saving robots, our digital and software technology experts think outside the box as they provide unique R&D and engineering services across all industries. Join us for a career full of opportunities. Where you can make a difference. Where no two days are the same.

Job Description

Design and build scalable data solutions that power analytics, AI applications, agents, and RAG ecosystems while ensuring data governance, quality, and security.

Key Responsibilities

  • Develop and maintain scalable data pipelines and data products.
  • Curate and prepare enterprise data for AI, GenAI, agent-based applications, and RAG solutions.
  • Ensure data quality, lineage, metadata management, governance, and security standards.
  • Build agentic solutions to automate data discovery, transformation, and enrichment processes.
  • Collaborate with Data Science, AI, and business teams to enable production-ready AI use cases.

Technical Requirements

  • Strong experience with Python and SQL.
  • Hands-on experience with Azure Data Factory, Databricks, Synapse Analytics, and Fabric.
  • Knowledge of Data Lakehouse architectures and ETL/ELT frameworks.
  • Experience with Generative AI, LLMs, RAG architectures, vector databases, and AI agents.
  • Familiarity with Azure OpenAI, LangChain, Semantic Kernel, or similar AI frameworks.
  • Understanding of data governance, lineage, metadata management, and data security.
  • Experience with Git, CI/CD, and DataOps practices.

#LI-DC10

#LI-Remote

Capgemini is a global business and technology transformation partner, helping organizations to accelerate their dual transition to a digital and sustainable world, while creating tangible impact for enterprises and society. It is a responsible and diverse group of 340,000 team members in more than 50 countries. With its strong over 55-year heritage, Capgemini is trusted by its clients to unlock the value of technology to address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, generative AI, cloud and data, combined with its deep industry expertise and partner ecosystem.

Similar roles