Citi

Senior Data Engineer - Assistant Vice President

Citi Chennai, Tamil Nadu, India

Financial Services · 10,001+ employees

21 h ago
data-engineer Principal (10+ yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Design, develop, and maintain robust ETL/ELT pipelines while contributing to the architecture of scalable data lakes and warehouses. Implement data quality frameworks and optimize system performance across cloud platforms using tools like Spark and Kafka.

What they look for

Python SQL Apache Spark Hadoop ETL/ELT Pipelines Data Architecture Data Lakes Data Warehousing Lakehouse Patterns NoSQL Machine Learning Artificial Intelligence Distributed Computing Data Quality Kafka Databricks

Requirements

Requires 8-10 years of data engineering experience with proficiency in Python and SQL. Candidates must have hands-on experience with distributed computing frameworks and hold a bachelor's degree or equivalent.

Full description

Are you passionate about building scalable, high-impact data solutions that drive critical business insights? We are seeking an experienced and forward-thinking Lead Data Engineer to join our dynamic team.

In this pivotal role, you will be at the forefront of designing and constructing our next-generation data architecture. You will tackle complex challenges in data ingestion, transformation, and quality, directly shaping our ability to make data-driven decisions. If you thrive on building robust systems from the ground up and want to see your work make a tangible difference, this is the opportunity for you.

Key Responsibilities

  • Design, develop, and maintain robust ETL/ELT pipelines to ingest, transform, and deliver data from diverse sources.
  • Contribute to the design of scalable  data architectures, including data lakes, data warehouses, and lakehouse patterns.
  • Implement data quality checks, validation frameworks, and monitoring solutions to ensure data accuracy, completeness, and lineage compliance.
  • Build and optimize data solutions on cloud platforms  leveraging  services such as Spark, Kafka, Databricks.
  • Profile and tune SQL queries, Spark jobs, and pipeline workflows for efficiency and cost effectiveness.

Experience and Education:

  • Must to have 8-10 years of experience in Data Engineering 
  • proficiency in Python, strong SQL skills
  • Hands on experience with Apache Spark, Hadoop, or equivalent distributed computing frameworks
  • Experience with relational databases and NoSQL systems

Knowledge on Machine Learning, AI would be added advantage

  • Bachelor’s degree/University degree or equivalent experience

This job description provides a high-level review of the types of work performed. Other job-related duties may be assigned as required.

------------------------------------------------------

Job Family Group:

Technology------------------------------------------------------

Job Family:

Applications Development------------------------------------------------------

Time Type:

Full time------------------------------------------------------

Most Relevant Skills

Please see the requirements listed above.------------------------------------------------------

Other Relevant Skills

For complementary skills, please see above and/or contact the recruiter.------------------------------------------------------

Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.

 

If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi.

View Citi’s EEO Policy Statement and the Know Your Rights poster.

Similar roles