Air Products

Senior Data Engineer

Air Products pune, Maharashtra, India

Chemical Manufacturing · 10,001+ employees

13 h ago Closes in 3d
data-engineer Senior (5-10 yrs) Full-time India
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Senior Data Engineer will design, develop, and optimize cloud-native data pipelines using Databricks and Apache Spark. They will also lead the migration of enterprise data workloads from AWS to Databricks while implementing robust data quality and governance practices.

What they look for

Databricks Apache Spark PySpark Delta Lake Azure AWS Structured Streaming SQL Python Terraform Data Engineering Data Pipelines Cloud-native Architectures Data Governance Data Security CI/CD

Requirements

Candidates must have 8+ years of experience in data engineering with strong hands-on expertise in Databricks and PySpark. Proficiency in Delta Lake, cloud-native architectures, and data migration strategies is essential for this role.

Full description

At Air Products, we reimagine what’s possible. By tapping into the motivation of our people and our collective experience, we create the ideas and innovations that drive us forward. When we come together – where every voice is heard and everyone knows they belong and matter – we create solutions that launch people into space, support lifesaving care in hospitals, and enable the construction of groundbreaking, world scale production facilities.

Reimagine What’s Possible 

Location: Pune (Hybrid)

Experience: 8+ Years

Skills: Databricks | Spark | Azure or AWS | Delta Lake

 

About the Role

We are looking for a Senior Data Engineer with strong hands-on experience in Databricks to build and scale modern, cloud-native data platforms. You will work on high-impact data engineering initiatives, collaborating with analytics, AI/ML, and business teams to deliver reliable and performant data solutions.

This role will also be involved in migrating enterprise data workloads from AWS-based platforms to Databricks, contributing to our broader data modernization journey.

 

Key Responsibilities

  • Must have experience in requirement gathering, design, develop, and optimize data pipelines using Databricks and Apache Spark (PySpark).
  • Build and manage Delta Lake Medallion architectures.
  • Develop batch and streaming pipelines using Structured Streaming.
  • Optimize Spark workloads for performance and cost efficiency.
  • Implement Databricks Jobs, Workflows, and SQL.
  • Ingest data from various relation databases, SAP, API’s, etc into databricks.
  • Contribute to AWS (Glue/Redshift) → Databricks data migration initiatives.
  • Implement data quality, security, and governance best practices.
  • Collaborate with cross-functional teams and mentor junior engineers.

 

Required Skills & Experience

  • 8+ years of experience in Data Engineering
  • Strong hands-on experience with Databricks.
  • Knowledge and experience on security and governance topics is a plus.
  • Advanced knowledge of Apache Spark (PySpark preferred)
  • Experience with Terraform (IAC), Delta Lake, SQL, and Python
  • Solid understanding of cloud-native data architectures on Databricks.
  • Experience building scalable, production-grade data pipelines
  • Strong problem-solving and communication skills.

 

Nice to Have

  • Experience with Unity Catalog and data governance
  • Familiarity with CI/CD pipelines and DevOps best practices for data solutions.
  • Familiarity with Terraform, dbt, or Airflow
  • Databricks certification (preferred, not mandatory)

 

Founded in 1940, Air Products is a world-leading industrial gases company and has a proud history of innovation, operational excellence, with an unwavering commitment to safety and environmental stewardship. Working together, we are taking our passion and diverse backgrounds forward to reimagine what’s possible and generate a cleaner future for our customers, our communities, and the world.

Similar roles