通用磨坊股份有限公司

CPW Data Engineer II

通用磨坊股份有限公司 · Mumbai, Maharashtra, India

Manufacturing · 10,001+ employees

Yesterday
Senior (5-10 yrs) Full-time India
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Develop and maintain robust ETL/ELT pipelines using Azure Databricks and PySpark to build high-quality data assets. Partner with stakeholders to design scalable data models and ensure data quality, security, and consistency across business hierarchies.

What they look for

Azure Databricks PySpark SQL Azure Data Factory ETL ELT Data Modeling Data Governance Delta Lake Data Warehousing Business Intelligence Data Quality Azure Storage Azure Key Vault Azure DevOps Data Integration

Requirements

Requires a Bachelor's degree in Computer Science or a related field and a minimum of 7 years of relevant data engineering experience. Candidates must have strong hands-on experience with SQL, PySpark, Azure Databricks, and Azure Data Factory.

Full description

COMPANY OVERVIEW

We exist to make food the world loves. But we do more than that. Our company is a place that prioritizes being a force for good, a place to expand learning, explore new perspectives and reimagine new possibilities, every day. We look for people who want to bring their best — bold thinkers with big hearts who challenge one another and grow together. Because becoming the undisputed leader in food means surrounding ourselves with people who are hungry for what’s next.​


OVERVIEW

Cereal Partners Worldwide (CPW) is a joint venture between General Mills and Nestlé, two of the world’s leading food organizations. Working at CPW offers the best of both worlds: the scale and capabilities of large organizations combined with the agility and entrepreneurial spirit of a smaller company.

CPW builds cutting-edge capabilities, sets industry benchmarks, partners with technology start-ups, implements advanced processes, and creates strong technical and analytical foundations. This is an opportunity to learn, grow, and contribute to an exciting, data-driven team.

KEY ACCOUNTABILITIES

  • Build high-quality data assets that provide a competitive advantage.
  • Develop and maintain robust ETL/ELT pipelines using Azure Databricks and PySpark.
  • Design and maintain Bronze, Silver, and Gold data-layer architectures.
  • Implement incremental processing, deduplication, data validation, and performance optimization.
  • Establish and maintain data governance and stewardship practices, including data quality, security, validation metrics, and accurate data flow.
  • Integrate data from multiple sources, including Nielsen, Circana, internal systems, panel data, RMS data, and other relevant datasets.
  • Develop scalable data models to support reporting, analytics, and business intelligence use cases.
  • Ensure consistency across product, period, market, and other business hierarchies.
  • Review existing data models and processes to identify sustainable, scalable, and automated approaches.
  • Harmonize data and processes while maintaining an effective balance between speed, quality, and business needs.
  • Partner with stakeholders and internal teams to understand data requirements and deliver timely solutions to business questions.
  • Support reporting and visualization teams with backend data architecture and data model development.
  • Optimize queries, Delta storage formats, pipelines, workflows, scalability, and cost efficiency.
  • Identify and resolve data-quality issues and proactively address data opportunities.
  • Maintain strong working relationships with peers and contribute effectively as a team member.

MINIMUM QUALIFICATIONS

  • Bachelor’s degree in Computer Science, Information Technology, Electronics and Telecommunications, or a related field.
  • Minimum 7 years of relevant experience in Data Engineering.
  • Mandatory experience working with data lakes and multiple data sources.
  • Strong hands-on experience with - SQL, PySpark, Azure Databricks, Azure Data Factory.
  • Knowledge of data warehousing concepts, ETL frameworks, and business intelligence platforms.
  • Experience with Delta Lake architecture and large-scale data processing.
  • Strong understanding of data modeling, data integration, and data quality principles.
  • Good communication and stakeholder-management skills.

PREFERRED SKILLS

  • Master’s degree in Computer Science, Information Technology, Electronics and Telecommunications, or a related field.
  • Experience in the FMCG industry.
  • Experience developing business intelligence platforms or 360-degree data views.
  • Experience with Nielsen, Circana, panel data, RMS data, or similar consumer and market datasets.
  • Experience with Azure Storage Accounts, Azure Key Vault, and Azure DevOps.
  • Familiarity with data governance, data security, and stewardship frameworks.
  • Strong focus on data accuracy, reliability, and attention to detail.
  • Ability to manage ambiguous data issues and drive timely resolution.
  • Continuous-improvement and ownership mindset.
  • Ability to lead change, make quality decisions, and deliver results within established timelines.
  • Curiosity and willingness to learn new tools, technologies, and functional capabilities.
  • Ability to build strong peer relationships, collaborate effectively, and positively influence others.


ELIGIBILITY

Applicants must meet minimum age qualifications in the country in which the job is located.