Meta

Data Engineer, Meta Superintelligence Labs (Privacy)

Meta Menlo Park, California, United States · $177K–$247K/yr

Software Development · 10,001+ employees

Yesterday
data-engineer Senior (5-10 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will own the privacy architecture for large-scale data pipelines and design structural controls for data usage, consent, and anonymization. Additionally, you will collaborate with cross-functional teams to build scalable data models and visualizations that support AI model development and product strategy.

What they look for

Data Engineering SQL Python Data Modeling ETL Privacy Engineering Data Governance Data Architecture Machine Learning Compliance Data Security System Design Algorithmic Optimization Data Visualization Technical Leadership Mentorship

Requirements

Candidates must have at least 7 years of experience in data engineering or related fields, including proficiency in SQL, ETL, and at least one programming language. A bachelor's degree in a technical field and proven experience implementing privacy or data-governance controls in large-scale production environments are required.

Benefits

Bonus Equity Health Insurance

Full description

Meta’s Products & Applied Research (PAR) team is where product-focused research meets real-world impact, taking breakthrough AI research and transforming it into products that reach billions. As part of Meta Superintelligence Labs (MSL), we’re driving the transformation of Meta’s core experiences—across Facebook, Instagram, WhatsApp, Threads, and beyond—by applying cutting-edge research to real-world products at massive scale.

We are looking for a Data Engineer to join our PAR organization where your technical skills and analytical mindset will be utilized designing and building some of the world's most extensive data sets, helping to craft experiences for billions of people and hundreds of millions of businesses worldwide.

In this role, you will collaborate with software engineering, data science, privacy, research, and product management teams to design/build scalable data solutions across Meta to optimize growth, strategy, and user experience.

You will own the privacy architecture for large-scale data pipelines that support AI model development. You will design and enforce the controls that determine whether data is permissible to use — consent and jurisdiction logic, personal-information filtering and anonymization, deletion and retention semantics, and access controls for sensitive fields. You will build these guarantees to be structural rather than manual: enforced automatically at write time, validated in continuous integration, and designed so that processing halts rather than proceeds when a control cannot be applied.

You will set the privacy engineering standards that other data engineers build against, define how new data surfaces onboard to those standards, and mentor engineers into independent ownership of parts of the architecture.

You will be at the forefront of identifying and solving some of the most interesting data challenges at a scale few companies can match. By joining Meta, you will become part of a world-class data engineering community dedicated to skill development and career growth in data engineering and beyond.

Data Engineering: You will guide teams by building optimal data artifacts (including datasets and visualizations) to address key questions. You will refine our systems, design logging solutions, and create scalable data models. Ensuring data security and quality, and with a strong focus on efficiency, you will suggest architecture and development approaches and data management standards to address complex analytical problems.

Product leadership: You will use data to shape product development, identify new opportunities, and tackle upcoming challenges. You'll ensure our products add value for users and businesses, by prioritizing projects, and driving innovative solutions to respond to challenges or opportunities.

Communication and influence: You won't simply present data, but tell data-driven stories. You will convince and influence your partners using clear insights and recommendations. You will build credibility through structure and clarity, and be a trusted strategic partner.

Responsibilities

  • Conceptualize and own the data architecture for multiple large-scale projects, while evaluating design and operational cost-benefit tradeoffs within systems
  • Create and contribute to frameworks that improve the efficacy of logging data, while working with data infrastructure to triage issues and resolve
  • Collaborate with engineers, product managers, and data scientists to understand data needs, representing key data insights visually in a meaningful way
  • Define and manage Service Level Agreements for all data sets in allocated areas of ownership
  • Determine and implement the security model based on privacy requirements, confirm safeguards are followed, address data quality issues, and evolve governance processes within allocated areas of ownership
  • Design, build, and launch collections of sophisticated data models and visualizations that support multiple use cases across different products or domains
  • Solve our most challenging data integration problems, utilizing optimal Extract, Transform, Load (ETL) patterns, frameworks, query techniques, sourcing from structured and unstructured data sources
  • Assist in owning existing processes running in production, optimizing complex code through advanced algorithmic concepts
  • Optimize pipelines, dashboards, frameworks, and systems to facilitate easier development of data artifacts
  • Influence product and cross-functional teams to identify data opportunities to drive impact
  • Mentor team members by giving/receiving actionable feedback

Minimum Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
  • 7+ years of experience where the primary responsibility involves working with data. This could include roles such as data analyst, data scientist, data engineer, or similar positions
  • 7+ years of experience with SQL, ETL, data modeling, and at least one programming language (e.g., Python, C++, C#, Scala or others.)
  • Experience designing and implementing privacy or data-governance controls in large-scale production data pipelines, including consent or eligibility logic, personal-information filtering or anonymization, and data deletion and retention requirements
  • Experience implementing access controls and data-governance mechanisms for sensitive data, including automated or programmatic enforcement rather than manual remediation
  • Experience translating privacy, legal, or regulatory requirements into verifiable technical controls, working with policy, legal, or compliance partners

Preferred Qualifications

  • Master's or Ph.D degree in a STEM field
  • Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
  • Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
  • Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
  • Experience building privacy or compliance controls for data used in machine learning or model training
  • Experience establishing data-governance standards and onboarding processes that multiple teams or product surfaces adopt
  • Experience implementing data deletion propagation or retroactive data-remediation processes across historical data
  • Experience serving as a technical lead, including setting technical direction and mentoring other engineers

$177,000/year to $247,000/year + bonus + equity + benefits

Similar roles