Numentica LLC

Senior Data Engineer

Numentica LLC San Diego, California, United States

IT Services and IT Consulting · 201-500 employees

Yesterday
data-engineer Senior (5-10 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Lead the architecture and engineering of the Microsoft Fabric data lake-house while providing technical leadership and mentoring to the engineering team. Partner with cross-functional business stakeholders to translate complex needs into scalable data solutions and establish robust platform operational processes.

What they look for

Microsoft Fabric Data Engineering Technical Leadership PySpark Python T-SQL ETL/ELT Medallion Architecture CI/CD Azure DevOps Data Governance Microsoft Purview Data Lake-house Power BI Infrastructure-as-code

Requirements

Requires a bachelor's degree in a relevant field and 7–10 years of experience in data engineering or analytics infrastructure. Candidates must possess strong hands-on proficiency in Microsoft Fabric, PySpark, Python, and modern data engineering practices.

Full description

We are seeking a hands-on Technical Lead to lead the development and operation of its Microsoft Fabric-based Platform. This role combines data platform architecture, hands-on engineering, technical leadership, and business partnership.

Key Responsibilities

  • Lead the architecture and engineering of the Microsoft Fabric data lake-house, including OneLake, medallion architecture, Spark/SQL, and CI/CD.
  • Provide hands-on technical leadership, mentoring, code reviews, and architectural guidance to data and analytics engineers.
  • Design and build scalable data pipelines integrating 50+ enterprise source systems.
  • Partner with business stakeholders across R&D, Commercial, Manufacturing, Supply Chain, and Corporate functions to translate business needs into data solutions.
  • Manage and technically oversee external implementation partners, ensuring engineering quality and adherence to standards.
  • Establish platform support models, SLAs, monitoring, data quality, incident management, and operational processes.
  • Define and track platform KPIs such as pipeline reliability, data freshness, query performance, onboarding velocity, and incident resolution.
  • Develop the semantic layer, including Power BI datasets, gold-layer views, and self-service analytics.
  • Implement data governance, security, classification, RBAC, sensitivity labeling, and audit requirements.
  • Establish engineering standards for development, testing, branching, documentation, and deployment.

Required Qualifications

  • Bachelor’s degree in Computer Science, Data Engineering, Information Systems, or related field.
  • 7–10 years of experience in data engineering, data platforms, or analytics infrastructure.
  • 2+ years of technical leadership experience with data/analytics engineers.
  • Strong hands-on experience with Microsoft Fabric; Databricks or Azure Synapse experience may be considered.
  • Strong proficiency in PySpark, Python, T-SQL, ETL/ELT, and modern data engineering practices.
  • Experience with medallion/layered data architecture at enterprise scale.
  • Experience with DevOps, CI/CD, Azure DevOps/GitHub, and infrastructure-as-code.
  • Knowledge of data governance and quality frameworks such as Microsoft Purview.
  • Experience managing technical delivery from external vendors.

Preferred

  • Biopharma, life sciences, or other regulated-industry experience, including GxP and validation requirements.
  • Experience with Veeva, LIMS, ELN/Benchling, CTMS, EDC, MasterControl, or HRIS integrations.
  • Experience with AI/ML workloads on a lake-house.
  • Microsoft Fabric Data Engineering certification.

Similar roles