Personify Health

Data Engineer II

Personify Health Tempe, Arizona, United States

Health and Human Services · 1,001-5,000 employees

Yesterday
data-engineer Mid (2-5 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Design and maintain robust ETL/ELT pipelines to transform raw healthcare and TPA data into reliable information for business operations. Collaborate with cross-functional teams to ensure data quality, compliance with HIPAA/CMS regulations, and effective resolution of pipeline issues.

What they look for

Python SQL Airflow ETL/ELT pipelines AWS Data modeling PostgreSQL Oracle Docker CI/CD EDI transactions Data warehousing Healthcare data API integration Tableau Power BI

Requirements

Requires a bachelor's degree in computer science or a related field and at least 3 years of experience in data engineering, preferably within the healthcare or insurance industry. Candidates must be proficient in Python, SQL, and data orchestration tools like Airflow.

Benefits

Medical insurance Dental insurance Unlimited PTO Mental health support Retirement planning Financial protection Professional development Learning budgets

Full description

Overview

Who We Are

Because health is personal. That's why Personify Health created the first and only personalized health platform—bringing health plan administration, holistic wellbeing solutions, and comprehensive care navigation together in one place. We serve employers, health plans, and health systems with data-driven solutions that reduce costs while actually improving health outcomes. Together, our team is on a mission to empower people to lead healthier lives.

Learn even more about the work that drives us at personifyhealth.com.

Responsibilities

Ready to build the data infrastructure that keeps healthcare claims moving?

Why This Role Matters

Every claim, eligibility check, and provider record that flows through our systems depends on pipelines that work — every time, without exception. As a Data Engineer II, you're the person who builds and maintains that backbone, turning raw healthcare and TPA data into clean, reliable information that analysts, business teams, and ultimately our clients and members can count on. When a pipeline breaks or data goes bad, real people feel it — a claim gets delayed, a report goes out wrong, a decision gets made on bad numbers. Your work closing that gap between messy source data and trustworthy systems is what lets the rest of the organization move fast with confidence. Get the pipelines right, and you're not just moving data — you're moving outcomes for the members and clients who rely on us.

What You'll Actually Do

  • Build ETL/ELT pipelines: Design and maintain pipelines that ingest and transform healthcare and TPA data — including claims, provider, and eligibility sources — turning raw feeds into usable, trustworthy data sets.
  • Orchestrate and monitor workflows: Develop, schedule, and monitor pipelines using Airflow, CloudWatch, ECS, and DAGs, following established CI/CD, observability, and governance practices to keep data flowing reliably.
  • Develop data applications: Build workflows and applications using Python, SQL, and Django, working across PostgreSQL, Oracle, and cloud-native databases to power downstream reporting and analysis.
  • Support core healthcare data processes: Manage EDI file transfers, claims adjudication support, audits, and reporting workflows that keep compliance and operations on track.
  • Own data extraction and documentation: Extract, cleanse, and load new data sets, investigating and documenting source systems so the team always knows what it's working with.
  • Translate requirements into solutions: Partner with Data Analysts, Data Scientists, Product, Reporting, and Account Management to turn business needs into working data pipelines.
  • Safeguard data quality and compliance: Implement quality assurance rules and automated validation that keep data accurate, complete, and secure under HIPAA and CMS regulations.
  • Troubleshoot and resolve pipeline issues: Monitor pipeline health, triage incoming bugs and incidents, and resolve data flow issues — escalating complex architectural problems to senior engineers when needed.
  • Contribute to data modeling decisions: Apply data modeling patterns and standards set by senior engineers, contributing your perspective to pipeline design discussions.
  • Guide junior engineers: Help onboard and mentor Data Engineer I team members, sharing what you learn so the whole team levels up.
  • Automate routine work: Build automation for repetitive operational workflows and reporting, cutting out manual busywork wherever possible.

Qualifications

What You Bring to Our Team

Education & Experience:

  • Bachelor's degree in computer science, information systems, or a related field, or equivalent experience
  • 3+ years in data engineering or analytics engineering, ideally within TPA, healthcare, insurance, or claims processing
  • 1+ years in system/data analysis or process improvement
  • 1+ years in the healthcare industry preferred
  • AWS Certification preferred: AWS Certified Cloud Practitioner, Developer – Associate, or Data Engineer – Associate

Technical Skills:

  • Proficient in Python and SQL, including complex queries (pivots, window functions, date calculations)
  • Hands-on experience with orchestration tools (Airflow), containers (Docker), and CI/CD pipelines
  • Exposure to healthcare EDI transactions (834, 835, 837, 2222, 2223, 999) preferred
  • Exposure to data files like OMS or HL7, or ability to parse delimited files into a database using Python scripting
  • Familiarity with REST APIs, JSON, and non-relational data models, including converting that data into relational models
  • Experience with JIRA, BitBucket Git, and BitBucket Pipelines
  • Proficient in Excel; familiar with analytical tools like Tableau, Power BI, or MicroStrategy
  • Solid understanding of data modeling concepts (star/snowflake schemas, dimensional modeling) and relational vs. non-relational models
  • Working knowledge of AWS services (S3, Glue, EC2, MWAA, Lambda, ECS) a plus; Infrastructure as Code (Terraform) a plus
  • Experience with relational databases (PostgreSQL, Oracle, AWS RDS); exposure to modern data warehouses (Snowflake, Redshift) a plus

Benefits

The Highlights:

  • Competitive base salary and benefits effective day one
  • Comprehensive medical and dental through our own health solutions (yes, we use what we build)
  • Unlimited PTO—rest and recharge time is non-negotiable
  • Mental health support, retirement planning, and financial protection
  • Professional development with clear career progression and learning budgets
  • Mission-driven culture where diverse perspectives drive real impact on people's health

Want the full picture? Visit personifyhealthbenefits.com to explore our complete benefits package, wellness programs, and other employee perks.

Compensation: This position offers a competitive base salary that varies based on location, skills, and experience.

You're eligible for our full benefits package starting day one.

Our Commitment: Personify Health is an equal opportunity employer committed to diversity, equity, inclusion, and belonging. We cultivate a work environment where differences are celebrated, and employees of all backgrounds are empowered to thrive—because diversity is core to who we are and critical to our work in health and wellbeing.

Stay Safe: Personify Health will never ask for payment or sensitive personal information like social security numbers during hiring. All official communication comes from verified company email addresses and or our secure applicant tracking system. Suspicious requests? Report them to talent@personifyhealth.com. View all legitimate openings at personifyhealth.com/careers.

Similar roles