Y-Prime, LLC

Data Engineer I

Y-Prime, LLC Malvern, Pennsylvania, United States

Software Development · 201-500 employees

4 d ago
data-engineer Mid (2-5 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Build and maintain data pipelines to ingest data into Snowflake while developing dbt models for analytics-ready datasets. Collaborate with cross-functional teams to ensure data reliability, security, and performance across the organization.

What they look for

Snowflake Dbt SQL Git GitHub Python Data modeling Dimensional modeling ELT ETL Data pipelines CI/CD Data governance Cloud computing Troubleshooting Analytical skills

Requirements

Requires 2+ years of experience in data engineering with proficiency in Snowflake, dbt, and SQL. Candidates should have strong version control skills using GitHub and a solid understanding of data modeling and ELT/ETL methodologies.

Benefits

Flexible paid time off Comprehensive benefits package 401(k) with company match Professional growth and advancement

Full description

About YPrime

At YPrime, we pioneer solutions that streamline the clinical trial journey, increasing certainty from study design to data lock. With a foundation built on decades of industry insight and expertise, we are inspired by the life-altering outcomes unlocked by clinical trials. Our dedication to quality is pivotal in propelling the groundbreaking endeavors of our partners, researchers, and investigators. With a technology platform that enables speed, flexibility, and certainty for large and emerging pharma companies alike, we provide eConsent, eCOA, IRT, and patient engagement solutions that solve with certainty in clinical research.

About the Role

YPrime is seeking a Data Engineer I to join our Data Analytics and Engineering team. In this role, you will help build and maintain the data platform that powers analytics, reporting, and business insights across the organization. You'll work with modern data technologies, including Snowflake and dbt, to transform raw data into trusted, analytics-ready datasets while ensuring the reliability, security, and performance of our data ecosystem.

This position is ideal for someone who enjoys solving complex data challenges, collaborating with cross-functional teams, and contributing to the development of scalable data solutions that support business decision-making.

Responsibilities

  • Build, maintain, and monitor data pipelines that ingest data from internal and third-party source systems into Snowflake.
  • Develop and maintain dbt models, tests, and documentation to transform raw data into curated, analytics-ready datasets.
  • Write efficient, scalable SQL for data transformation, validation, and analysis.
  • Utilize GitHub for version control, including branching, pull requests, code reviews, and CI/CD workflows.
  • Implement data quality checks, monitor data reliability, and troubleshoot pipeline failures and data discrepancies.
  • Collaborate with analysts, business stakeholders, and technical teams to understand data requirements and deliver high-quality datasets.
  • Support Snowflake administration activities, including access management, warehouse optimization, and performance monitoring.
  • Document data models, pipelines, and operational processes while contributing to team standards and best practices.
  • Participate in Agile ceremonies, including sprint planning, backlog refinement, code reviews, and retrospectives.
  • Ensure compliance with company policies, data governance standards, security requirements, and applicable SOPs.

Skills

  • Working knowledge of Snowflake, including databases, schemas, virtual warehouses, roles, stages, streams, and tasks.
  • Experience developing and maintaining dbt models, tests, snapshots, macros, and documentation.
  • Strong SQL skills with the ability to write, troubleshoot, and optimize complex queries.
  • Proficiency with Git and GitHub workflows, including pull requests, code reviews, and CI/CD processes.
  • Working knowledge of Python for scripting, automation, and data processing.
  • Understanding of data modeling concepts, including dimensional modeling, star schemas, and slowly changing dimensions.
  • Familiarity with ELT/ETL methodologies and common orchestration or ingestion tools.
  • Basic understanding of cloud computing platforms and data security best practices.
  • Strong analytical, troubleshooting, and problem-solving skills.
  • Excellent attention to detail and ability to communicate effectively with both technical and non-technical audiences.
  • Ability to work collaboratively in a fast-paced, team-oriented environment.

Education & Experience

  • 2+ years of experience in data engineering, analytics engineering, or a related field.
  • Demonstrated experience building and maintaining production data pipelines using Snowflake and dbt.
  • Experience working within a collaborative, version-controlled development environment utilizing GitHub.
  • Experience writing and optimizing SQL for large-scale data transformation and analysis.
  • Familiarity with data modeling, data warehousing, and modern ELT/ETL best practices.
  • Experience supporting analytics, reporting, or business intelligence initiatives.
  • Experience working in a regulated or data-governed environment is a plus.
  • No formal degree required. Equivalent practical experience, relevant certifications (such as SnowPro Core or dbt Analytics Engineering), and demonstrated technical expertise will be considered.

What are the Perks?

  • Flexible paid time off
  • Comprehensive benefits package
  • 401(k) with company match
  • Friendly, smart, passionate, and hard-working coworkers
  • Opportunities for professional growth and advancement

Similar roles