Data Engineer Consultant
Indicium AI New York, New York, United States · $120K–$175K/yr
Information Technology & Services · 501-1,000 employees
About the role
Build and deploy production-grade data pipelines using modern cloud-native Lakehouse architectures. Collaborate with global engineering squads to modernize legacy data systems and ensure high-performance data delivery for AI applications.
What they look for
Requirements
Requires strong hands-on experience with Python, SQL, Databricks, and dbt for data transformation and infrastructure automation. Candidates should possess a consulting mindset with the ability to communicate technical solutions clearly to enterprise clients.
Benefits
Full description
Data Engineer Consultant (US Delivery Team)
The Opportunity
As a Data Engineer Consultant on our US team, you will build and deploy production-grade data pipelines for major enterprise clients. You will spend your day-to-day writing code, modernizing brittle legacy setups into cloud-native Lakehouses, and making client data clean, fast, and ready for AI applications.
In this role, you will work directly with modern technologies like Databricks, dbt, Python, and Terraform, collaborating with technical clients and our global engineering squads to ship solutions in weeks, not months.
Key Responsibilities
• Build Lakehouse Modernization Pipelines: Convert legacy data setups (Informatica, PL/SQL, legacy data warehouses) into high-performance PySpark, Databricks SQL, and Delta Live Tables (DLT) using Medallion Architecture standards (Bronze, Silver, Gold).
• Data Modeling & Transformation: Write clean, modular, and tested dbt models and advanced SQL to transform raw client data into trusted, business-ready datasets.
• Automate Infrastructure (DataOps): Use Terraform to provision cloud resources (AWS/GCP) programmatically and enforce data governance using Unity Catalog.
• Integrate & Orchestrate: Build and monitor automated ELT/ETL workflows using tools like Apache Airflow and Databricks Workflows, ensuring data quality, lineage, and strict SLA compliance.
• Client Engagement & Collaboration: Embedded directly within client projects, participating in daily standups, code reviews, and working closely with our nearshore delivery squads to deliver on time.
• Troubleshoot & Optimize: Debug failing pipelines, tune complex SQL/Spark queries for performance, and reduce cloud computing costs for our clients.
Qualifications & Requirements
Technical Stack
• Hands-On Data Engineering: Strong background writing production code in Python and advanced SQL.
• Databricks Experience: Hands-on experience building on the Databricks Lakehouse platform (Delta Lake, PySpark, Unity Catalog, Workflows).
• dbt Mastery: Solid experience using dbt for data transformation, testing, and documentation.
• Infrastructure as Code (IaC): Practical experience writing Terraform scripts to deploy AWS or GCP cloud data infrastructure.
• Orchestration & Version Control: Proficient with Git (GitHub/GitLab) for CI/CD workflows and experience with schedulers like Airflow.
Consulting & Delivery Mindset
• Client-Facing Comfort: Ability to communicate technical decisions clearly to client teams and collaborate in a fast-paced consulting environment.
• Problem-Solving Grit: Ability to jump into unfamiliar legacy codebases, figure out how the data flows, and rebuild it cleanly without needing a rigid playbook.
• Quality Focus: Passion for writing testable, documented, and reusable code.
Nice-to-Haves
• Databricks, AWS, or dbt certifications.
• Experience working on migration projects from legacy warehouses (Teradata, Netezza, Snowflake, Redshift) to Databricks.
• Familiarity with streaming tools like Apache Kafka or CDC (Change Data Capture) pipelines.
The anticipated base salary range for this role is $120,000 - $175,000. In addition to base pay, this position may be eligible for an annual discretionary bonus. An individual's final salary offer will be determined based on a variety of factors, including geographic location, experience, specialized skills, and qualifications. This compensation range is subject to updates or modifications at the company’s discretion
Similar roles
-
Senior Data Engineer, Economy
Roblox San Mateo, California, United States · $243K–$295K/yr
-
Fabric Senior Data Engineer
EXL Pune, Maharashtra, India
-
Cloud Data Engineer Senior
Paradigma Digital - Nuestras ofertas de Empleo Pozuelo de Alarcón, Community of Madrid, Spain
-
Data Engineer Senior
Paradigma Digital - Nuestras ofertas de Empleo Pozuelo de Alarcón, Community of Madrid, Spain
-
Data Engineer
EXL Atlanta, Georgia, United States
-
Senior Data Engineer
Clarity Innovations Tampa, Florida, United States · $77K–$180K/yr