HEXAWARE

Big Data Engineer

HEXAWARE · Mexico, Chihuahua, Mexico

IT Services and IT Consulting · 10,001+ employees

9 h ago
Junior (0-2 yrs) Full-time Mexico
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The role involves designing, developing, and maintaining cloud-based data pipelines and ETL solutions using AWS services. You will collaborate with engineers and stakeholders to extract, transform, and validate data while ensuring effective monitoring and documentation.

What they look for

Python SQL PySpark AWS Glue AWS Lambda Step Functions S3 Redshift RDS Oracle GitLab Terraform Data Engineering ETL Cloud Computing Data Pipelines

Requirements

Candidates should have 0-4 years of software development experience with foundational knowledge in Python, SQL, and PySpark. A strong understanding of AWS data services and ETL concepts is required to support development and deployment activities.

Full description

" Jr Dev - AWS Data Engineer We are seeking a **Jr Dev - AWS Data Engineer** with **0-4 years of software development experience** to support the design, development, and maintenance of cloud-based data pipelines and ETL solutions on AWS. This role is ideal for candidates who are building foundational skills in Python, SQL, PySpark, and AWS data services. Key Responsibilities - Assist in building and maintaining ETL pipelines using Python and PySpark. - Support development of workflows using AWS Glue, Lambda, and Step Functions. - Work with cloud data storage platforms such as S3, Redshift, RDS and Oracle. - Write SQL queries for data extraction, transformation, validation, and reporting. - Help implement basic monitoring, logging, and error handling for pipelines. - Support ingestion and processing of data from APIs and JSON payloads. - Collaborate with engineers, analysts, and stakeholders to understand requirements. - Contribute to code management, documentation, and deployment support activities. Required Skills and Qualifications - 0-4 years of software development experience across the appropriate platform. - Good knowledge on Python and SQL. - Good understanding to AWS services such as S3, SNS/SQS, EMR, Glue, Lambda, Redshift, and Step Functions. - Good understanding of ETL concepts and data processing fundamentals. - Familiarity with GitLab/Terraform and SDLC from development to production. - Familiarity to PySpark, AWS managed services, data engineering best practices, and code optimization. - Good analytical, problem-solving, and communication skills. Preferred Skills - Exposure to PySpark, Athena, CloudWatch, SNS and SQS - Internship, project, or academic experience in cloud, analytics, or data engineering. - Good understanding of using AI tools like Github Copilot or similar for code productivity"