Alignity Solutions

PySpark Developer / Senior Data Engineer

Alignity Solutions Hyderabad, Telangana, India · ₹800K–₹2M/yr

IT Services and IT Consulting · 11-50 employees

13 h ago
data-engineer Senior (5-10 yrs) Contractor India
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Design, build, and maintain scalable data pipelines using PySpark and Apache Spark. Develop efficient ETL/ELT workflows while collaborating with stakeholders to ensure data quality and performance.

What they look for

PySpark Apache Spark Python SQL ETL ELT Big Data Distributed Computing Data Pipelines Git Linux Unix Data Engineering Performance Tuning Data Validation

Requirements

Requires over 6 years of experience with strong hands-on expertise in big data processing and distributed computing. Candidates must possess deep knowledge of Spark concepts, Python programming, and SQL.

Full description

Do you love a career where you Experience, Grow & Contribute at the same time, while earning at least 10% above the market? If so, we are excited to have bumped onto you.

Learn how we are redefining the meaning of work , and be a part of the team raved by Clients, Job-seekers and Employees.

  • Jobseeker Video Testimonials
  • Employee Glassdoor Reviews

If you are a PySpark Developer / Senior Data Engineer looking for excitement, challenge and stability in your work, then you would be glad to come across this page.

We are an IT Solutions Integrator/Consulting Firm helping our clients hire the right professional for an exciting long-term project. Here are a few details.

Check if you are up for maximizing your earning/growth potential, leveraging our Disruptive Talent Solution.

Role: PySpark Developer / Senior Data Engineer Location: HYDERABAD | BANGALORE | PUNE | CHENNAI Experience: 6+ Years Employment Type: Contract to hire Notice Period:0-30 days(If you have negotiable notice period or buyout option please apply)

Requirements

We are looking for an experienced PySpark Developer with strong hands-on expertise in big data processing, distributed computing, and data engineering. The ideal candidate will have deep experience building scalable data pipelines, transforming large datasets, and working with Spark-based ecosystems in production environments.

Key Responsibilities

  • Design, build, and maintain scalable data pipelines using PySpark and Apache Spark
  • Develop efficient ETL/ELT workflows for batch and near-real-time processing
  • Optimize Spark jobs for performance, reliability, and cost efficiency
  • Work with large structured and unstructured datasets
  • Integrate data from multiple sources such as databases, APIs, files, and cloud storage
  • Write reusable, modular, and maintainable PySpark code
  • Troubleshoot job failures, data quality issues, and performance bottlenecks
  • Collaborate with data architects, analysts, platform teams, and business stakeholders
  • Implement data validation, monitoring, and logging frameworks
  • Support deployment, scheduling, and orchestration of data pipelines
  • Participate in design reviews, code reviews, and technical discussions
  • Mentor junior engineers and contribute to team best practices

Required Skills

  • Strong hands-on experience with PySpark and Apache Spark
  • Deep understanding of Spark concepts such as RDDs, DataFrames, datasets, partitioning, caching, shuffling, joins, and window functions
  • Strong Python programming skills
  • Experience with SQL and relational databases
  • Knowledge of big data concepts and distributed data processing
  • Hands-on experience with ETL/ELT pipeline development
  • Good understanding of performance tuning and optimization techniques in Spark
  • Experience with version control tools like Git
  • Familiarity with Linux/Unix environments
  • Strong debugging and analytical skills

Benefits

Visit us at http://alignity.io/careers. Alignity Solutions is an Equal Opportunity Employer, M/F/V/D. CEO Message: Click Here Clients Testimonial: Click Here

Similar roles