The Hong Kong Jockey Club

Data Engineer

The Hong Kong Jockey Club · Sha Tin District, Hong Kong, China

Non-profit Organizations · 10,001+ employees

13 h ago
Senior (5-10 yrs) Full-time China
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Design, develop, and operationalize scalable ETL/ELT data pipelines using PySpark and SQL on AWS. Manage cloud data lakes and optimize data storage solutions for performance and reliability.

What they look for

Pyspark Sql Aws Etl/elt pipelines Cloud data lake Medallion architecture Data modelling Terraform Cloudformation Kafka Solace Power bi Tableau Devops Dataops Ci/cd

Requirements

Requires 5-8 years of experience in data engineering with a focus on cloud ETL/ELT pipelines and large-scale data processing. Candidates must have strong expertise in PySpark, SQL, and AWS cloud infrastructure.

Benefits

Competitive salary Benefits packages Development opportunities

Full description

The Hong Kong Jockey Club

Founded in 1884, The Hong Kong Jockey Club (“the Club”) is a world-class racing club that acts continuously for the betterment of our society. The Club has a unique integrated business model, comprising racing and racecourse entertainment, a membership club, responsible sports wagering and lottery, and charities and community contribution. Through this model, the Club generates economic and social value for the community and supports the HKSAR Government in combatting illegal gambling.

The Department

Since 1884, The Hong Kong Jockey Club has been a cornerstone of Hong Kong’s sports and entertainment landscape, driving innovation while contributing to the betterment of society. We are seeking motivated individuals eager to help shape the future of sports and entertainment in a fast-paced, dynamic environment. If you value creativity, collaboration, innovation, and love a challenge, this opportunity is for you.

The Job

  • Design, develop, test, and operationalise reliable and scalable ETL/ELT data pipelines using PySpark and SQL on the AWS platform
  • Implement and manage a cloud data lake and orchestrate data flows following the medallion architecture
  • Optimise data pipelines and storage solutions for performance, cost, and reliability while handling huge volumes of data
  • Communication with business users and product owners may be required
  • Proven expertise in PySpark for large-scale data processing (terabyte- to petabyte-scale)
  • Strong experience in building and managing production-level ETL workflows and frameworks
  • Hands-on experience with AWS or other cloud data analytics services
  • Hands-on experience with data visualisation tools (e.g., Power BI, Tableau).

About you

  • 5-8 years of hands-on experience in data engineering with a focus on designing, building, and maintaining scalable Cloud ETL/ELT pipelines.
  • 5-8 years of production experience using PySpark and SQL for large-scale data processing.
  • 3+ years of experience with end-to-end data pipeline lifecycle management, including monitoring, observability, and production support.
  • 3+ years of experience architecting, implementing, and managing cloud data platforms on AWS.
  • Experience with data governance and data security best practices and compliance requirements.
  • Experience with real-time data streaming technologies such as Kafka or Solace.
  • Knowledge of Medallion architecture principles, design patterns and data modelling.
  • Experience with infrastructure-as-code (Terraform, CloudFormation).
  • Experience with DevOps/DataOps practices (CI/CD, automated testing, pipeline monitoring).
  • Background working with on-premises data platforms and migration to the cloud.
  • AWS Certifications: Data Analytics Speciality, Solutions Architect Associate/Professional.
  • Data Visualisation Certifications: Power BI, Tableau, QlikView, etc.
  • Core Data Processing: PySpark, SQL, Apache Airflow/Glue Orchestration, Delta Lake/Iceberg
  • Cloud & Infrastructure: AWS (S3, Glue, EMR, Redshift, Athena, Lake Formation), Terraform/CloudFormation
  • Development & Testing: Python, Jupyter, GitHub, Data pipeline unit/integration testing frameworks
  • Monitoring & Observability: AWS CloudWatch, Datadog, custom pipeline monitoring and alerting

Apply Now!

We offer competitive salary and benefits packages, a dynamic working environment and development opportunities.

Add horsepower to your career today. If you do not meet all of the requirements but still believe you can make a difference, please apply.

Equal Opportunity and Inclusive Hiring

We are an equal opportunity employer and strive to create an inclusive workplace for all. Applicants from diverse backgrounds are welcomed to apply. If you have any special needs or require accommodations during the interview process, please e-mail us via careers@hkjc.org.hk. Personal data provided by job applicants will be used strictly in accordance with the Club's notice to employees and job applicants relating to the Personal Data (Privacy) Ordinance. A copy of which will be provided immediately upon request.