The Hong Kong Jockey Club

Data Engineer

The Hong Kong Jockey Club Sha Tin District, Hong Kong, China

Non-profit Organizations · 10,001+ employees

Sep 02
data-engineer Senior (5-10 yrs) Full-time China
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Design, develop, and operationalize reliable and scalable ETL/ELT data pipelines using PySpark and SQL on AWS. Implement and manage cloud data lakes while optimizing storage solutions for performance and reliability.

What they look for

Pyspark Sql Aws Etl/elt Data pipelines Cloud data lake Medallion architecture Data modelling Terraform Cloudformation Ci/cd Kafka Solace Data governance Data security Python

Requirements

Requires 5-8 years of hands-on experience in data engineering with a focus on scalable cloud ETL/ELT pipelines and PySpark. Candidates should have strong expertise in AWS cloud platforms, data governance, and DevOps practices.

Benefits

Competitive salary Benefits packages Dynamic working environment Development opportunities

Full description

The Hong Kong Jockey Club

Founded in 1884, The Hong Kong Jockey Club (“the Club”) is a world-class racing club that acts continuously for the betterment of our society. The Club has a unique integrated business model, comprising racing and racecourse entertainment, a membership club, responsible sports wagering and lottery, and charities and community contribution. Through this model, the Club generates economic and social value for the community and supports the HKSAR Government in combatting illegal gambling.

The Department

Join us as we shape the future of wagering at the Club through a strategic mission to transform our digital capabilities. Our flagship programme, operating under the internal acronym ‘AOP’, aims to Accelerate our Platform, Product, People, Performance and Potential, bringing this strategy to life through innovation and delivery excellence.

AOP's work spans product, technology, operations, delivery and change, creating next-generation experiences and strengthening the capabilities that support the Club's future growth. We are a diverse team passionate about solving complex challenges, delivering meaningful impact, and creating value for customers, members, the Club and the wider community.

If you are looking for a rewarding career where purpose, innovation and collaboration come together, your next opportunity starts here.

The Job

  • Design, develop, test, and operationalise reliable and scalable ETL/ELT data pipelines using PySpark and SQL on the AWS platform
  • Implement and manage a cloud data lake and orchestrate data flows following the medallion architecture
  • Optimise data pipelines and storage solutions for performance, cost, and reliability while handling huge volumes of data
  • Communication with business users and product owners may be required
  • Proven expertise in PySpark for large-scale data processing (terabyte- to petabyte-scale)
  • Strong experience in building and managing production-level ETL workflows and frameworks
  • Hands-on experience with AWS or other cloud data analytics services
  • Hands-on experience with data visualisation tools (e.g., Power BI, Tableau).

About you

  • 5-8 years of hands-on experience in data engineering with a focus on designing, building, and maintaining scalable Cloud ETL/ELT pipelines.
  • 5-8 years of production experience using PySpark and SQL for large-scale data processing.
  • 3+ years of experience with end-to-end data pipeline lifecycle management, including monitoring, observability, and production support.
  • 3+ years of experience architecting, implementing, and managing cloud data platforms on AWS.
  • Experience with data governance and data security best practices and compliance requirements.
  • Experience with real-time data streaming technologies such as Kafka or Solace.
  • Knowledge of Medallion architecture principles, design patterns and data modelling.
  • Experience with infrastructure-as-code (Terraform, CloudFormation).
  • Experience with DevOps/DataOps practices (CI/CD, automated testing, pipeline monitoring).
  • Background working with on-premises data platforms and migration to the cloud.
  • AWS Certifications: Data Analytics Speciality, Solutions Architect Associate/Professional.
  • Data Visualisation Certifications: Power BI, Tableau, QlikView, etc.
  • Core Data Processing: PySpark, SQL, Apache Airflow/Glue Orchestration, Delta Lake/Iceberg
  • Cloud & Infrastructure: AWS (S3, Glue, EMR, Redshift, Athena, Lake Formation), Terraform/CloudFormation
  • Development & Testing: Python, Jupyter, GitHub, Data pipeline unit/integration testing frameworks
  • Monitoring & Observability: AWS CloudWatch, Datadog, custom pipeline monitoring and alerting

Apply Now!

We offer competitive salary and benefits packages, a dynamic working environment and development opportunities.

Add horsepower to your career today. If you do not meet all of the requirements but still believe you can make a difference, please apply.

Equal Opportunity and Inclusive Hiring

We are an equal opportunity employer and strive to create an inclusive workplace for all. Applicants from diverse backgrounds are welcomed to apply. If you have any special needs or require accommodations during the interview process, please e-mail us via careers@hkjc.org.hk. Personal data provided by job applicants will be used strictly in accordance with the Club's notice to employees and job applicants relating to the Personal Data (Privacy) Ordinance. A copy of which will be provided immediately upon request.

Similar roles