Middle Data Engineer
Suntech Innovation Dubai, Dubai Emirate, United Arab Emirates
Technology, Information and Internet · 51-200 employees
About the role
You will build and operate robust batch and streaming data pipelines to power experimentation and production systems. You will also collaborate with data scientists and backend engineers to ensure data contract reliability and pipeline performance.
What they look for
Requirements
Candidates must have 3+ years of experience in data or software engineering with strong proficiency in Python and PySpark. Experience with Kafka, Databricks, and orchestration tools like Airflow is essential for this role.
Benefits
Full description
About the team We are a focused engineering team of 5 to 9 people at a multi-product tech SaaS company. We work across personalization, recommendation, fraud detection, and some GenAI, through real-time, near-real-time, and batch pipelines. The team is a blend of Python backend engineers, ML engineers, QA (manual and automation), a data engineer, and data scientists. Our shared mandate is to take experimentation from our data scientists and turn it into robust, production-grade systems, either as standalone services or integrated into wider platforms.
The role You will build the data pipelines that power our experimentation and production systems, the backbone behind fraud detection, recommendations, and personalization. You will work primarily with Kafka (sometimes Flink), Spark on Databricks, and Airflow, modelling and optimizing data for both experimentation and production serving, and partnering closely with data scientists and backend engineers on data contracts and reliability.
Responsibilities
- Build and operate batch and streaming data pipelines (structured & unstructured data)
- Work primarily with Kafka, Spark on Databricks, and Airflow
- Model and optimize data for both experimentation and production serving
- Partner with data scientists and backend engineers on data contracts and reliability
- Own pipeline quality: testing, monitoring, and performance
- Take part in the teams on-call rotation for the pipelines you build and operate
Requirements
- 3+ years of data engineering or software engineering for data
- Strong Python, including PySpark
- Kafka and stream processing
- Databricks experience
- Airflow or similar orchestration tools
- Strong SQL and data modeling, plus solid relational database skills
- Software engineering mindset
Nice to have
- Flink
- ML Feature stores
- MLOps and ML pipeline exposure
- Docker and Kubernetes
- SaaS or high-scale product domain experience
Why join us
- Work on high-scale, high-throughput systems
- Broad range of technologies, stacks, and problems
- Contribute to turning experimentation into production systems
- A strong engineering culture that values technical depth and ownership
- Flexibility through hybrid and remote working
- A distributed team across Dubai and Europe
Similar roles
-
Data Engineer - Splunk and Azure
Bosch Group Bengaluru, Karnataka, India
-
Data Engineer
BID Operations Shenzhen, Guangdong Province, China
-
Data Engineer
Prodigal Bengaluru, Karnataka, India
-
Lead Data Engineer
Bristlecone Noida, Uttar Pradesh, India
-
Quantitative Data Engineer
Qube Research & Technologies Hong Kong, Hong Kong Island, Hong Kong S.A.R.
-
Data Engineer
CipherHealth United States