Blis

Senior Data Engineer

Blis London, England, United Kingdom

Advertising Services · 201-500 employees

13 h ago
data-engineer Senior (5-10 yrs) Full-time United Kingdom
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Design, build, monitor, and support large-scale data processing pipelines while ensuring data quality and efficient cloud compute usage. Collaborate with product teams to explore new data streams and mentor team members to enhance overall technical capabilities.

What they look for

Python GCP Apache Airflow Apache Druid Imply Spark Data Engineering Data Pipelines DevOps Linux Pandas Scikit-learn BigQuery Data Governance Distributed Processing Statistical Analysis

Requirements

Requires 5+ years of experience in building robust data pipelines and mastery of Python and GCP technologies. Proven expertise in Apache Druid, Imply, and large-scale system re-architecture is essential for this role.

Full description

Senior Data Engineer

Hybrid working, London

Come work on fantastically high-scale systems with us! Blis is an award-winning, global leader and technology innovator in big data analytics and advertising. We help brands such as McDonald's, Samsung, and Mercedes Benz to understand and effectively reach their best audiences.

We are looking for solid and experienced Data Engineers to work on building out secure, automated, scalable pipelines on GCP. We receive over 350gb of data an hour and respond to 400,000 decision requests each second, with petabytes of analytical data to work with.

We tackle challenges across almost every major discipline of data science, including classification, clustering, optimisation, and data mining. You will be responsible for building stable production level pipelines maximising the efficiency of cloud compute to ensure that data is properly enabled for operational and scientific cause.

This is a growing team with big responsibilities and exciting challenges ahead of it, as we look to reach the next 10x level of scale and intelligence.

At Blis, Data Engineers are a combination of software engineers, cloud engineers, and data processing engineers. They actively design and build production pipeline code, typically in Python, whilst having practical experience in ensuring, policing, and measuring for good data governance, quality, and efficient consumption. To run an efficient landscape we are ideally looking for candidates that are comfortable with event- driven automation across also aspects of our operational pipelines.

As a Blis data engineer, we seek to understand the data and problem definition and find efficient solutions, so critical thinking is a key component to efficient pipelines and effective reuse, this must include defining the pipelines for the correct controls and recovery points not only function and scale. The team are almost always adherents of Lean Development and work well in environments with significant amounts of freedom and ambitious goals.

Key responsibilities

  • Design, build, monitor, and support large scale data processing pipelines.
  • Support, mentor, and pair with other members of the team to advance our team’s capabilities and capacity.
  • Help Blis explore and exploit new data streams to innovative and support commercial and technical growth
  • Work closely with Product and be comfortable with taking, making and delivering against fast paced decisions to delight our customers.

This ideal candidate will be comfortable with fast feature delivery with a robust engineered follow up.

Skills and requirements

  • 5+ years direct experience delivering robust performant data pipelines within the constraints of direct SLA’s and commercial financial footprints.
  • Proven experience in architecting, developing, and maintaining Apache Druid and Imply platforms, with a focus on DevOps practices and large-scale system re-architecture
  • Mastery of building Pipelines in GCP maximising the use of native and native supporting technologies e.g. Apache Airflow
  • Mastery of Python for data and computational tasks with fluency in data cleansing, validation and composition techniques.
  • Hands-on implementation and architectural familiarity with all forms of data sourcing i.e streaming data, relational and non-relational databases, and distributed processing technologies (e.g. Spark)
  • Fluency with all appropriate python libraries typical of data science e.g. pandas, scikit-learn, scipy, numpy, MLlib and/or other machine learning and statistical libraries
  • Advanced knowledge of cloud based services specifically GCP
  • Excellent working understanding of server-side Linux
  • Professional in managing and updating on tasks ensuring appropriate levels of documentation, testing and assurance around their solutions.

Desired

  • Experience optimising both code and config in Spark, Hive, or similar tools
  • Practical experience working with relational databases, including advanced operations such as partitioning and indexing
  • Knowledge and experience with tools like AWS Athena or Google BigQuery to solve data-centric problems
  • Understanding and ability to innovate, apply, and optimise complex algorithms and statistical techniques to large data structures

Experience with Python Notebooks, such as Jupyter, Zeppelin, or Google Datalab to analyse, prototype, and visualise data and algorithmic output

About us

Blis is the only omnichannel DSP that unites telco data, real-world movement patterns, and transactions to deliver a complete view of the consumer. Powered by T-Mobile and built for precision at scale, Blis' omnichannel platform helps marketers map the full purchase journey – from impression to transaction – and expand their audience reach, driving incremental results across every screen.

Founded in the UK in 2004, Blis employs over 300 global employees across 14 offices in 11 countries.

Similar roles