Senior Data Scientist (Single Cell)
relationrx London, England, United Kingdom
Biotechnology Research · 51-200 employees
About the role
You will process and analyze large-scale single-cell multi-omics data to build generative and predictive models of cellular behavior. Additionally, you will collaborate with wet lab scientists and ML modelers to integrate multimodal datasets and evaluate model performance.
What they look for
Requirements
Candidates must hold a PhD or equivalent industry experience in computational biology, bioinformatics, or a related quantitative field. Proficiency in Python, workflow management frameworks like Nextflow, and deep experience with single-cell multi-omics analysis are required.
Full description
About Relation
Relation is a sector defining TechBio company developing transformational medicines, with technology at our core. Our ambition is to understand human biology in unprecedented ways, discovering therapies to treat some of life’s most devastating diseases. We leverage single-cell multi-omics from patient tissue, functional assays, and machine learning to drive disease understanding, from cause to cure.
We are scaling rapidly and building a team of exceptional individuals to push the boundaries of drug discovery. You will work in highly interdisciplinary teams where biology, computation, and engineering come together to solve complex problems that have not been solved before. Our state-of-the-art wet and dry labs in the heart of London are designed to accelerate this integration and translate insight into impact.
We are committed to building diverse and inclusive teams. Relation is an equal opportunities employer and does not discriminate on the basis of gender, sexual orientation, marital or civil partnership status, gender reassignment, race, colour, nationality, ethnic or national origin, religion or belief, disability, or age.
By joining Relation, you will help define how medicines are discovered and deliver meaningful impact for patients.
The opportunity
Relation is offering an outstanding opportunity for a Senior Data Scientist to help build the next generation of generative and predictive models of cellular behaviour, with a focus on single cell multi-omics. Your work will be central to our mission to understand and control cellular decision-making, enabling novel therapeutic strategies grounded in generative models.
You'll be joining a team with access to cutting-edge multiomic and interventional datasets, advanced computational infrastructure, and deep interdisciplinary expertise, and a culture that embraces modern ML tooling, including agentic workflows, to accelerate research iteration. You will contribute advanced data analysis and domain expertise to challenge ML models that aim to predict and explain cellular decision-making in disease.
Day to day, you will
- Process and QC paired single-cell RNA-seq and scATAC-seq at scale from thousands of perturbations across diverse cell types and biological contexts.
- Develop and maintain Nextflow pipelines compatible with custom combinatorial indexing chemistries and including methods such as STARsolo, Alevin-fry, Chromap, and SnapATAC2.
- Integrate scRNA-seq and scATAC-seq with extracellular proteomics and high-content imaging to build multimodal training datasets.
- Manage data in memory-efficient formats (e.g. Zarr) for GPU-accelerated processing and model training at scale.
- Partner closely with wet lab scientists from screen and cohort design through to interpretation, so experiments are analytically tractable and pipelines maximise data value.
- Contribute to the evaluation framework that determines whether ML models capture meaningful biology: design metrics, held-out benchmarks, and biological sanity checks.
- Collaborate closely with ML modellers to develop model architectures.
- Present findings and methodologies to internal stakeholders and contribute to publications.
Professionally, you will have
- A PhD (or equivalent industry experience) in computational biology, bioinformatics, or a related quantitative field.
- Deep hands-on experience processing and analysing large-scale single-cell multiomics data, ideally including paired RNA + ATAC (10x Multiome or equivalent).
- Proficiency with single-cell analysis toolkits (e.g. scverse ecosystem) and workflow management frameworks (Nextflow, Snakemake).
- Experience with arrayed CRISPR perturbation screen data, including well-level perturbation assignment via sample hashing, perturbation-specific QC (e.g. MOI estimation), and single-cell effect quantification.
- Proficiency in Python and familiarity with high-performance computing environments.
- Strong communication skills to bridge experimental and ML teams across a fast-moving programme
Bonus Experience:
- Experience integrating multiple data modalities (transcriptomics, chromatin accessibility, imaging and proteomics) for joint analysis.
- Exposure to foundation model development or perturbation prediction models.
- A background in statistical modelling and algorithm development.
- Experience working in interdisciplinary teams.
Personally, you
- Are comfortable working in a matrixed environment, balancing multiple stakeholders and contributing effectively across teams.
- Take ownership of your work, proactively seek opportunities to contribute, and enable others to do their best work.
- Communicate openly and directly, give and receive feedback constructively, and handle challenging conversations with respect.
- Actively seek out diverse perspectives, build strong working relationships, and contribute to shared goals across teams.
- Embrace challenges with openness and resilience, set high standards for yourself, and strive to deliver meaningful outcomes.
Working Style & Culture at Relation
At Relation, we operate in a matrixed, interdisciplinary environment, where impact is driven through collaboration across scientific, technical, and operational domains. We collaborate, and you will partner with colleagues across multiple teams and projects, contributing your expertise while aligning to shared company priorities. We work together and win together! The patient is waiting!
Recruitment Agencies
Please note that Relation does not accept unsolicited resumes from agencies. Resumes should not be forwarded to our job aliases or employees. Relation will not be liable for any fees associated with unsolicited CVs.
Similar roles
-
Business Data Scientist, gTech Users and Products
Google Dublin, Leinster, Ireland · $138K–$197K/yr
-
Senior Product Data Scientist, Consumer Payments
Google Singapore
-
Environmental Data Scientist
Uni Systems Ispra, Lombardy, Italy
-
Senior Data Scientist, Member Lifecycle
Vinted Vilnius, Vilnius County, Lithuania · €55K–€75K/yr
-
Senior Data Scientist / Analytics Manager - Pharmaceutical domain
WNS Global Services Gurgaon, Haryana, India
-
Data Scientist
Ford Motor Company Chennai, Tamil Nadu, India