Site Reliability Engineer
Genesis London, England, United Kingdom
Software Development · 11-50 employees
About the role
You will design and maintain large-scale, fault-tolerant distributed systems to ensure the reliability and performance of the Genesis platform. Additionally, you will drive infrastructure improvements and automation to eliminate toil while fostering a culture of reliability across engineering teams.
What they look for
Requirements
Candidates must possess strong software engineering skills in languages like Python or Go and have extensive experience with distributed systems and cloud platforms. You should also demonstrate leadership in complex technical projects and the ability to solve ambiguous problems at scale.
Full description
What You'll Do
- Take on ambiguous reliability, scalability, and efficiency challenges and drive solutions across SRE and development teams.
- Build and run large-scale, massively distributed, fault-tolerant systems that keep Genesis platform reliable and performant for our customers.
- Optimize existing systems, build infrastructure, and eliminate toil through automation to continuously improve uptime and rate of change.
- Cultivate a culture of reliability throughout the organization, guiding technical decisions that balance system health with fast-moving product priorities.
- Ensure the long-term health, maintainability, and reliability of services through capacity planning, performance analysis, and proactive incident prevention.
What You'll Bring
- Strong software engineering skills (e.g., in Python, Go, or similar) with extensive experience designing, analyzing, and troubleshooting distributed systems.
- Deep expertise with cloud computing platforms (e.g., Kubernetes, Cloud Functions) and Non-Abstract Large Systems Design (NALSD).
- Experience leading complex, large-scale technical projects and providing technical leadership across teams.
- Ability to apply coding, algorithms, and complexity analysis to solve ambiguous problems at scale with minimal disruption.
- A collaborative, intellectually curious mindset — comfortable working across a wide variety of backgrounds and bringing cross-team perspective to build robust, reusable solutions.
Similar roles
-
Senior Site Reliability Engineer
GetFrankly Cluj-Napoca, Romania
-
SRE (Site Reliability Engineer) - H/F
Devoteam Levallois-Perret, Ile-de-France, France
-
Ingénieur Observabilité / SRE - H/F
Devoteam Levallois-Perret, Ile-de-France, France
-
SRE
Radware Tel-Aviv, Tel-Aviv District, Israel
-
SRE [Antifraud]
Plata Card Osnabrück, Lower Saxony, Germany
-
Senior SRE & Monitoring Developer
Ford Motor Company India