Software Engineer (SRE)
Geordie AI London, England, United Kingdom
Software Development · 11-50 employees
About the role
You will build and improve product capabilities while owning the reliability and observability of the services you ship. You will also collaborate with stakeholders to design solutions, manage incident responses, and drive technical excellence through automation and infrastructure improvements.
What they look for
Requirements
You must be fluent in at least one of Java, Python, TypeScript, or Rust and have experience leading the delivery of production-grade software. Additionally, you should be comfortable in cloud environments like AWS and possess strong skills in observability, CI/CD, and infrastructure as code.
Full description
The Role
We are looking for a seasoned software engineer with a background in Site Reliability Engineering (SRE) to build the platform that enterprises use to understand and govern their AI agents.
Geordie is building category-defining tooling for agent visibility, governance, and risk. The ecosystem around us is moving fast, our customers are moving fast, and we need to work fast and sustainably to keep up. You will work directly on impactful customer problems to understand what they need, design solutions together, and get changes in front of them quickly.
We use a variety of technologies across the Java, Python, Node and Rust ecosystems and cloud-native infrastructure in AWS. We make heavy use of agents in our day-to-day work and dogfood our own product extensively.
Our engineering culture takes influence from Extreme Programming: collaborative by default, close to customers, iterating in small steps, and shipping at a swift but sustainable pace. We pair often, share ownership of what we build, and lean on automated tests and observability so we can keep moving without breaking what we have already shipped.
What You'll Do
- Build and improve product capabilities for customers, shipping consistently and impactfully.
- Own the reliability of what you ship: define what "healthy" means for your services, instrument them, and act on what the data tells you.
- Grow our observability, alerting, and incident practice so problems surface before customers notice them, and are diagnosed quickly when they do.
- Strengthen the infrastructure, delivery pipelines, and automation that let a small team ship safely and often. Infrastructure as code, repeatable deploys, fast and honest feedback.
- Respond to customer asks and unblock the Go to Market team.
- Actively improve team processes and ways of working.
- Discuss, design, and iterate solutions collaboratively with business stakeholders and customers.
- Drive technical excellence through simplicity: refactoring internal systems, reducing operational toil, and streamlining customer workflows.
- Take part in the on-call rota for incident response and turn what we learn into changes that stick to avoid them happening again.
- Learn and explore the future of agents and agent risks alongside the rest of the team.
- Drive the technical direction of the product, the platform, and the way we work.
Behaviours We Are Looking For
- The most important thing in this role is care for customer outcomes. You build software, solve problems, and keep systems running smoothly for customers, and you own ensuring that what you have built meets their needs.
- You are methodical, calm and curious - able to jump into anything quickly and get up to speed.
- You care deeply about technical quality, automated tests, observability, and the automation that lets us ship fast for customers. You are also pragmatic, favouring simple solutions that work now over engineering for anticipated future needs.
- You enjoy working closely with other engineers, business stakeholders and with customers. You are communicative, comfortable pairing, comfortable thinking out loud, and comfortable being wrong in front of other people.
- You are continually learning, seeking to understand the domain you are working in, the technologies you are working with, and better ways of working. You are always looking to improve yourself and the rest of the team.
- You give feedback freely and kindly, and receive feedback gratefully.
- You are comfortable working in ambiguity and know how to quickly get it to clarity.
Technical Requirements
- You write and read code fluently in at least one of the languages we use (Java, Python, TypeScript, Rust), and you can pick your way through an unfamiliar codebase.
- You have led the delivery of several pieces of significant production grade of software, both greenfield and expanding on an existing software estate.
- You have been on-call for something that mattered, debugged it while it was misbehaving, and made it better afterwards.
- You build with security in mind and recognise common software vulnerabilities.
- You are comfortable in a cloud environment (ours is AWS) and with the tooling that keeps deployments repeatable - infrastructure as code, CI/CD pipelines, containers.
- You know what useful observability looks like: metrics, logs, and traces, and alerts that mean something rather than dashboards nobody reads.
- You take a high degree of ownership, adaptability, and willingness to operate in fast-moving environments with evolving processes and tooling.
- Familiarity with AI agents, workflow orchestration, or operational automation using modern AI tooling.
Why This Role
AI agents are changing how companies operate, and the question of how to govern them safely is only going to grow. Geordie is building the platform enterprises rely on to answer that question.
You will work on a product where the problem space is still being defined, alongside customers who are figuring it out at the same time we are. You will ship real things to real users every week, work closely with people you can learn from, and help shape both the product and the way an early-stage company chooses to build software.
Our Values
Solve What Matters
We listen intently and stay focused on what truly matters to our customers. If it doesn't help them navigate real-world agentic adoption and risk, it's probably not worth building.
Build, Learn, Iterate
We favour momentum over perfection - testing, learning, and evolving through action. It's how we move fast and get better, together.
Kind, Not Comfortable
We're honest, respectful, and unafraid to challenge each other. We believe doing great work should feel good, too. Kindness makes that possible.
A Platform for Your Best Work
We want this to be the most meaningful chapter of your career. Geordie is the place where you grow, lead and lift others, while building something that truly matters.
Similar roles
-
AI Platform / SRE Lead
Glint Tech Solutions LLC New York, New York, United States
-
Engineering Manager, Site Reliability Engineering
Google Sydney, New South Wales, Australia
-
Site Reliability Engineer (SRE)
Devexperts Porto, Portugal
-
Telco SRE Engineer
AST SpaceMobile Rīga, Latvia
-
SITE RELIABILITY ENGINEER II
Stone - Linkedin Hamburg, Germany
-
Principal Site Reliability Engineer
LivePerson Sofia, Sofia-City, Bulgaria