Staff DevOps Engineer
what3words · Ho Chi Minh City, Ho Chi Minh, Vietnam · $4K/yr
Software Development · 51-200 employees
About the role
You will own the infrastructure as code and build self-service platforms to enable product teams to provision resources independently. Additionally, you will implement reliability and security measures, such as service level objectives and shift-left security, while participating in an on-call rotation.
What they look for
Requirements
The role requires substantial production experience with multi-account AWS, Kubernetes cluster management, and advanced Terraform module authoring. Candidates must also possess strong scripting skills in Python or Go and a deep understanding of GitOps and CI/CD pipelines.
Benefits
Full description
Our mission is to become the global standard for addressing. Street addresses weren’t designed for 2026. They aren’t accurate enough to specify building entrances, and they don’t exist for parks, rural areas and many parts of the world. This makes it hard to find places and causes problems and inefficiencies on a global scale.
That’s why we created what3words. We divided the world into 3m squares and gave each square a unique combination of three words. It’s the easiest way to find and share precise locations.
Over the last year, what3words has been used in 193 countries, and our monthly active users continue to grow at an impressive pace. Our tech is used by emergency services, delivery companies, eCommerce businesses, ride-hailing apps and NGOs, and is integrated into the navigation systems of millions of cars around the world.
The role:
At what3words, we're proud of our tech stack. We have built a massively scalable, microservices-based architecture, based around Kubernetes on AWS. You'll be exposed to technologies such as EKS, Terraform, Terragrunt, Istio, Karpenter, Flux, GitHub Actions, Prometheus, Grafana, Go, Python, Lambda and CloudFront — with a little GCP alongside.
Our DevOps team is deliberately broad: the team owns platform engineering, SRE, observability and security. The remit is wide and the ownership is end-to-end, so you'll carry work from the design call through to how it behaves in production.
You'll work closely with colleagues in our UK office, so you'll need to be effective asynchronously and comfortable making your case in writing.
Responsibilities:
- Own our infrastructure as code as a governed product rather than a pile of configuration — a versioned, tested module library with clear ownership, change control and automated quality gates.
- Build the self-service paths that let product teams provision what they need without waiting on us, then measure adoption and iterate until they're genuinely used.
- Make reliability a property of the delivery path rather than something inspected afterwards, with service level objectives, dashboards and alerting wired in by default.
- Shift Left Security approach for workflows — policy as code, least privilege and secure-by-default templates, so that our teams can build on a foundation of good decisions.
- Treat cloud cost as an engineering responsibility, with commitment and allocation decisions made deliberately.
- Take open-ended platform problems through to durable outcomes against multi-quarter goals, and write the design docs behind decisions that outlast any single project.
- Share an on-call rotation with the team, and improve how we alert and how we respond.
Essential Skills:
- Multi-account AWS: Substantial production experience running AWS at scale across multiple accounts, with the identity and networking discipline that goes with them.
- Kubernetes in Production: Deep, hands-on experience owning cluster lifecycle and version upgrades, tuning auto scaling under real load, and debugging the failures that follow.
- Advanced Terraform: You've authored and versioned modules other people depend on, made deliberate state-management decisions, and run an orchestration or automation layer on top.
- GitOps: Practical experience with declarative delivery, and a clear grasp of what reconciliation gives you and what it takes away when something is wrong.
- CI/CD as a Platform: A track record of owning the pipelines other engineers ship through, including self-hosted runners or build agents and the build times and failure modes that come with them.
- Observability & Incident Response: Experience managing alert systems that the team can trust and be well prepared to respond to and resolve.
- Automation & Security: Strong scripting skills (Shell, plus Python or Go), with least privilege, secrets handling and supply chain awareness built into platform defaults.
- AI tools: Experience and enthusiasm in leveraging Agentic AI workflows to improve developer experience.
- Soft Skills: Excellent communication in spoken and written English.
Diversity & Inclusion
Our mission is to help everyone talk about everywhere, and we believe diverse perspectives make for a better company and better products too. We strongly encourage applications from underrepresented groups and are committed to equality and inclusivity in our hiring processes and company culture.
Benefits
We offer the following benefits to all permanent employees of what3words:
- Competitive salary
- Flexible working
- 6 week remote working (work from anywhere) policy
- 25 days holiday: plus the option to buy more!
- Share options
- Private health insurance
- Wellbeing Days
- Generous parental leave policies
- Family friendly policies
- Employee Assistance Programme (EAP)
- Lunch & learn sessions
- Team social budget