Site Reliability Engineering Lead
Jobgether United States · $112K–$264K/yr
Internet Marketplace Platforms · 11-50 employees
About the role
Lead and support a distributed SRE team across the United States and India while establishing effective architectural patterns for cloud-native environments. Drive the design, implementation, and automation of scalable infrastructure to ensure reliability, security, and operational excellence.
What they look for
Requirements
Requires extensive experience in cloud architecture, infrastructure migration, and hands-on proficiency with tools like Terraform and GitHub Actions. Candidates must demonstrate strong technical leadership, mentoring capabilities, and a systematic approach to solving complex engineering challenges.
Benefits
Full description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Site Reliability Engineering Lead based in United States.
This role leads a distributed SRE team across the United States and India, shaping reliable and secure cloud-native environments. You will guide the design and implementation of scalable infrastructure while helping teams adopt modern engineering practices. The position serves as a key technical partner for SRE and DevOps teams, removing blockers and establishing effective architectural patterns. A strong focus on automation, infrastructure as code, security, observability, and operational excellence is central to the role. You will mentor engineers through complex infrastructure and coding challenges while promoting systematic, collaborative ways of working. The role offers the opportunity to influence new cloud landscapes, SaaS solutions, and large-scale migration initiatives. This is a full-time U.S.-based position with a distributed, cross-functional working environment.
\n
Accountabilities:
- Lead and support a distributed SRE team operating across the United States and India.
- Serve as a key point of contact for the Core SRE organization and other product-focused DevOps and SRE teams.
- Deliver highly elastic, resilient, and cloud-native infrastructure solutions and reusable architectural patterns.
- Design and implement cloud environments that meet scalability, reliability, security, and business requirements.
- Support the migration of applications and infrastructure from on-premises environments to the cloud.
- Help teams adopt new cloud landscapes, platforms, SaaS solutions, and infrastructure approaches.
- Establish self-service pipelines and automation that empower stakeholders while reducing operational toil.
- Ensure infrastructure is fully managed through Infrastructure as Code and can be rebuilt from the ground up without manual cloud-portal configuration.
- Mentor engineers in resolving deep technical issues, advanced cloud infrastructure challenges, and complex coding problems.
- Provide technical guidance and promote consistent engineering and architectural best practices.
- Conduct thorough code reviews with a focus on quality, rigor, maintainability, and best practices.
- Maintain a strong focus on information security across diverse technologies and solutions.
- Collaborate with project managers and stakeholders to provide clear project status, reporting, and progress updates.
- Act as an ambassador across teams, building common ground and establishing clear working agreements.
- Identify and remove technical or organizational blockers affecting delivery.
- Drive projects toward agreed schedules and milestones.
- Apply continuous improvement techniques across engineering and operational processes.
- Automate repetitive operational activities and continuously identify opportunities to improve performance, cost, and reliability.
- Encourage knowledge sharing, cross-training, and collaborative problem-solving across the team.
- Help establish a supportive and forward-thinking engineering environment that leverages team members’ diverse strengths.
Requirements:
- Demonstrated experience leading or designing application and/or infrastructure migration projects from on-premises environments to the cloud.
- Proven ability to partner with and lead technical resources in solving complex business and technology challenges.
- Strong experience designing and implementing cloud architectures aligned with business needs and engineering objectives.
- Hands-on experience with Terraform and GitHub Actions.
- Strong understanding of cloud-native architecture, scalable infrastructure, automation, and reliability engineering practices.
- Experience with Infrastructure as Code and designing environments that can be consistently rebuilt without manual configuration.
- Experience working with distributed SRE, DevOps, infrastructure, or platform engineering teams.
- Strong technical leadership and mentoring capabilities, with the ability to guide engineers through complex technical problems.
- Ability to evaluate architectural approaches and establish practical engineering standards and patterns.
- Strong understanding of security considerations across cloud infrastructure and modern technology environments.
- Experience working cross-functionally with engineering teams, project managers, stakeholders, and business partners.
- Strong communication and collaboration skills, with the ability to establish alignment across teams.
- Systematic and methodical approach to technical execution, troubleshooting, and operational improvement.
- Ability to manage competing priorities, identify blockers, and drive projects through to completion.
- Experience working with Azure and related cloud-native technologies is highly relevant.
- Experience with Kubernetes, particularly Azure Kubernetes Service, is preferred.
- Familiarity with Argo, Helm, Docker, JFrog, Grafana, networking, information security, and Redis is preferred.
- Ability to operate effectively in a fast-paced, evolving environment while maintaining a focus on reliability, security, and quality.
Benefits:
- U.S. national base salary range of $118,300–$219,800.
- Geographic salary differentials may apply depending on location.
- Location-specific ranges include:
- Colorado: $118,300–$219,800.
- Illinois: $124,200–$230,800.
- Chicago, Illinois: $130,200–$241,800.
- Maryland: $124,200–$230,800.
- New York: $130,200–$241,800.
- New York City: $142,000–$263,800.
- Rochester, New York: $118,300–$219,800.
- Ohio: $112,400–$208,800.
- New Jersey: $149,765–$239,235.
- Eligibility for an annual incentive bonus.
- Medical inpatient and outpatient insurance.
- Life assurance coverage.
- Family benefits supporting maternity, paternity, and adoption.
- Long-service recognition awards.
- Celebratory allowances and gifts for eligible occasions.
- Flexible benefits plan providing access to a broader range of services and products.
- Employee Assistance Program for personal and work-related support.
- Flexible working arrangements.
- Access to learning and professional development resources.
- Support for building technical skills through team learning and cross-training.
- Remote work opportunity for U.S.-based employees.
- Application deadline: October 5, 2026.
\nHow Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1
Similar roles
-
ART 1486 - Application Engineer (Site Reliability)
FPT Asia Pacific Pte Ltd Singapore, Singapore
-
Senior Site Reliability Engineer
Gen Digital Inc. Kuala Lumpur, Kuala Lumpur, Malaysia
-
Site Reliability Engineer
Chevron Makati City, National Capital District, Philippines
-
Site Reliability Engineer
Binance Asiago, Veneto, Italy
-
Lead Site Reliability Engineer
JPMorgan Chase & Co. Columbus, Ohio, United States
-
Site Reliability Engineer (Europe)
Arango Lund, Sweden