Jobgether

Senior Site Reliability Engineer

Jobgether India

Internet Marketplace Platforms · 11-50 employees

19 h ago
sre Principal (10+ yrs) Full-time India
Log in to apply, save this posting, or score it against your profile with AI.

About the role

You will architect and operate secure, highly available AWS infrastructure while driving reliability, scalability, and cost optimization. The role involves leading technical initiatives, managing containerized environments, and mentoring engineering teams to ensure operational excellence.

What they look for

AWS Kubernetes Terraform Jenkins DataDog Python Bash Infrastructure as Code CI/CD Cloud Architecture Site Reliability Engineering Ansible Security Compliance Mentorship System Troubleshooting

Requirements

Candidates must have 9–12 years of professional experience in SRE or Cloud Engineering with deep expertise in AWS and Infrastructure as Code. Strong proficiency in Kubernetes, CI/CD pipelines, and scripting languages like Python or Bash is required, along with experience in regulated environments.

Benefits

Health insurance Accidental insurance Life insurance Complimentary office meals Wellness allowance Paid leave Parental leave Bereavement leave Medical leave Celebration leave Company-paid holidays

Full description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer based in India.

As a Senior Site Reliability Engineer II, you will help architect and operate secure, highly available cloud infrastructure supporting business-critical healthcare applications.You will take strategic ownership of AWS environments, driving reliability, scalability, performance, and cost optimization.The role combines hands-on engineering with technical leadership across Kubernetes, CI/CD, observability, and infrastructure automation.You will strengthen cloud operations through Infrastructure as Code, proactive monitoring, and resilient deployment practices.Security and compliance are central, with responsibility for maintaining rigorous HIPAA, GDPR, and SOC 2 standards.You will also mentor engineers, lead complex initiatives, and influence technical strategy across cross-functional teams.This is an opportunity to make a direct impact on healthcare technology while working in a collaborative, innovation-focused environment.

\n

Accountabilities:

  • Cloud Architecture & Reliability: Design, deploy, and continuously improve secure, scalable, and fault-tolerant AWS infrastructure, with a focus on availability, resilience, performance, and cost efficiency.
  • Infrastructure Operations: Manage and optimize AWS services including EC2, S3, Lambda, and RDS while improving resource utilization and operational efficiency.
  • Observability: Enhance and maintain monitoring and observability capabilities using DataDog, enabling proactive issue detection, performance analysis, and deep visibility across cloud environments.
  • CI/CD & Deployment: Lead the evolution of Jenkins-based deployment pipelines, improving automation, release reliability, deployment velocity, and engineering confidence.
  • Kubernetes: Manage and optimize containerized environments, improving scalability, consistency, resilience, and deployment practices.
  • Infrastructure as Code: Champion automation through Terraform, Ansible, and related technologies to reduce manual effort, standardize infrastructure, and improve operational efficiency.
  • Security & Compliance: Ensure cloud operations and infrastructure adhere to stringent security and regulatory requirements, including HIPAA, GDPR, and SOC 2.
  • Technical Leadership: Lead complex engineering initiatives, influence the technical roadmap, and make architecture decisions that strengthen long-term platform reliability.
  • Mentorship: Coach and mentor engineers, share technical expertise, encourage strong engineering practices, and foster a collaborative culture.
  • Problem Resolution & Continuous Improvement: Proactively identify reliability risks, investigate complex incidents, and architect durable solutions that prevent recurring operational issues.

Requirements:

  • Experience: 9–12 years of professional experience in Site Reliability Engineering, Cloud Engineering, or a closely related discipline, with demonstrated ownership of large-scale AWS environments.
  • AWS Expertise: Strong hands-on knowledge of AWS services such as EC2, S3, Lambda, and RDS, including experience with cost optimization and resource management.
  • SRE Tooling: Proven experience with Kubernetes, DataDog, Jenkins, and modern cloud-native operational practices.
  • Automation & Coding: Strong scripting capabilities in Python or Bash and professional experience with Infrastructure as Code tools, particularly Terraform.
  • Cloud & DevOps: Strong understanding of cloud architecture, deployment automation, CI/CD, containerization, scalability, availability, and production operations.
  • Security & Compliance: Experience implementing secure cloud operations and working with regulated environments or compliance frameworks is highly valuable.
  • Problem Solving: Strong analytical and troubleshooting skills, combined with an ownership mindset and the ability to design systems that proactively prevent failures.
  • Leadership: Demonstrated ability to lead technical initiatives, mentor engineers, influence engineering standards, and communicate effectively with technical and business stakeholders.
  • Communication: Excellent written and verbal communication skills, with the ability to translate complex technical concepts into clear objectives and recommendations.
  • Preferred Qualifications: AWS professional-level certifications, experience with serverless architectures or technologies such as Kafka, Kinesis, or Redshift, and previous healthcare technology or high-security environment experience are advantageous.

Benefits:

  • Hybrid working environment in Hyderabad, with flexibility designed to support effective ways of working.
  • Competitive benefits package designed to support health, well-being, and financial security.
  • Comprehensive health, accidental, and life insurance coverage, including family coverage.
  • Complimentary office lunches and dinners on select days, along with healthy workplace snacks.
  • Annual wellness allowance supporting employee well-being and productivity.
  • Earned, casual, and sick leave to support work-life balance.
  • Paid parental leave covering maternity, paternity, adoption, surrogacy, and abortion leave.
  • Bereavement and extended medical leave options.
  • Celebration leave and company-paid holidays.
  • Opportunity to influence cloud and SRE strategy at a senior technical level.
  • Meaningful impact on healthcare technology and systems that support improved patient outcomes.
  • Strong opportunities for technical leadership, innovation, mentorship, and professional growth.
  • Collaborative culture focused on engineering excellence, customer impact, and continuous improvement.

\nHow Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

#LI-CL1

Similar roles