Zensar

SRE

Zensar Bangalore South, Karnataka, India

IT Services and IT Consulting · 10,001+ employees

5 h ago Closes in 6d
sre Senior (5-10 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Site Reliability Engineer will ensure high availability, scalability, and performance of production systems while managing incident response and root cause analysis. They will also automate operational tasks using Infrastructure as Code and maintain cloud infrastructure across AWS, Azure, or GCP.

What they look for

Linux Python Shell Scripting Go Java AWS Azure GCP Kubernetes Docker Terraform Ansible Jenkins Prometheus Grafana SQL

Requirements

Candidates must have a bachelor's degree in Computer Science or a related field and at least 5 years of experience in SRE, DevOps, or Infrastructure Engineering. Proficiency in Linux, cloud platforms, containerization, and scripting languages is required.

Full description

Key Responsibilities

Reliability & Operations

  • Ensure high availability, scalability, and performance of production systems.
  • Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
  • Proactively identify and resolve system bottlenecks and performance issues.
  • Perform capacity planning and infrastructure optimization.

Monitoring & Incident Management

  • Implement and manage monitoring, logging, and alerting solutions.
  • Lead incident response, root cause analysis (RCA), and post-incident reviews.
  • Develop automated remediation and self-healing mechanisms.
  • Manage on-call support rotations and production support activities.

Automation & Infrastructure

  • Automate operational tasks using scripting and Infrastructure as Code (IaC).
  • Design and implement CI/CD pipelines to enhance deployment efficiency.
  • Standardize infrastructure provisioning and configuration management.
  • Drive infrastructure modernization initiatives.

Cloud & Platform Engineering

  • Manage cloud infrastructure across AWS, Azure, or GCP environments.
  • Optimize cloud resource utilization, security, and cost management.
  • Implement containerization and orchestration solutions using Docker and Kubernetes.
  • Support hybrid and multi-cloud deployments.

Security & Compliance

  • Ensure platform compliance with organizational security standards.
  • Implement security best practices, vulnerability remediation, and access controls.
  • Participate in disaster recovery planning and business continuity initiatives.

Required Skills

Technical Skills

  • Strong experience with Linux/Unix administration.
  • Proficiency in one or more programming/scripting languages:• Python
  • Shell Scripting
  • Go
  • Java
  • Experience with cloud platforms:• AWS
  • Microsoft Azure
  • Google Cloud Platform (GCP)
  • Hands-on experience with:• Kubernetes
  • Docker
  • Terraform
  • Ansible
  • Experience with CI/CD tools:• Jenkins
  • GitHub Actions
  • GitLab CI/CD
  • Azure DevOps

Monitoring & Observability

  • Prometheus
  • Grafana
  • ELK Stack (Elasticsearch, Logstash, Kibana)
  • Splunk
  • Datadog
  • New Relic

Database Knowledge

  • SQL Server
  • PostgreSQL
  • MySQL
  • MongoDB
  • Redis

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or related field.
  • 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
  • Experience supporting large-scale enterprise applications.
  • Understanding of networking concepts, DNS, load balancing, and security principles.

Preferred Qualifications

  • AWS Certified Solutions Architect / DevOps Engineer.
  • Azure Administrator or Azure DevOps Engineer Certification.
  • Google Professional Cloud DevOps Engineer Certification.
  • Kubernetes certifications (CKA/CKAD).
  • Experience in enterprise retail, eCommerce, or digital transformation projects.

Soft Skills

  • Strong troubleshooting and analytical skills.
  • Excellent communication and stakeholder management abilities.
  • Ability to work in a fast-paced production environment.
  • Strong collaboration and cross-functional teamwork skills.
  • Continuous learning and improvement mindset.

Experience

5-10+ Years

Location

Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)

Employment Type

Full-Time

Provide your feedback on BizChat

Add preferred certificationsInclude salary range details

Responsibilities

Key Responsibilities

Reliability & Operations

  • Ensure high availability, scalability, and performance of production systems.
  • Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
  • Proactively identify and resolve system bottlenecks and performance issues.
  • Perform capacity planning and infrastructure optimization.

Monitoring & Incident Management

  • Implement and manage monitoring, logging, and alerting solutions.
  • Lead incident response, root cause analysis (RCA), and post-incident reviews.
  • Develop automated remediation and self-healing mechanisms.
  • Manage on-call support rotations and production support activities.

Automation & Infrastructure

  • Automate operational tasks using scripting and Infrastructure as Code (IaC).
  • Design and implement CI/CD pipelines to enhance deployment efficiency.
  • Standardize infrastructure provisioning and configuration management.
  • Drive infrastructure modernization initiatives.

Cloud & Platform Engineering

  • Manage cloud infrastructure across AWS, Azure, or GCP environments.
  • Optimize cloud resource utilization, security, and cost management.
  • Implement containerization and orchestration solutions using Docker and Kubernetes.
  • Support hybrid and multi-cloud deployments.

Security & Compliance

  • Ensure platform compliance with organizational security standards.
  • Implement security best practices, vulnerability remediation, and access controls.
  • Participate in disaster recovery planning and business continuity initiatives.

Required Skills

Technical Skills

  • Strong experience with Linux/Unix administration.
  • Proficiency in one or more programming/scripting languages:• Python
  • Shell Scripting
  • Go
  • Java
  • Experience with cloud platforms:• AWS
  • Microsoft Azure
  • Google Cloud Platform (GCP)
  • Hands-on experience with:• Kubernetes
  • Docker
  • Terraform
  • Ansible
  • Experience with CI/CD tools:• Jenkins
  • GitHub Actions
  • GitLab CI/CD
  • Azure DevOps

Monitoring & Observability

  • Prometheus
  • Grafana
  • ELK Stack (Elasticsearch, Logstash, Kibana)
  • Splunk
  • Datadog
  • New Relic

Database Knowledge

  • SQL Server
  • PostgreSQL
  • MySQL
  • MongoDB
  • Redis

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or related field.
  • 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
  • Experience supporting large-scale enterprise applications.
  • Understanding of networking concepts, DNS, load balancing, and security principles.

Preferred Qualifications

  • AWS Certified Solutions Architect / DevOps Engineer.
  • Azure Administrator or Azure DevOps Engineer Certification.
  • Google Professional Cloud DevOps Engineer Certification.
  • Kubernetes certifications (CKA/CKAD).
  • Experience in enterprise retail, eCommerce, or digital transformation projects.

Soft Skills

  • Strong troubleshooting and analytical skills.
  • Excellent communication and stakeholder management abilities.
  • Ability to work in a fast-paced production environment.
  • Strong collaboration and cross-functional teamwork skills.
  • Continuous learning and improvement mindset.

Experience

5-10+ Years

Location

Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)

Employment Type

Full-Time

Provide your feedback on BizChat

Add preferred certificationsInclude salary range details

Qualifications

Key Responsibilities

Reliability & Operations

  • Ensure high availability, scalability, and performance of production systems.
  • Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
  • Proactively identify and resolve system bottlenecks and performance issues.
  • Perform capacity planning and infrastructure optimization.

Monitoring & Incident Management

  • Implement and manage monitoring, logging, and alerting solutions.
  • Lead incident response, root cause analysis (RCA), and post-incident reviews.
  • Develop automated remediation and self-healing mechanisms.
  • Manage on-call support rotations and production support activities.

Automation & Infrastructure

  • Automate operational tasks using scripting and Infrastructure as Code (IaC).
  • Design and implement CI/CD pipelines to enhance deployment efficiency.
  • Standardize infrastructure provisioning and configuration management.
  • Drive infrastructure modernization initiatives.

Cloud & Platform Engineering

  • Manage cloud infrastructure across AWS, Azure, or GCP environments.
  • Optimize cloud resource utilization, security, and cost management.
  • Implement containerization and orchestration solutions using Docker and Kubernetes.
  • Support hybrid and multi-cloud deployments.

Security & Compliance

  • Ensure platform compliance with organizational security standards.
  • Implement security best practices, vulnerability remediation, and access controls.
  • Participate in disaster recovery planning and business continuity initiatives.

Required Skills

Technical Skills

  • Strong experience with Linux/Unix administration.
  • Proficiency in one or more programming/scripting languages:• Python
  • Shell Scripting
  • Go
  • Java
  • Experience with cloud platforms:• AWS
  • Microsoft Azure
  • Google Cloud Platform (GCP)
  • Hands-on experience with:• Kubernetes
  • Docker
  • Terraform
  • Ansible
  • Experience with CI/CD tools:• Jenkins
  • GitHub Actions
  • GitLab CI/CD
  • Azure DevOps

Monitoring & Observability

  • Prometheus
  • Grafana
  • ELK Stack (Elasticsearch, Logstash, Kibana)
  • Splunk
  • Datadog
  • New Relic

Database Knowledge

  • SQL Server
  • PostgreSQL
  • MySQL
  • MongoDB
  • Redis

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or related field.
  • 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
  • Experience supporting large-scale enterprise applications.
  • Understanding of networking concepts, DNS, load balancing, and security principles.

Preferred Qualifications

  • AWS Certified Solutions Architect / DevOps Engineer.
  • Azure Administrator or Azure DevOps Engineer Certification.
  • Google Professional Cloud DevOps Engineer Certification.
  • Kubernetes certifications (CKA/CKAD).
  • Experience in enterprise retail, eCommerce, or digital transformation projects.

Soft Skills

  • Strong troubleshooting and analytical skills.
  • Excellent communication and stakeholder management abilities.
  • Ability to work in a fast-paced production environment.
  • Strong collaboration and cross-functional teamwork skills.
  • Continuous learning and improvement mindset.

Experience

5-10+ Years

Location

Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)

Employment Type

Full-Time

Provide your feedback on BizChat

Add preferred certificationsInclude salary range details

At Zensar, we’re “experience-led everything”. We are committed to conceptualizing, designing, engineering, marketing, and managing digital solutions and experiences for over 130 leading enterprises. We are a company driven by a bold purpose: Together, we shape experiences for better futures. Whether for our clients, our people, or the world around us, this belief powers everything we do. At the heart of our culture is ONE with Client - a set of four core values that reflect who we are and how we work: One Zensar, Nurturing, Empowering, and Client Focus.

Part of the $4.8 billion RPG Group, we’re a community of 10,000+ innovators across 30+ global locations, including Milpitas, Seattle, Princeton, Cape Town, London, Zurich, Singapore, and Mexico City. Explore Life at Zensar and join us to Grow. Own. Achieve. Learn. to be the best version of yourself.

We believe the best work happens when individuality is celebrated, growth is encouraged, and well-being is prioritized. We are an equal employment opportunity (EEO) and affirmative action employer, committed to creating an inclusive workplace. All qualified applicants will be considered without regard to race, creed, color, ancestry, religion, sex, national origin, citizenship, age, sexual orientation, gender identity, disability, marital status, family medical leave status, or protected veteran status.

Similar roles