SRE
Zensar Bangalore South, Karnataka, India
IT Services and IT Consulting · 10,001+ employees
About the role
The Site Reliability Engineer will ensure high availability, scalability, and performance of production systems while managing incident response and root cause analysis. They will also automate operational tasks using Infrastructure as Code and maintain cloud infrastructure across AWS, Azure, or GCP.
What they look for
Requirements
Candidates must have a bachelor's degree in Computer Science or a related field and at least 5 years of experience in SRE, DevOps, or Infrastructure Engineering. Proficiency in Linux, cloud platforms, containerization, and scripting languages is required.
Full description
Key Responsibilities
Reliability & Operations
- Ensure high availability, scalability, and performance of production systems.
- Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
- Proactively identify and resolve system bottlenecks and performance issues.
- Perform capacity planning and infrastructure optimization.
Monitoring & Incident Management
- Implement and manage monitoring, logging, and alerting solutions.
- Lead incident response, root cause analysis (RCA), and post-incident reviews.
- Develop automated remediation and self-healing mechanisms.
- Manage on-call support rotations and production support activities.
Automation & Infrastructure
- Automate operational tasks using scripting and Infrastructure as Code (IaC).
- Design and implement CI/CD pipelines to enhance deployment efficiency.
- Standardize infrastructure provisioning and configuration management.
- Drive infrastructure modernization initiatives.
Cloud & Platform Engineering
- Manage cloud infrastructure across AWS, Azure, or GCP environments.
- Optimize cloud resource utilization, security, and cost management.
- Implement containerization and orchestration solutions using Docker and Kubernetes.
- Support hybrid and multi-cloud deployments.
Security & Compliance
- Ensure platform compliance with organizational security standards.
- Implement security best practices, vulnerability remediation, and access controls.
- Participate in disaster recovery planning and business continuity initiatives.
Required Skills
Technical Skills
- Strong experience with Linux/Unix administration.
- Proficiency in one or more programming/scripting languages:• Python
- Shell Scripting
- Go
- Java
- Experience with cloud platforms:• AWS
- Microsoft Azure
- Google Cloud Platform (GCP)
- Hands-on experience with:• Kubernetes
- Docker
- Terraform
- Ansible
- Experience with CI/CD tools:• Jenkins
- GitHub Actions
- GitLab CI/CD
- Azure DevOps
Monitoring & Observability
- Prometheus
- Grafana
- ELK Stack (Elasticsearch, Logstash, Kibana)
- Splunk
- Datadog
- New Relic
Database Knowledge
- SQL Server
- PostgreSQL
- MySQL
- MongoDB
- Redis
Qualifications
- Bachelor's degree in Computer Science, Information Technology, or related field.
- 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
- Experience supporting large-scale enterprise applications.
- Understanding of networking concepts, DNS, load balancing, and security principles.
Preferred Qualifications
- AWS Certified Solutions Architect / DevOps Engineer.
- Azure Administrator or Azure DevOps Engineer Certification.
- Google Professional Cloud DevOps Engineer Certification.
- Kubernetes certifications (CKA/CKAD).
- Experience in enterprise retail, eCommerce, or digital transformation projects.
Soft Skills
- Strong troubleshooting and analytical skills.
- Excellent communication and stakeholder management abilities.
- Ability to work in a fast-paced production environment.
- Strong collaboration and cross-functional teamwork skills.
- Continuous learning and improvement mindset.
Experience
5-10+ Years
Location
Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)
Employment Type
Full-Time
Provide your feedback on BizChat
Add preferred certificationsInclude salary range details
Responsibilities
Key Responsibilities
Reliability & Operations
- Ensure high availability, scalability, and performance of production systems.
- Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
- Proactively identify and resolve system bottlenecks and performance issues.
- Perform capacity planning and infrastructure optimization.
Monitoring & Incident Management
- Implement and manage monitoring, logging, and alerting solutions.
- Lead incident response, root cause analysis (RCA), and post-incident reviews.
- Develop automated remediation and self-healing mechanisms.
- Manage on-call support rotations and production support activities.
Automation & Infrastructure
- Automate operational tasks using scripting and Infrastructure as Code (IaC).
- Design and implement CI/CD pipelines to enhance deployment efficiency.
- Standardize infrastructure provisioning and configuration management.
- Drive infrastructure modernization initiatives.
Cloud & Platform Engineering
- Manage cloud infrastructure across AWS, Azure, or GCP environments.
- Optimize cloud resource utilization, security, and cost management.
- Implement containerization and orchestration solutions using Docker and Kubernetes.
- Support hybrid and multi-cloud deployments.
Security & Compliance
- Ensure platform compliance with organizational security standards.
- Implement security best practices, vulnerability remediation, and access controls.
- Participate in disaster recovery planning and business continuity initiatives.
Required Skills
Technical Skills
- Strong experience with Linux/Unix administration.
- Proficiency in one or more programming/scripting languages:• Python
- Shell Scripting
- Go
- Java
- Experience with cloud platforms:• AWS
- Microsoft Azure
- Google Cloud Platform (GCP)
- Hands-on experience with:• Kubernetes
- Docker
- Terraform
- Ansible
- Experience with CI/CD tools:• Jenkins
- GitHub Actions
- GitLab CI/CD
- Azure DevOps
Monitoring & Observability
- Prometheus
- Grafana
- ELK Stack (Elasticsearch, Logstash, Kibana)
- Splunk
- Datadog
- New Relic
Database Knowledge
- SQL Server
- PostgreSQL
- MySQL
- MongoDB
- Redis
Qualifications
- Bachelor's degree in Computer Science, Information Technology, or related field.
- 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
- Experience supporting large-scale enterprise applications.
- Understanding of networking concepts, DNS, load balancing, and security principles.
Preferred Qualifications
- AWS Certified Solutions Architect / DevOps Engineer.
- Azure Administrator or Azure DevOps Engineer Certification.
- Google Professional Cloud DevOps Engineer Certification.
- Kubernetes certifications (CKA/CKAD).
- Experience in enterprise retail, eCommerce, or digital transformation projects.
Soft Skills
- Strong troubleshooting and analytical skills.
- Excellent communication and stakeholder management abilities.
- Ability to work in a fast-paced production environment.
- Strong collaboration and cross-functional teamwork skills.
- Continuous learning and improvement mindset.
Experience
5-10+ Years
Location
Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)
Employment Type
Full-Time
Provide your feedback on BizChat
Add preferred certificationsInclude salary range details
Qualifications
Key Responsibilities
Reliability & Operations
- Ensure high availability, scalability, and performance of production systems.
- Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
- Proactively identify and resolve system bottlenecks and performance issues.
- Perform capacity planning and infrastructure optimization.
Monitoring & Incident Management
- Implement and manage monitoring, logging, and alerting solutions.
- Lead incident response, root cause analysis (RCA), and post-incident reviews.
- Develop automated remediation and self-healing mechanisms.
- Manage on-call support rotations and production support activities.
Automation & Infrastructure
- Automate operational tasks using scripting and Infrastructure as Code (IaC).
- Design and implement CI/CD pipelines to enhance deployment efficiency.
- Standardize infrastructure provisioning and configuration management.
- Drive infrastructure modernization initiatives.
Cloud & Platform Engineering
- Manage cloud infrastructure across AWS, Azure, or GCP environments.
- Optimize cloud resource utilization, security, and cost management.
- Implement containerization and orchestration solutions using Docker and Kubernetes.
- Support hybrid and multi-cloud deployments.
Security & Compliance
- Ensure platform compliance with organizational security standards.
- Implement security best practices, vulnerability remediation, and access controls.
- Participate in disaster recovery planning and business continuity initiatives.
Required Skills
Technical Skills
- Strong experience with Linux/Unix administration.
- Proficiency in one or more programming/scripting languages:• Python
- Shell Scripting
- Go
- Java
- Experience with cloud platforms:• AWS
- Microsoft Azure
- Google Cloud Platform (GCP)
- Hands-on experience with:• Kubernetes
- Docker
- Terraform
- Ansible
- Experience with CI/CD tools:• Jenkins
- GitHub Actions
- GitLab CI/CD
- Azure DevOps
Monitoring & Observability
- Prometheus
- Grafana
- ELK Stack (Elasticsearch, Logstash, Kibana)
- Splunk
- Datadog
- New Relic
Database Knowledge
- SQL Server
- PostgreSQL
- MySQL
- MongoDB
- Redis
Qualifications
- Bachelor's degree in Computer Science, Information Technology, or related field.
- 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
- Experience supporting large-scale enterprise applications.
- Understanding of networking concepts, DNS, load balancing, and security principles.
Preferred Qualifications
- AWS Certified Solutions Architect / DevOps Engineer.
- Azure Administrator or Azure DevOps Engineer Certification.
- Google Professional Cloud DevOps Engineer Certification.
- Kubernetes certifications (CKA/CKAD).
- Experience in enterprise retail, eCommerce, or digital transformation projects.
Soft Skills
- Strong troubleshooting and analytical skills.
- Excellent communication and stakeholder management abilities.
- Ability to work in a fast-paced production environment.
- Strong collaboration and cross-functional teamwork skills.
- Continuous learning and improvement mindset.
Experience
5-10+ Years
Location
Bangalore / Hyderabad / Chennai / Pune (Hybrid/Remote)
Employment Type
Full-Time
Provide your feedback on BizChat
Add preferred certificationsInclude salary range details
At Zensar, we’re “experience-led everything”. We are committed to conceptualizing, designing, engineering, marketing, and managing digital solutions and experiences for over 130 leading enterprises. We are a company driven by a bold purpose: Together, we shape experiences for better futures. Whether for our clients, our people, or the world around us, this belief powers everything we do. At the heart of our culture is ONE with Client - a set of four core values that reflect who we are and how we work: One Zensar, Nurturing, Empowering, and Client Focus.
Part of the $4.8 billion RPG Group, we’re a community of 10,000+ innovators across 30+ global locations, including Milpitas, Seattle, Princeton, Cape Town, London, Zurich, Singapore, and Mexico City. Explore Life at Zensar and join us to Grow. Own. Achieve. Learn. to be the best version of yourself.
We believe the best work happens when individuality is celebrated, growth is encouraged, and well-being is prioritized. We are an equal employment opportunity (EEO) and affirmative action employer, committed to creating an inclusive workplace. All qualified applicants will be considered without regard to race, creed, color, ancestry, religion, sex, national origin, citizenship, age, sexual orientation, gender identity, disability, marital status, family medical leave status, or protected veteran status.
Similar roles
-
Site Reliability Engineer
Armor Defense Inc Pune, Maharashtra, India
-
Senior Site Reliability Engineer
Formation Bio San Francisco, California, United States · $186K–$232K/yr
-
Staff Site Reliability Engineer
Okta Dublin, Leinster, Ireland · €92K–€126K/yr
-
Senior Site Reliability Engineer
Okta Dublin, Leinster, Ireland · €76K–€104K/yr
-
Site Reliability Engineer Engineer
Modus Create United States
-
Staff Site Reliability Engineer
Renesas Electronics San Diego, California, United States · $170K–$210K/yr