Site Reliability Engineer
Deutsche Bank Bengaluru, Karnataka, India
Financial Services · 10,001+ employees
About the role
The Site Reliability Engineer will ensure the reliability, availability, and performance of production systems through monitoring, incident response, and automation. They will also collaborate with development teams to design scalable infrastructure and manage CI/CD pipelines.
What they look for
Requirements
Candidates must have strong expertise in Google Cloud Platform, Kubernetes, and Infrastructure as Code using Terraform. Proficiency in scripting languages and experience with CI/CD tools and service mesh technologies are also required.
Benefits
Full description
Job Description:
Job Title: Site Reliability Engineer
Corporate Title: Assistant Vice President
Location: Bangalore, India
Role Description
We are looking for Site Reliability Engineer candidate with below requirement. This role is combination of Production support + SRE + Devops. So majorly looking for GCP experience and Kubernetes to support the design, deployment, automation, and operational excellence of enterprise-grade cloud applications.
What we’ll offer you
As part of our flexible scheme, here are just some of the benefits that you’ll enjoy,
- Best in class leave policy.
- Gender neutral parental leaves
- 100% reimbursement under childcare assistance benefit (gender neutral)
- Sponsorship for Industry relevant certifications and education
- Employee Assistance Program for you and your family members
- Comprehensive Hospitalization Insurance for you and your dependents
- Accident and Term life Insurance
- Complementary Health screening for 35 yrs. and above
Your key responsibilities
- System Reliability: Ensure the reliability, availability, and performance of production systems by implementing best practices in monitoring, alerting, and incident response.
- System Maintenance: Understand thoroughly the end-to-end application support process and escalation procedures, become fully conversant with all support tools. Maintain an end-to-end view of the application and infrastructure landscape.
- Automation: Develop and maintain automation tools and scripts to streamline deployment, scaling, and operational tasks.
- Incident Management: Act as a primary responder to system outages and incidents, ensuring rapid resolution and thorough post-mortem analysis to prevent recurrence.
- Monitoring & Alerting: Design and implement robust monitoring and alerting systems to proactively identify and address potential issues.
- Performance Optimization: Identify and resolve performance bottlenecks across the stack, from application code to infrastructure.
- Collaboration: Work closely with development teams and other stakeholders to ensure that new features and services are designed with reliability and scalability in mind.
- Documentation: Maintain comprehensive documentation of systems, processes, and procedures to ensure knowledge sharing and continuity.
- Continuous Improvement: Continuously evaluate and improve our infrastructure, tools, and processes to enhance system reliability and operational efficiency.
- Design, implement, and manage CI/CD pipelines using GitHub Actions.
- Deploy and operate applications on Google Kubernetes Engine (GKE).
- Develop and maintain Helm charts for complex application deployments.
- Manage Kubernetes infrastructure including node management, auto-scaling, configuration management, and secrets management.
- Configure and support service networking components such as gateways, virtual services, and service mesh technologies (Anthos Service Mesh preferred).
Your skills and experience
- Proficiency in Infrastructure as Code - Terraform (must)
- Proficiency in cloud platforms such Google Cloud (preferred), Openshift Cloud
- Usage of enterprise Security Management solutions including GCP Secret Manager.
- Expertise in Kubernetes (GKE) administration and operations
- Experience in CI/CD tools
- GitHub Actions – CI/CD experience is must
- Experience with Docker/Kubernetes (creating images, deployments)
- Experience into developing Helm Charts ( templates, hooks, packaging)
- Exposure to delivering good quality code within enterprise scale development
- Working knowledge of environment monitoring tools such as GCO, Prometheus, Grafana
- Strong experience in software development processes, models, lifecycles and methodologies.
- Expert hands-on experience with service-mesh technology such as Istio or Anthos Service Mesh
- Experience in software development and scripting in at least one language (Java, JavaScript, Python, Go, Bash)
Proven ability to leverage AI tools to enhance productivity, optimise workflows to solve business problems, while applying critical judgment to ensure responsible and ethical use of data and AI outputs.
How we’ll support you
- Training and development to help you excel in your career.
- Coaching and support from experts in your team.
- A culture of continuous learning to aid progression.
- A range of flexible benefits that you can tailor to suit your needs.
About us and our teams
Please visit our company website for further information:
https://www.db.com/company/company.html
We strive for a culture in which we are empowered to excel together every day. This includes acting responsibly, thinking commercially, taking initiative and working collaboratively.
Together we share and celebrate the successes of our people. Together we are Deutsche Bank Group.
We welcome applications from all people and promote a positive, fair and inclusive work environment.
Similar roles
-
Staff Site Reliability Engineer
KEV Group Toronto, Ontario, Canada · $150K–$180K/yr
-
Site Reliability Engineer - Data Platform
IMC Amsterdam, North Holland, Netherlands
-
Staff Software Engineer, Site Reliability Engineering, Vertex AI
Google Warsaw, Masovian Voivodeship, Poland · PLN 480K–PLN 492K/yr
-
Site Reliability Engineering Manager
Conifers.ai Tel-Aviv, Tel-Aviv District, Israel
-
Senior Site Reliability Engineer
Precisely International Jobs Bielsko-Biała, Silesian Voivodeship, Poland
-
Network SRE
JPMorgan Chase & Co. Buenos Aires, Argentina