Deutsche Bank

Site Reliability Engineer

Deutsche Bank Bengaluru, Karnataka, India

Financial Services · 10,001+ employees

20 h ago
sre Senior (5-10 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Site Reliability Engineer will ensure the reliability, availability, and performance of production systems through monitoring, incident response, and automation. They will also collaborate with development teams to design scalable infrastructure and manage CI/CD pipelines.

What they look for

GCP Kubernetes Terraform GitHub Actions Docker Helm Istio Anthos Service Mesh Java JavaScript Python Go Bash Prometheus Grafana CI/CD

Requirements

Candidates must have strong expertise in Google Cloud Platform, Kubernetes, and Infrastructure as Code using Terraform. Proficiency in scripting languages and experience with CI/CD tools and service mesh technologies are also required.

Benefits

Leave policy Gender neutral parental leaves Childcare assistance Industry relevant certifications sponsorship Employee Assistance Program Hospitalization insurance Accident insurance Term life insurance Health screening

Full description

Job Description:

Job Title: Site Reliability Engineer

Corporate Title: Assistant Vice President

Location: Bangalore, India

Role Description

We are looking for Site Reliability Engineer candidate with below requirement. This role is combination of Production support + SRE + Devops. So majorly looking for GCP experience and Kubernetes to support the design, deployment, automation, and operational excellence of enterprise-grade cloud applications.

What we’ll offer you

As part of our flexible scheme, here are just some of the benefits that you’ll enjoy,

  • Best in class leave policy.
  • Gender neutral parental leaves
  • 100% reimbursement under childcare assistance benefit (gender neutral)
  • Sponsorship for Industry relevant certifications and education
  • Employee Assistance Program for you and your family members
  • Comprehensive Hospitalization Insurance for you and your dependents
  • Accident and Term life Insurance
  • Complementary Health screening for 35 yrs. and above

Your key responsibilities

  • System Reliability: Ensure the reliability, availability, and performance of production systems by implementing best practices in monitoring, alerting, and incident response.
  • System Maintenance: Understand thoroughly the end-to-end application support process and escalation procedures, become fully conversant with all support tools. Maintain an end-to-end view of the application and infrastructure landscape.
  • Automation: Develop and maintain automation tools and scripts to streamline deployment, scaling, and operational tasks.
  • Incident Management: Act as a primary responder to system outages and incidents, ensuring rapid resolution and thorough post-mortem analysis to prevent recurrence.
  • Monitoring & Alerting: Design and implement robust monitoring and alerting systems to proactively identify and address potential issues.
  • Performance Optimization: Identify and resolve performance bottlenecks across the stack, from application code to infrastructure.
  • Collaboration: Work closely with development teams and other stakeholders to ensure that new features and services are designed with reliability and scalability in mind.
  • Documentation: Maintain comprehensive documentation of systems, processes, and procedures to ensure knowledge sharing and continuity.
  • Continuous Improvement: Continuously evaluate and improve our infrastructure, tools, and processes to enhance system reliability and operational efficiency.
  • Design, implement, and manage CI/CD pipelines using GitHub Actions.
  • Deploy and operate applications on Google Kubernetes Engine (GKE).
  • Develop and maintain Helm charts for complex application deployments.
  • Manage Kubernetes infrastructure including node management, auto-scaling, configuration management, and secrets management.
  • Configure and support service networking components such as gateways, virtual services, and service mesh technologies (Anthos Service Mesh preferred).

Your skills and experience

  • Proficiency in Infrastructure as Code - Terraform (must)
  • Proficiency in cloud platforms such Google Cloud (preferred), Openshift Cloud
  • Usage of enterprise Security Management solutions including GCP Secret Manager.
  • Expertise in Kubernetes (GKE) administration and operations
  • Experience in CI/CD tools
  • GitHub Actions – CI/CD experience is must
  • Experience with Docker/Kubernetes (creating images, deployments)
  • Experience into developing Helm Charts ( templates, hooks, packaging)
  • Exposure to delivering good quality code within enterprise scale development
  • Working knowledge of environment monitoring tools such as GCO, Prometheus, Grafana
  • Strong experience in software development processes, models, lifecycles and methodologies.
  • Expert hands-on experience with service-mesh technology such as Istio or Anthos Service Mesh
  • Experience in software development and scripting in at least one language (Java, JavaScript, Python, Go, Bash)

Proven ability to leverage AI tools to enhance productivity, optimise workflows to solve business problems, while applying critical judgment to ensure responsible and ethical use of data and AI outputs.

How we’ll support you

  • Training and development to help you excel in your career.
  • Coaching and support from experts in your team.
  • A culture of continuous learning to aid progression.
  • A range of flexible benefits that you can tailor to suit your needs.

About us and our teams

Please visit our company website for further information:

https://www.db.com/company/company.html

We strive for a culture in which we are empowered to excel together every day. This includes acting responsibly, thinking commercially, taking initiative and working collaboratively.

Together we share and celebrate the successes of our people. Together we are Deutsche Bank Group.

We welcome applications from all people and promote a positive, fair and inclusive work environment.

Similar roles