SimCorp

Lead Site Reliability Engineer

SimCorp Hyderabad, Telangana, India

Software Development · 1,001-5,000 employees

Yesterday
sre Senior (5-10 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Lead Site Reliability Engineer will build and operate Azure-based cloud infrastructure, Kubernetes platforms, and CI/CD pipelines while managing incident response and system reliability. They will also define SLIs/SLOs, drive cost optimization through FinOps, and mentor junior engineers to foster operational excellence.

What they look for

Azure Kubernetes CI/CD Site Reliability Engineering Cloud infrastructure FinOps Incident response Monitoring Automation Capacity planning Root cause analysis Mentorship Scalability Performance optimization AI-driven engineering

Requirements

The role requires hands-on experience in cloud infrastructure, container orchestration, and automation to ensure system scalability and performance. Candidates should possess strong leadership skills to guide teams, manage on-call rotations, and implement AI-driven engineering solutions to reduce operational toil.

Benefits

Attractive salary Bonus scheme Pension Flexible working hours Hybrid work model Professional development opportunities

Full description

Lead Site Reliability Engineer

WHAT MAKES US, US

Join some of the most innovative thinkers in FinTech as we lead the evolution of financial technology. If you are an innovative, curious, collaborative person who embraces challenges and wants to grow, learn, and pursue outcomes with our prestigious financial clients, say Hello to SimCorp!

At its foundation, SimCorp is guided by our values — caring, customer success-driven, collaborative, curious, and courageous. Our people-centered organization focuses on skills development, relationship building, and client success. We take pride in cultivating an environment where all team members can grow, feel heard, valued, and empowered.

If you like what we’re saying, keep reading!

WHY THIS ROLE IS IMPORTANT TO US

As a Lead Site Reliability Engineer, you will join one of our Product Areas and become part of a collaborative team focused on operating and improving our Azure-based SaaS platform. You’ll work with experienced engineers and cross-functional colleagues to help ensure reliability, scalability, and operational excellence for both onboarding and running client environments.

This role is a launchpad for your development as an SRE—combining hands-on engineering, automation, monitoring, and platform support—with strong mentorship and continuous learning opportunities built in.

What you will be responsible for

  • Work hands-on across platform and infrastructure SRE activities — cloud infrastructure, Kubernetes/container platforms, CI/CD, and production systems — not just reviewing or delegating this work, but building and operating it directly.
  • Define, own, and report on SLIs, SLOs, and SLAs for critical services; manage error budgets and use them to drive concrete engineering and release decisions.
  • Actively participate in cost optimization efforts across cloud and platform infrastructure, identifying and implementing efficiency and rightsizing opportunities (FinOps practices).
  • Lead the design and implementation of strategies to ensure the reliability, scalability, and performance of critical systems and services.
  • Manage incident response, troubleshooting, and root cause analysis for system outages and performance issues, ensuring timely resolution and prevention of future incidents.
  • Develop and maintain monitoring, alerting, and automation systems to enhance service reliability and reduce manual intervention, including automation and AI-driven engineering approaches (e.g. AI-assisted anomaly detection, automated remediation, AI-supported incident triage) to reduce toil and speed up resolution.
  • Collaborate with development, operations, and product teams to optimize application performance and system infrastructure.
  • Drive initiatives for capacity planning, resource management, and scalability to meet growing business needs.
  • Implement continuous improvement processes to enhance the reliability and efficiency of systems and services.

What we value

  • Mentor and guide junior site reliability engineers, fostering a culture of collaboration and knowledge-sharing, while remaining personally hands-on with the underlying systems.
  • Participate in on-call rotations, providing leadership during critical incidents and ensuring minimal downtime.
  • Create and maintain documentation for incident management, system configurations, and operational processes.
  • Drive the adoption of industry best practices, tools, and technologies — including automation and AI-driven engineering tooling — to enhance system reliability and operational performance.

BENEFITS 

Attractive salary, bonus scheme, and pension are essential for any work agreement. However, in SimCorp we believe we can offer more. Therefore, in addition to the traditional benefit scheme, we provide a good work-life balance: flexible working hours and a hybrid model. Simcorp follows a global hybrid policy, asking employees to work from the office two days each week while allowing remote work on other days.

Simcorp does offer opportunities for professional development: there is never just only one route - we offer an individual approach to professional development to support the direction you want to take.

NEXT STEPS  

Please send us your application in English via our career site as soon as possible, we process incoming applications continually. Please note that only applications sent through our system will be processed. At SimCorp, we recognize that bias can unintentionally occur in the recruitment process. To uphold fairness and equal opportunities for all applicants, we kindly ask you to exclude personal data such as photos, age, or any non-professional information from your application. Thank you for aiding us in our endeavor to mitigate biases in our recruitment process. 

If you are interested in being a part of SimCorp but are not sure this role is suitable, submit your CV anyway. SimCorp is on a positive growth journey, and our Talent Acquisition Team is ready to assist you discover the right role for you. The approximate time to consider your CV is three weeks.  

We are focused on continually improving our talent acquisition process and making everyone’s experience positive and valuable. Therefore, during the process we will ask you to provide your feedback, which is highly appreciated. 

WHO WE ARE 

For over 50 years, we have worked closely with investment and asset managers to become the world’s leading provider of integrated investment management solutions. We are 3,000+ colleagues with a broad range of nationalities, educations, professional experiences, ages, and backgrounds. 

SimCorp is an independent subsidiary of the Deutsche Börse Group. Following the recent merger with Axioma, we leverage the combined strength of our brands to provide an industry-leading, full, front-to-back offering for our clients. 

SimCorp is an equal opportunity employer and welcome applicants from all backgrounds, without regard to race, gender, age, disability, or any other protected status under applicable law. We are committed to building a culture where diverse perspectives and expertise are integrated into our everyday work. We believe in the continual growth and development of our employees, so that we can provide best-in-class solutions to our clients.  

#Li-Hybrid

Similar roles