73 Strings

Senior Site Reliability Engineer - Application Support (SRE)

73 Strings Bengaluru, Karnataka, India

Financial Services · 201-500 employees

2 h ago
sre Senior (5-10 yrs) Full-time India
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Senior Site Reliability Engineer will design resilient systems, define Service Level Objectives, and lead incident response efforts. They will also collaborate with engineering teams to troubleshoot production issues and maintain infrastructure monitoring and automation.

What they look for

Site Reliability Engineering Incident Management System Architecture Cloud Services Observability Datadog Dynatrace Java Microservices Kafka PostgreSQL Snowflake ETL Agentic Development Troubleshooting Automation

Requirements

Candidates must have strong technical knowledge in frontend, backend, or data engineering technologies and experience with large cloud services. Proven experience in incident management, observability tools, and excellent communication skills are required.

Full description

OVERVIEW OF 73 STRINGS:

73 Strings is an innovative platform providing comprehensive data extraction, monitoring, and valuation solutions for the private capital industry. The company's AI-powered platform streamlines middle-office processes for alternative investments, enabling seamless data structuring and standardization, monitoring, and fair value estimation at the click of a button. 73 Strings serves clients globally across various strategies, including Private Equity, Growth Equity, Venture Capital, Infrastructure and Private Credit.

Our 2025 $55M Series B, the largest in the industry, was led by Goldman Sachs, with participation from Golub Capital and Hamilton Lane, with continued support from Blackstone, Fidelity International Strategic Ventures and Broadhaven Ventures.

Role Summary

A Senior Site Reliability Engineer is responsible for ensuring the reliability, scalability, and operational excellence of our production systems. You operate at the intersection of software engineering and infrastructure. You have a senior technical role with broad impact across product engineering teams.

Responsibilities:

  • Drive the design of resilient systems and develop processes that prevent incidents.
  • Define Service Level Objectives (SLOs) and influence system architecture, system design and development to meet these SLOs.
  • Lead incident response and raise the reliability bar in your feature area.
  • Actively participate in incident management and work towards resolution with engineering.
  • Troubleshoot and resolve complex production issues, collaborating with engineering teams.
  • Engage with customers to provide technical support and address their concerns.

As a Senior Individual Contributor, you will own and lead for the whole team one specific process (e.g. incident management response, automatic alerting) and contribute to some other processes:

  • Participate in application deployments, upgrades, and infrastructure monitoring.
  • Help maintain code quality, organization, and automation.
  • Document support procedures, runbooks, and incident reports.
  • Contribute to documentation and knowledge sharing with both internal teams and customers

Requirements:

  • Strong technical knowledge in at least one of these technologies:
  • Front end technologies (e.g., HTML, CSS, JavaScript, Angular)
  • Backend technologies (e.g. Java, microservices, Kafka)
  • Data engineering (e.g. PostgreSQL, Snowflake, ETL)
  • Experience with large cloud services and application infrastructure.
  • Experience working with observability tools like datadog, dynatrace, or similar.
  • Strong analytical and troubleshooting skills in a SaaS environment.
  • Proven experience in incident management and resolution.
  • Proficient in agentic development and operation
  • Excellent communication and collaboration skills with technical and non-technical stakeholders.

Similar roles