Butterfly Network

Staff , Site Reliability Engineer - Cloud Platform

Butterfly Network · New York, New York, United States · $190K–$210K/yr

Medical Equipment Manufacturing · 51-200 employees

3 d ago
Senior (5-10 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Staff Site Reliability Engineer will own end-to-end observability and drive production reliability through SLOs and incident response standards. They will also build automation for Kubernetes/EKS workloads and mentor engineering teams on operational excellence.

What they look for

Site reliability engineering Cloud platform Observability Kubernetes AWS Incident response Metrics Logging Distributed tracing SLOs Automation Scripting Programming Operational excellence Mentoring

Requirements

Candidates must have 8+ years of experience managing production systems with deep expertise in AWS and Kubernetes. Strong programming skills and a proven track record in observability and incident management are required.

Benefits

Health insurance Dental coverage Vision coverage Health savings account Employee assistance program 401k plan Employee stock purchase plan Unlimited paid time off Parental leave Equity

Full description

Company Description

Butterfly Network, Inc. (NYSE: BFLY) is driving a digital revolution in ultrasound imaging and sensing with its proprietary Ultrasound-on-Chip™ semiconductor technology and software solutions. Butterfly first proved its technology in the point-of-care ultrasound market – commercializing the world’s first single-probe, whole-body portable ultrasound device, which is now on its best-selling, third-generation: Butterfly iQ3™. The Company combines its advanced hardware with cloud software and AI, an enterprise workflow solution (Compass AI™) and other offerings to drive adoption of affordable, accessible ultrasound. Butterfly also enables third-party development of imaging AI apps through Butterfly Garden™, its software development kit and AI partnership initiative.

In addition to its medical imaging products, Butterfly Embedded™ is the Company’s Ultrasound-on-Chip™ licensing and co-development program designed to enable a new wave of ultrasound-enabled technologies across non-competitive healthcare markets and beyond. Through Butterfly Embedded™, partners can build and scale novel ultrasound applications powered by Butterfly’s proprietary semiconductor chip and software platform. Butterfly’s innovations have been recognized by Prix Galien USA, Fierce 50, TIME’s Best Inventions and Fast Company’s World Changing Ideas, among other achievements.

We’re a team of bold thinkers, problem-solvers, and innovators ready to shape the future of medical imaging. Let’s build something extraordinary together!

Job Description

We're looking for a Staff Site Reliability Engineer to raise the bar for how we observe and operate our production cloud platform. Our cloud platform powers clinical workflows on top of one of the largest ultrasound image repositories in the world. In this role you'll be a hands-on technical contributor who leads by execution.

What you’ll do

  • Own observability end to end. Evolve and streamline our metrics, logs, and distributed-tracing strategy. Building telemetry standards that give every team a clear, real-time picture of production health
  • Drive reliability through SLOs. Establish meaningful service-level objectives with product and engineering teams, and use error budgets to balance velocity against stability.
  • Lead incident response. Help set the standards for our incident response feedback loop. Manage the on-call rotations for production systems, respond to incidents with calm and rigor, manage blameless postmortems – then systematically reduce toil.
  • Engineer for reliability. Operate and improve our Kubernetes/EKS workloads on AWS. Build automation and tooling that removes manual operational work and improves developer productivity.
  • Be a force multiplier. Mentor engineers on operational excellence, set standards for how services are built to be observable and reliable, and influence architecture decisions across teams.

Qualifications

Baseline skills/experiences/attributes:

  • 8+ years of managing production systems
  • Deep, hands-on production experience with AWS and Kubernetes
  • Strong programming/scripting skills and experience building operational automation
  • Proven ownership of a full stack observability platform (NewRelic/Datadog/etc)
  • A track record of leading incident response and delivering improvements in reliability
  • Clear written and verbal communication skills
  • Strong desire for being a mentor to others

Ideally, you also have these skills/experiences/attributes (but it’s ok if you don’t!):

  • Background operating in a highly regulated or compliance-sensitive environment (e.g., HIPAA, SOC 2, FedRAMP, SaMD, etc).
  • Background in technical Healthcare Protocols (DICOM, HL7, FHIR) and systems integrations (PACS, VNA, EMR, etc)

Values

Innovation is what we do. Our values are how we make it happen. Butterflies are and believe in…

  • Patient-Centric Innovators: Our mission is THE mission.
  • Empowered to Impact: Every voice matters.
  • One Team, One Goal: Unity fuels progress.
  • Growth Champions: We embrace challenges.
  • Action-Oriented Achievers: We follow through, every time.

Location

Butterfly offers a hybrid work model for most positions, with team members spending two or more days a week in the office. While flexibility is key, we value in-person connections that spark creativity and teamwork. Our offices are designed for collaboration, with comfortable workspaces, stocked kitchens, and opportunities to connect with peers.

This is a hybrid position (2 - 3 days a week in office) and will be based out of our office in New York City, NY.

Benefits and Perks

  • Comprehensive health insurance, encompassing dental and vision coverage, is provided to all our employees. As a health-tech company, we prioritize the well-being of our teams. We also contribute to Health Savings Account (HSA) accounts for all enrolled employees on an annual basis.
  • Comprehensive Employee Assistance Program - we provide access to tools and resources to support your emotional health and day-to-day needs.
  • 401k plan and match - we facilitate your retirement goals.
  • Eligible employees will have the opportunity to participate in Employee Stock Purchase Plan (ESPP)
  • Unlimited Paid Time Off + 10 Holiday Days a Year - recharge and come back ready to make an impact
  • Parental Leave - we aim to provide our employees with time to bond with their growing family, along with additional support for primary caregivers to help transition back to work
  • Competitive salaried compensation - we value our employees and show it
  • Equity - we want every employee to be a stakeholder
  • The opportunity to build a revolutionary healthcare product and save millions of lives!

Compensation

Our estimated salary for this role based in NYC is around $190,000 - $210,000 + bonus + equity + benefits. Actual pay is determined by multiple factors such as skills, qualifications, experience and market demand.

Candidates must be authorized to work in the United States. Butterfly Network will not provide immigration sponsorship of any kind for this position, including but not limited to H-1B, TN, O-1, L-1, or green card sponsorship (this includes applicants that are in the process of obtaining a greencard), now or in the future.

Butterfly Network does not accept agency resumes.

Butterfly Network is an E-Verify Company.

Butterfly Network is an equal opportunity employer. Regardless of race, traits associated with race, color, ancestry, religion, gender, national origin, sexual orientation, age, citizenship, marital status, disability or Veteran status. All your information will be kept confidential according to EEO guidelines.

Butterfly requires security adherence responsibilities from all employees. These include: adhering to all company security policies and procedures, utilize provided company assets securely, and complete all required security awareness training programs. Safeguarding company data and systems from unauthorized access, modification, or destruction, contributing to the overall security posture of the organization. Immediately reporting any suspected or actual security incidents, including phishing attempts, malware infections, or unauthorized access, following the established incident response procedur

#LI-KG

#KG-LI