Senior/Staff Software Engineer – Site Reliability & Infrastructure
General Intuition & Medal New York, New York, United States · $180K–$275K/yr
Computer Games · 51-200 employees
About the role
You will own the on-call rotation, drive incident response, and manage infrastructure scaling to ensure high reliability. You will work directly with engineering teams to address infrastructure needs and conduct postmortems to prevent future issues.
What they look for
Requirements
Candidates must have strong fluency in Terraform and hands-on experience scaling relational databases and Elasticsearch in production environments. You should possess deep knowledge of GCP, Kubernetes, and a proven track record of managing infrastructure at scale in startup environments.
Benefits
Full description
About The Company
General Intuition is the frontier lab for acting in space and time. We build large action models and world models that can perceive, predict, and act across virtual and physical environments. General Intuition builds on the strength of Medal, the world's largest and fastest-growing platform for gaming clips, where millions of gamers capture, share, and discover new games every year. We've raised over $650M from Khosla, GC, Valor, and Point72 since October 2025, and recently closed our latest round at a $6.2B valuation.
The Role
Medal's infrastructure handles billions of clips, video ingestion pipelines, and social features at a massive scale most engineers never get to touch. The work centers on reliability, incident response, scaling, and making sure our infrastructure keeps up with our growth. You'll own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infra needs. The right person probably came through startups and scale-ups and has been in the room when things broke at 2am, has scaled databases under pressure.
What We're Looking For
- Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scale
- Elastic search depth: Hands-on experience running ES for user-facing features, not just as a log sink
- GCP depth: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystem
- Database scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in production
- Incident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrence
- CI/CD: You've worked with GitHub Actions in a production environment
- Communication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortems
- Experience at startups: You are comfortable in an environment of rapid growth where scaling up is a priority
- Great judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week
Our Stack
Electron, React, Redux, Styled Components & other modern web-based technologies C# and C++ for native Windows recording & more Swift for iOS, Kotlin for Android Java, Redis, RabbitMQ, Kubernetes for backend Terraform, Salt, GitHub Actions, CircleCI for IaC and CI/CD
Similar roles
-
Site Reliability Engineer III (DBA)
Backblaze External Website United States · $125K–$150K/yr
-
Site Reliability Engineer II (AI Platform)
OpenTable Toronto, Ontario, Canada · CA$110K–CA$130K/yr
-
Site Reliability Engineer
Apple Hyderabad, Telangana, India
-
Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)
Encora Perímetro Urbano Santiago de Cali, Valle del Cauca, Colombia
-
Senior Site Reliability Engineer
Salesforce Dublin, Leinster, Ireland
-
Sr Staff Site Reliability Engineer, AI Infrastructure
d-Matrix Santa Clara, California, United States · $175K–$265K/yr