S

MTS, Engineering Manager

SkyPilot San Mateo, California, United States

Technology, Information and Internet · 1 employee

8 h ago
engineering-manager Principal (10+ yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will lead and scale a high-performance engineering team while driving technical architecture and operational standards. Additionally, you will collaborate with founders to shape the product roadmap by synthesizing customer feedback and feature requests.

What they look for

Engineering management Distributed systems AI infrastructure Kubernetes Slurm GPU workloads Cloud platforms Product management Team leadership Recruiting Architecture design Reliability engineering Security standards Roadmap planning Technical strategy

Requirements

The ideal candidate has over 8 years of software engineering experience, including at least 3 years in a management role for teams of 10 or more. You must possess hands-on experience with distributed systems, cloud platforms, and AI infrastructure technologies like Kubernetes or Slurm.

Benefits

Competitive compensation Equity Medical coverage Dental coverage Vision coverage Gourmet lunch and dinner

Full description

About SkyPilot

SkyPilot accelerates the world's most ambitious AI teams. SkyPilot turns fragmented AI compute across clusters into one optimized, highly available and easy-to-use pool: a single AI supercomputer.

SkyPilot (10k+ GitHub stars, 18M+ downloads) manages the GPU fleets of 100s of companies, from Fortune 500s to top AI-natives like Abridge, Applied Compute, Mistral, Unconventional AI, H Company, and Nubank, with usage growing exponentially. Born in the UC Berkeley lab behind Spark and Databricks, our growing team includes top-tier talent from Databricks, Google, Berkeley, MIT, CMU, and Cornell. To date, SkyPilot has raised over $20M in seed funding from top investors (incl. Lux, Coatue, Amplify) and operators (incl. Ali Ghodsi, Jeff Dean, Guillermo Rauch, Amjad Masad, Clem Delangue, Aaron Levie).

The role

We are hiring our very first Engineering Manager to grow and scale our engineering team. You will report directly to our CTO and work closely with the founders.

We are looking for someone who is hands-on, hungry for impact, and takes pride in building high-performing teams. The ideal candidate understands the output of a manager is the sum of the outputs of all of their reports.

You'll lead the engineering team, stay involved in technical work, and shape product directions alongside the founders: reviewing customer call recordings and notes, aggregating requests across accounts, and deciding what the team builds next. You'll have direct exposure to our largest customers.

What you'll do

Build and lead a high-performance engineering team

  • Recruit, coach, and grow a team of distributed systems and AI infrastructure engineers.
  • Set the engineering culture and practices needed as the team scales, driving high performance and ownership.
  • Give fast, direct feedback.

Drive technical directions

  • Participate in design reviews, code reviews, and architecture decisions.
  • Own reliability, security, and operational standards for a control plane that frontier AI teams depend on
  • Guide the team as the platform expands into more use cases.

Shape product directions for the platform

  • Work with the CTO and the founders to set priorities.
  • Review customer call recordings and notes, aggregate feature requests and bugs across customers, and turn them into a prioritized roadmap, weighing retention of existing customers and growth of new accounts.

Deliver

  • Translate the roadmap into milestones, sequence work across parallel projects, track execution, and serve as the escalation point for major incidents.

What we're looking for

  • 8+ years of software engineering experience with 3+ years managing engineers (teams of 10+), ideally at a fast-growing infrastructure or developer-tools company.
  • Experience building and operating large-scale distributed systems or cloud platforms in production, and still able to contribute technically.
  • Hands-on with Kubernetes or Slurm, and familiar with GPU workloads. Experience with schedulers, orchestration, or AI infrastructure is a plus.
  • Comfortable acting as the product manager for your area: synthesizing customer input, running an intake and triage process, and making prioritization calls with incomplete information.
  • Strong product intuition and care for both developer experience and enterprise needs, and able to make and communicate hard tradeoffs.
  • Experience as an early engineering leader at a startup, prior PM experience, or building enterprise platforms (multi-tenancy, RBAC, HA control planes) is a strong plus.

What we offer

  • Competitive compensation and equity
  • Comprehensive medical, dental, vision coverage for you and your dependents
  • The chance to work with some of the best minds in AI infrastructure and distributed systems, with significant autonomy and ownership.
  • A front-row seat at the latest open-source infra startup from Berkeley (lineage: Databricks).
  • Gourmet lunch & dinner for the team to do their best work

Location: San Mateo, CA.

Similar roles