StackAI

Senior DevOps Engineer

StackAI · San Francisco, California, United States · $130K–$210K/yr

Software Development · 51-200 employees

Yesterday
Remote Senior (5-10 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Senior DevOps Engineer will design, build, and maintain scalable cloud infrastructure while managing Kubernetes operations. They will also streamline CI/CD pipelines and collaborate with engineering teams to ensure secure and reliable deployments.

What they look for

Kubernetes Terraform Docker Infrastructure as Code AWS Azure GCP CI/CD GitHub Actions Python Bash Grafana Prometheus Datadog GitOps FluxCD

Requirements

Candidates must have 5+ years of experience in DevOps or infrastructure engineering with strong expertise in Kubernetes, Terraform, and Docker. Proficiency in cloud platforms like AWS, Azure, or GCP and experience with CI/CD pipelines and scripting languages are required.

Benefits

Equity

Full description

About Stack AI

Stack AI is building cutting-edge infrastructure at the intersection of AI and enterprise. Our mission is to enable companies to safely and reliably deploy AI at scale. As we grow, we’re looking for a Senior DevOps Engineer who can design, build, and manage the systems that power our platform.

This is a hands-on role where you’ll shape our infrastructure, scale our cloud environments, and partner closely with engineering to ensure our deployments are seamless, secure, and reliable.

What You’ll Do

  • Own cloud infrastructure: Design, implement, and maintain our infrastructure with a strong focus on scalability, security, and cost-efficiency.
  • Lead Kubernetes operations: Manage and optimize containerized applications in Kubernetes (deployment, scaling, monitoring, troubleshooting).
  • Build automation: Drive Infrastructure as Code to standardize, automate, and improve reliability of environments.
  • Streamline CI/CD: Architect and maintain pipelines to accelerate developer productivity and ensure zero-downtime deployments.
  • Enable observability & security: Implement monitoring, logging, and alerting and drive compliance best practices.
  • Collaborate across teams: Work closely with engineers to improve architecture, accelerate deployments, and contribute to strategic initiatives.
  • Experiment & improve: Explore new tools, optimize workflows, and continuously refine our DevOps practices.

What We’re Looking For

  • 5+ years of experience in DevOps, Cloud, or Infrastructure engineering.
  • Strong expertise with Kubernetes, Terraform, Docker, and Infrastructure as Code (IaC).
  • Hands-on experience with AWS, Azure, and/or GCP.
  • Solid background in CI/CD pipelines (GitHub Actions, Azure DevOps, GitLab CI, etc.).
  • Comfort with scripting/programming (Python, Bash, or similar).
  • Experience with monitoring & observability (Grafana, Prometheus, Datadog, etc.).
  • Strong problem-solving and communication skills, with the ability to thrive in a fast-paced startup environment.

Nice to have:

  • Experience with FluxCD or similar GitOps tools.
  • Familiarity with security/compliance frameworks (SOC 2, HIPAA).
  • Previous startup experience.

Why Stack AI?

  • Work at the cutting edge of AI + infrastructure.
  • Shape the foundation of our cloud and DevOps practices.
  • Join a small, fast-moving team where your work has an immediate impact.
  • Hybrid flexibility: collaborate in-person in our San Francisco office (next to Salesforce Park) while enjoying remote days or be fully remote