Associate Site Reliability Engineer/Site Reliability Engineer
C3 AI Redwood City, California, United States · $90K–$160K/yr
Software Development · 1,001-5,000 employees
Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
The role involves maximizing system uptime and availability while establishing comprehensive monitoring and alerting for critical services. You will also lead automation efforts to streamline system updates, upgrades, and deployment cycles across the platform.
What they look for
Requirements
Candidates must have a Bachelor's or Master's degree in Computer Science or equivalent experience, along with expertise in Linux, Kubernetes, and cloud infrastructure. Proficiency in scripting languages like Python or Bash and experience with IaC tools such as Ansible or Terraform are required.
Benefits
Full description
C3 AI (NYSE: AI), is the Enterprise AI application software company. C3 AI delivers a family of fully integrated products including the C3 Agentic AI Platform, an end-to-end platform for developing, deploying, and operating enterprise AI applications, C3 AI applications, a portfolio of industry-specific SaaS enterprise AI applications that enable the digital transformation of organizations globally, and C3 Generative AI, a suite of domain-specific generative AI offerings for the enterprise. Learn more at: C3 AI
We are looking for Associate Site Reliability Engineer/Site Reliability Engineer to join our team at our HQ in Redwood City, CA.
Responsibilities:
- Maximize system uptime and availability, ensuring functional and performance SLAs.
- Establish end-to-end monitoring and alerting on all critical aspects.
- Solve complex problems for critical services and build automation to prevent problem recurrence.
- Influence and create new designs, architectures, standards, and methods for supporting the platform.
- Initiate and lead scripting and automation to streamline system updates and upgrades.
- Set up critical infrastructure, tools, and framework to streamline the deployment cycle.
- Work cross-functionally with Services and Engineering teams.
Qualifications:
- BS or MS in Computer Science, related field, or equivalent professional experience.
- Demonstrated experience in deploying, managing, and operating scalable and fault-tolerant Linux/Kubernetes/JVM-based infrastructure in AWS, GCP, and other public clouds.
- Expertise in Linux Operating Systems, Networking, and Database concepts.
- Experience deploying, upgrading, and troubleshooting Kubernetes clusters and workloads.
- Experience with Cassandra (or another NoSQL alternative).
- Expertise in cloud providers, such as Amazon Web Services, Azure, and GCP.
- Experience with configuration management systems such as Puppet.
- Experience in Bash or Python; to automate and monitor systems.
- Experience with IaC tools like Ansible or Terraform.
- Excellent problem-solving, critical thinking, and communication skills.
- Experience supporting as a DevOps or sys admin for commercial SaaS solutions.
C3 AI provides excellent benefits, a competitive compensation package and generous equity plan.
California Base Pay Range
$90,000—$160,000 USD
C3 AI is proud to be an Equal Opportunity and Affirmative Action Employer. We do not discriminate on the basis of any legally protected characteristics, including disabled and veteran status.
Similar roles
-
Senior Site Reliability Engineer
UnitedHealth Group Schaumburg, Illinois, United States · $92K–$164K/yr
-
IC3 - Infra Engineer - SRE
Spin - Job Board External Ciudad de México, Mexico
-
Site Reliability Engineer II
Onapsis Dallas, Texas, United States
-
Vice President - Senior Manager of Site Reliability Engineering
JPMorgan Chase & Co. Jersey City, New Jersey, United States · $176K–$260K/yr
-
Sr Engineer, Site Reliability Engineer
Panera Bread Saint Louis, Louisiana, United States · $112K–$134K/yr
-
Director, Site Reliability Engineering - Paze
Early Warning® Scottsdale, Arizona, United States · $173K–$276K/yr