Lead Site Reliability Engineer (Job Code: JC14)
Qualys Raleigh, North Carolina, United States
Computer and Network Security · 1,001-5,000 employees
About the role
The Lead Site Reliability Engineer will co-develop and manage the lifecycle of Qualys Container Security within the cloud platform. They are responsible for system performance, efficiency, change management, monitoring, and emergency response.
What they look for
Requirements
Candidates must hold a bachelor's degree in Computer Science, Information Technology, or a related field. A minimum of 2 years of experience in the job offered or a related role is required, along with specific technical expertise in Kubernetes, automation tools, and cloud platform reliability.
Benefits
Full description
Come work at a place where innovation and teamwork come together to support the most exciting missions in the world!
Lead Site Reliability Engineer (Job Code: JC14) Raleigh, NC, United States
Responsibilities:
Co-develop and participate in the full lifecycle development of Qualys Container Security (CS) within the Qualys Cloud Platform from inception and design to deployment, operation and improvement. Responsible for the performance, efficiency, change management, monitoring, emergency response, and capacity planning for CS. Support Cloud Platform team before the technologies are pushed for production release through activities such as system design, capacity planning, automation of key deployments, and engage in building a strategy for production monitoring and alerting. Test and verify platform services for CS and integrated technologies before production release, identify performance bottlenecks and anomalous system behavior, and analyze root causes of incidents. Ensure that the cloud platform technologies are maintained properly by measuring and monitoring availability, latency, performance and system health. Participate in the development process by supporting new features, services, releases for the cloud platform technologies. Develop enhanced platform features and scale infrastructure. Develop tools and automate the process for achieving large-scale provisioning and deployment of cloud platform technologies. Work from home permitted 40% of time.
Requirements:
- Bachelor’s degree (or the foreign equivalent) in Computer Science, Information Technology, or related field.
- 2 years of experience in the job offered or related.
- Experience with:
- Qualys Container Security (CS);
- Kubernetes-based cloud platform reliability engineering;
- Using Prometheus/Grafana to detect availability & latency;
- Automating infrastructure provisioning using Ansible/Terraform; and
- Capacity planning & performance engineering for distributed production systems.
Salary, benefits and discretionary bonus.
Reference Job Code JC14 when applying.
Qualys is an Equal Opportunity Employer, please see our EEO policy.
Similar roles
-
Senior Staff Site Reliability Engineer
NVIDIA Bengaluru, Karnataka, India
-
Site Reliability Engineer I
Yum! Plano, Texas, United States · $96K–$120K/yr
-
SR SRE (Linux & Windows) - Latam
Pearster Buenos Aires, Argentina
-
Associate Site Reliability Specialist
Co-operators Calgary, Alberta, Canada · CA$50K–CA$84K/yr
-
Principal Site Reliability Engineer
Hewlett Packard Enterprise Aguadilla, Puerto Rico
-
SRE Platform Engineer
CACI International Washington, District of Columbia, United States · $115K–$252K/yr