Site Reliability Engineer
Leidos Honolulu, Hawaii, United States · $108K–$195K/yr
Defense and Space Manufacturing · 10,001+ employees
About the role
The Site Reliability Engineer will monitor and maintain system health while automating repetitive tasks to improve efficiency and reliability. They will also collaborate with development teams to ensure smooth deployments and perform essential data center tasks including hardware upgrades and security scans.
What they look for
Requirements
Candidates must possess a TS/SCI clearance and a Bachelor's degree with 8-12 years of relevant experience, or a Master's with 6-10 years. Proficiency in cloud platforms, scripting languages, and configuration management tools is required, along with DOD 8140 IAT II certification.
Full description
Site Reliability Engineer
Location: Hickam Air Force Base, Hawaii
Clearance: TS/SCI
Leidos has an opening for a highly qualified TS/SCI cleared Site Reliability Engineer at Hickam Air Force Base, Hawaii for the Decision Advantage Business Area in Defense sector. This is an exciting opportunity to bring your experience to support all-domain large-scale weapon systems, Information Technology Systems, and Command and Control Systems to realize the Department of Defense Joint All-Domain Command and Control (JADC2). In this role you will support the Advanced Battle Management System (ABMS) Digital Infrastructure (DI) Processing Node (PN) team to design and implement solutions that can be delivered at speed, scale, and with the necessary security to deliver operational advantages to the joint warfighter. ABMS is a top modernization priority for the Department of the Air Force and will be the backbone of a network-centric approach to battle management in partnership with all the services across JADC2. This position will work closely with Program Managers, domain engineers, and Government counterparts across Government and Industry partners.
Primary responsibilities:
- Monitor and maintain system health.
- Automate repetitive tasks to improve efficiency and reliability.
- Build and maintain effective monitoring and alerting systems.
- Collaborate with development teams to ensure smooth deployment of new features.
- Troubleshoot and resolve incidents to minimize downtime.
- Conduct post-incident reviews and implement improvements.
- Optimize system performance and scalability.
- Document processes and procedures for future reference.
- Must be able to perform the following (but not limited to) data center tasks:
- Perform system back ups
- Perform software installs
- Perform ACAS security scans
- Perform hardware upgrades (servers, switches, cabling, end user devices, and peripherals)
- Troubleshoot networking comms (LAN/WAN)
Required Qualifications:
- TS/SCI Clearance required
- Requires BS degree and 8 – 12 years of prior relevant experience or master’s with 6 – 10 years of prior relevant experience. Additional years of experience will be considered in lieu of degree.
- Proven experience as a Site Reliability Engineer or similar role.
- Strong knowledge of cloud platforms (e.g., AWS, Azure, Google Cloud).
- Proficiency in scripting and coding (e.g., Python, Bash, Go).
- Experience with configuration management tools (e.g., Ansible, Terraform).
- Familiarity with containerization and orchestration (e.g., Rancher/Harvester).
- Excellent problem-solving and analytical skills.
- Strong communication and collaboration abilities.
- Serve as a liaison between site operators and consortium DI team members.
- Ability to work in an on-call rotation.
- Familiar with Cisco/Juniper and Linux CLI and Logs
- Knowledge of IdAM and VSAN/Cloud Storage
- Required Certifications:
- DOD 8140 IAT II (Intermediate)
- GIAC Security Essentials Certification;
- GICSP: Global Industrial Cyber Security Professional: SSCP: Systems Security Certified Practitioner
Additional Qualifications/Certifications
- Experience with CI/CD pipelines and tools (e.g., Jenkins, GitLab CI/CD).
- Knowledge of networking and security best practice.
- Familiarity with database management and optimization.
- Experience with HAIPE devices/KG encryptors and security associations
- Experience with OSI Layers 1 – 4 troubleshooting methodologies
DABAOPP1
If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.
Original Posting:
August 12, 2026
For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.
Pay Range:
Pay Range $107,900.00 - $195,050.00
The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.
Similar roles
-
Senior Site Reliability Engineer (Digital Infrastructure)
Egis Group Melbourne, Victoria, Australia
-
IT Site Reliability Engineer — API Management Platforms
Texas Instruments Dallas, Texas, United States
-
Staff Site Reliability Engineer, GovCloud
Medallia Mclean, Virginia, United States
-
Lead SRE & Support Engineer
Providence Hyderabad, Telangana, India
-
Consultant Specialist(SRE)
HSBC Global Services Limited Tianhe District, Guangdong, China
-
Site Reliability Engineer
IFS Tokyo, Tokyo, Japan