Senior Site Reliability Engineer - 2394503
UnitedHealth Group Eden Prairie, Minnesota, United States · $150K–$184K/yr
Hospitals and Health Care · 10,001+ employees
Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
The Senior Site Reliability Engineer will design, develop, and maintain infrastructure automation using Chef while managing high-volume clusters for RabbitMQ, InfluxDB, and MySQL. The role also involves developing telemetry pipelines, creating Grafana dashboards, and participating in an on-call rotation to ensure system reliability.
What they look for
Requirements
Candidates must hold a Bachelor’s degree in Electronics Engineering or a related field and possess at least 5 years of progressive experience in an engineering-related occupation. Required technical expertise includes automation tools like Chef, time-series platforms, high-level programming languages like Ruby, and container technologies.
Benefits
Full description
Senior Site Reliability Engineer - 2394503
EMPLOYER: Optum Services, Inc.
JOB TITLE: Senior Site Reliability Engineer
LOCATION: 1 Optum Circle, Eden Prairie, MN 55344 (Telecommuting available from anywhere in the U.S.)
DUTIES: Design, develop, and maintain Chef cookbooks to automate the installation, upgrade, and configuration of infrastructure and monitoring components, including Telegraf, InfluxDB, Dynatrace OneAgent, and RabbitMQ; administer and support multiple high volume RabbitMQ clusters, including underlying infrastructure configuration, capacity management, performance tuning, and availability; manage and maintain a large scale InfluxDB cluster, including infrastructure provisioning, performance optimization, data retention, and reliability; manage multiple MySQL database clusters supporting core services, including monitoring, maintenance, and operational stability; administer and maintain multiple Grafana instances, including configuration, upgrades, and integration with telemetry data sources; operate, configure, and support HashiCorp Nomad container platform, including its supporting services such as HAProxy, HashiCorp Consul, and HashiCorp Vault; develop and maintain Grafana dashboards and alerting configurations to monitor server, container, and application health and performance; create and maintain automated CI/CD and operational workflows using GitHub Actions to support system reliability and deployment processes; develop and maintain Rubybased applications and services supporting the telemetry (metric) transport pipeline; and participate in an on-call rotation to respond to system incidents, troubleshoot issues, and ensure service availability. Telecommuting is available from anywhere in the U.S.
REQUIREMENTS: Employer will accept a Bachelor’s degree in Electronics Engineering or related field and 5 years of progressive, post baccalaureate experience in the job offered or in a Engineer-related occupation.
Position requires 5 (five) years of experience in the following:
- Write and maintain Chef Infra cookbooks or other automation tools such as Ansible, Puppet, or Salt;
- Time-series platforms like Telegraf, InfluxDB Prometheus, Thanos, Mimir, TimescaleDB, Graphite, or OpenTSDB;
- Code using a high-level programming language like Ruby, Sinatra or Rails;
- Git or GitHub;
- Container technologies like Docker, Nomad, Kubernetes, or OpenShift;
- SRE concepts;
- Utilize Linux skills like Red Hat or CentOS; and
- Data visualization Grafana or Kibana.
RATE OF PAY: $149,510- $183,834 per year
Please apply via careers.uhg.com and search for job #2394503
Careers with Optum. Here's the idea. We built an entire organization around one giant objective; make health care work better for everyone. So when it comes to how we use the world's large accumulation of health-related information, or guide health and lifestyle choices or manage pharmacy benefits for millions, our first goal is to leap beyond the status quo and uncover new ways to serve. Optum, part of the UnitedHealth Group family of businesses, brings together some of the greatest minds and most advanced ideas on where health care has to go in order to reach its fullest potential. For you, that means working on high performance teams against sophisticated challenges that matter. Optum, incredible ideas in one incredible company and a singular opportunity to do your life's best work.(sm)
UnitedHealth Group offers a full range of comprehensive benefits, including medical, dental and vision, as well as matching 401k and an employee stock purchase plan.
Diversity creates a healthier atmosphere: UnitedHealth Group is an Equal Employment Opportunity/Affirmative Action employer and all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, age, national origin, protected veteran status, disability status, sexual orientation, gender identity or expression, marital status, genetic information, or any other characteristic protected by law.
UnitedHealth Group is a drug-free workplace. Candidates are required to pass a drug test before beginning employment.
Similar roles
-
Senior Site Reliability Engineer (SRE)
LeoLabs, Inc. $171K–$192K/yr
-
Senior Site Reliability Engineer
2K Austin, Texas, United States
-
Lead Site Reliability Engineer
Sherwin-Williams Cleveland, Ohio, United States
-
Senior Site Reliability Engineer (SRE)
Tradeweb United States · $170K–$210K/yr
-
Senior Software Engineer, Site Reliability Engineering
Google New York, New York, United States · $174K–$252K/yr
-
Software Developer III, Site Reliability
Google Waterloo, Ontario, Canada · CA$150K–CA$153K/yr