eMed, LLC

Senior Software Engineer (SRE)

eMed, LLC Miami, Florida, United States

Pharmaceutical Manufacturing · 11-50 employees

Aug 21
sre Senior (5-10 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will lead reliability engineering efforts, design robust monitoring systems, and manage production Kubernetes and AWS environments. Additionally, you will collaborate with teams to improve service scalability, fault tolerance, and operational readiness through automation.

What they look for

Kubernetes AWS Terraform Infrastructure as Code Observability Incident Management Automation Cloud-native infrastructure EKS Networking IAM CloudWatch ELB VPC Scripting Reliability Engineering

Requirements

The role requires strong experience with Kubernetes, AWS, and Infrastructure as Code practices like Terraform. Candidates should possess deep knowledge of observability tools, incident management workflows, and the ability to build automation scripts.

Benefits

Health Care Plan Retirement Plan Life Insurance Paid Time Off Short Term Disability Long Term Disability Training & Development Catered Breakfast Catered Lunch Wellness Resources

Full description

Senior Software Engineer (SRE) Location: Onsite in Miami ·Reports to: Engineering Lead· Department: Engineering 

About eMed eMed is a leading digital health company specializing in cardio metabolic health through managed GLP-1 programs. Focused on delivering for employers and benefits administrators seeking to control skyrocketing healthcare costs driven by obesity-related conditions.   Our integrated approach — combining telehealth, medication management, coaching, and outcomes tracking — makes eMed a differentiated and proven solution. Delivering structured programs that make GLP-1 therapy safe, sustainable, and cost-effective. eMed solution empowers both the employer and employee to take control of their healthcare costs and their health.

About Role As a Senior Software Engineer in SRE at eMed, you will play a key role in ensuring our platform is highly available, secure, and performant. You’ll lead reliability engineering efforts across production systems, drive operational excellence, and collaborate closely with application and infrastructure teams to design resilient services. This role suits an engineer with a software mindset and deep operational experience, who thrives on improving systems through automation and proactive engineering.

WHAT YOU’LL DO

  • Design and implement robust monitoring, alerting, and observability systems across all services and infrastructure

 

  • Lead reliability reviews, incident response, and post-incident analysis—focusing on prevention, learning, and long-term improvements

 

  • Improve service scalability, fault tolerance, and performance through architectural input and systems optimisation

 

  • Build and maintain automation for infrastructure management using Terraform, and delivery pipelines using GitHub Actions

 

  • Partner with software engineers to improve the operational readiness and resilience of services, including capacity planning and runbooks

 

  • Lead initiatives to reduce operational toil through tooling, automation, and process improvement

 

  • Manage and optimise our production Kubernetes and AWS environments with a focus on reliability, security, and cost-effectiveness

 

  • Contribute to security hardening efforts, including network controls, secrets management, and compliance readiness

 

  • Participate in and lead in-person stand-ups, incident reviews, and cross-team planning sessions

 

  • Share knowledge and mentor engineers on best practices in observability, incident response, and operational engineering

QUALIFICATIONS

Technical Skills (Essential)• Strong experience operating Kubernetes and cloud-native infrastructure (preferably EKS on AWS) in production environments

  • Proficiency in AWS services, including networking, compute, IAM, and logging/monitoring tools (e.g. CloudWatch, ELB, VPC)
  • Skilled in Terraform and Infrastructure as Code practices
  • Deep understanding of observability tooling (metrics, logs, tracing) and incident management workflows
  • Strong coding skills for building tools, scripts, and automation
  • Ability to troubleshoot complex infrastructure issues and lead delivery of reliable cloud solutions

Preferred• Experience implementing SLAs, SLOs, and error budgets to guide operational priorities

  • Background in healthcare or other regulated industries with security and compliance requirements
  • Previous involvement in platform security reviews

Benefits• Health Care Plan (Medical, Dental & Vision)

  • Retirement Plan (401k with Company Match)
  • Life Insurance (Basic, Voluntary & AD&D)
  • Paid Time Off
  • Short Term & Long-Term Disability
  • Training & Development
  • Catered Breakfast and Lunch 5 days a Week
  • Wellness Resources

Similar roles