Aspire

Senior Software Engineer / SRE (Observability Focus)

Aspire

Staffing and Recruiting · 11-50 employees

Aug 02
Remote sre Senior (5-10 yrs) Full-time
Log in to apply, save this posting, or score it against your profile with AI.

About the role

You will support platform reliability, monitoring, and modernization initiatives by building and maintaining observability solutions. This involves configuring dashboards and alerts while driving scalability and performance improvements across Kubernetes-based environments.

What they look for

Python JavaScript Java Kubernetes DataDog AWS CI/CD Observability Microservices API integration Prometheus Grafana Go Automation Site Reliability Engineering

Requirements

Candidates must have strong proficiency in Python, JavaScript, or Java and hands-on experience with Kubernetes and AWS. You should also possess deep expertise in observability tools like DataDog and the ability to automate operational tasks.

Benefits

Competitive long-term total compensation package Performance-based bonus Remote-first work environment Technical and non-technical training programs Global exposure International technology conferences

Full description

This is a remote position.

About the Role

As a Senior SRE (Observability) at Aspire, you will support platform reliability, monitoring, and modernization initiatives across internal systems. This role blends software engineering (60–70%) with site reliability engineering (30–40%), with a strong emphasis on Kubernetes and observability platforms. You will build and maintain observability solutions, configure dashboards and alerts, and drive reliability, scalability, and performance improvements. This position is open to candidates in the APAC region with a preferred time zone overlap with JST.

What You'll Do

  • Support platform reliability, monitoring, and continuous improvement across internal systems.
  • Work in Kubernetes-based environments, deploying, operating, and monitoring containerized applications.
  • Build and maintain observability solutions, with a focus on DataDog (APM, metrics, logging, and tracing).
  • Configure dashboards, alerts, and monitoring for microservices-based architectures.
  • Integrate observability tools into AWS environments and CI/CD pipelines.
  • Automate monitoring and operational tasks using scripting (Python preferred).
  • Install and configure DataDog agents and integrations, managing API keys and secure configurations.
  • Manage user roles and access controls within observability platforms.
  • Lead maintenance efforts and platform improvements while proactively driving reliability, scalability, and performance.

What You'll Need

  • Strong proficiency in Python, JavaScript (Node.js), or Java.
  • Hands-on experience with API integrations (designing, consuming, and integrating).
  • Strong experience working in Kubernetes environments (deployment, operations, and monitoring).
  • Experience with DataDog (preferred) or similar tools (Prometheus, Grafana).
  • Ability to configure dashboards, alerts, and APM (tracing, metrics, logging).
  • Experience monitoring containerized and microservices architectures.
  • Hands-on experience with AWS.
  • Experience integrating observability tools into cloud environments.
  • Experience integrating observability into CI/CD pipelines.
  • Ability to automate monitoring and operational tasks using scripting (Python preferred).
  • Experience owning and operating an internal engineering platform.
  • Demonstrated ownership of reliability, scalability, and performance.
  • Proven ability to proactively lead maintenance efforts and platform improvements (not just reactive support).
  • Ability to work with JST time zone overlap.
  • Familiarity with Go (Golang).
  • Experience with additional observability tools such as New Relic, Dynatrace, Elastic, or Splunk Observability.

Why Aspire

In addition to a competitive long-term total compensation package with salary and performance-based bonus, we have a reward philosophy that goes beyond compensation.

  • Be part of a remote-first organization where flexibility is embraced.
  • Work and learn alongside talented engineers and technology leaders.
  • Explore opportunities to learn and grow through technical and non-technical training programs.
  • Gain global exposure by working on products with international teams and clients.
  • Attend virtual and in-person international technology conferences to expand your knowledge and network.

Similar roles