AI First DevOps Engineer
Tomorrow.io Tel Aviv, Tel-Aviv District, Israel
Software Development · 201-500 employees
About the role
You will evolve and maintain cloud infrastructure and self-service platforms to improve reliability, security, and efficiency. Additionally, you will partner with engineering teams to integrate AI tools and ensure production services meet their SLOs.
What they look for
Requirements
Candidates must have 4+ years of experience in DevOps or SRE, including 2+ years of hands-on cloud experience and production Kubernetes expertise. Proficiency in IaC tools like Terraform and experience with modern observability systems are also required.
Benefits
Full description
Tomorrow.io's Engineering department is focused on building life-changing software and products at scale, from infrastructure that handles massive amounts of data to outstanding customer-centric user experiences in B2B, B2C and B2D products that change billions of lives worldwide.
We're looking for a DevOps Engineer to power the reliability, security, and efficiency of the world's most impactful weather platform. You'll build self-service platforms that give developers and weather scientists true independence, weave AI into how we operate, and work side-by-side with R&D to push performance and scale further. You'll evolve our cloud infrastructure to match the pace of the business, hold the line on cost, and stay close to production through on-call. The people who thrive here bring a product mindset, take ownership without waiting to be asked, and leave the people and systems around them better than they found them.
As a DevOps Engineer at Tomorrow.io, you'll:
Evolve and maintain our cloud infrastructure, delivery and observability - and the self-service platforms above them - so they serve our business strategy and let scientists and developers work on their own as we grow at scaleDevelop and adopt tools, with AI at the center, that make development and operations processes measurably more efficientPartner with the engineering teams behind the Tomorrow.io platform and its API to improve service performance, reliability, scale, security and costKeep production available to its SLOs - taking your turn on the on-call rotation and following each incident through to whatever stops it from happening againWork as part of the team: share what you know, learn from the people around you, and help us keep improving how we work together
What you bring:
- 4+ years as a DevOps / Site Reliability Engineer, including 2+ years hands-on with a major cloud (GCP, AWS or Azure) and with IaC such as Terraform (Crossplane or another Kubernetes-native approach is a plus)
- Production Kubernetes depth - containerized deployments, scheduling, resource management and real troubleshooting - plus CI/CD pipelines you've built and owned end to end with modern tooling and advanced deployment methodologies
- Experience implementing and customizing monitoring and observability systems (Datadog, Grafana + Loki, Prometheus, ELK)
- Software engineering craft and the judgment to review, debug and improve what an agent produces
- Hands-on use of AI tooling in your engineering work - shaping and extending it (agents, rules, integrations), not only consuming it
- A product mindset and a deep sense of responsibility for service reliability: you build internal tools from developer feedback, measure success with clear metrics and KPIs, and treat production as yours to care forCuriosity, adaptability and a bias for action in a high-velocity, changing environment - with communication that connects infrastructure work to business impact and brings R&D stakeholders together around shared goals
Our stack: GCP (primary), Azure and AWS · Linux · Kubernetes on GKE and AKS with Helm, KEDA, Argo Workflows and External Secrets · Cloud Run and Cloud Functions · Terraform and Crossplane · GitHub Actions and ArgoCD (GitOps) · Datadog, Grafana + Loki · PostgreSQL/PostGIS, MongoDB Atlas, Redis · Pub/Sub · GCS and S3 · Cloudflare, Kong Gateway, Google LB, CloudFront · Python and Go, with Node.js + React on the frontend.
So if you're passionate about optimizing development processes and curious about how systems work behind the scenes, if continuous improvement is a way of life and automation is your tool of choice, this is the place for you. You'll help teams deliver faster and more safely, working with some of the sharpest minds in the industry on complex, high-scale problems, as we build the world's most trusted weather platform for millions of people every day. You'll grow quickly here, in both technical depth and responsibility, in a culture where engineers truly care for their production workloads end to end.
If your experience is close but doesn't fulfill all requirements, please apply. Tomorrow.io is on a mission to build a special company. To achieve our goal, we are focused on hiring people with different backgrounds, perspectives, and experiences.
___________________________________________________________________________________________
About tomorrow.io:
Selected by TIME Magazine as one of the Top 100 Most Influential Companies in the World, Tomorrow.io is the world's leading Resilience Platform™. Combining next-generation space technology, advanced generative AI, and proprietary weather modeling, Tomorrow.io delivers unmatched forecasting and decision-making capabilities. Trusted by six of the top ten Fortune 500 companies, Tomorrow.io empowers organizations to proactively manage weather-related risks, opportunities, and enhance operational efficiency. From cutting-edge weather intelligence to real-time early warning systems, Tomorrow.io enables predictive, impact-based action for a safer, more resilient future. Learn more at Tomorrow.io.
Ethos: Our ethos guides us in everything we do - The people of Tomorrow are here to make an impact, they show true grit, and always put people first.
How we roll: We work in an “one office” environment. We believe that magic happens when people work together. Together also includes Zoom meetings, flexible hours and unlimited vacation days. Your success is achieved by your impact and deliveries and not by the hours you put in. We believe in transparency and directness, putting work before ego and empathy. We grow fast and move faster but we always see people first. Each person has their own career growth path for we believe that the only way for the company to grow is if you grow.
Similar roles
-
Cloud DevOps Architect
CleanSlate Technology Group Carmel, Indiana, United States
-
DevOps Engineer
Swift Herndon, Virginia, United States
-
Architect / Technical Team Lead - DevOps & DevSecOps
BETSOL Bengaluru, Karnataka, India
-
AWS DevOps Engineer - Lead
Informed Solutions Altrincham, England, United Kingdom
-
Senior DevOps Engineer - OP02235
Dev.Pro Sofia, Sofia-City, Bulgaria
-
Staff DevOps Engineer
Ripple London, England, United Kingdom