About the role
The Site Reliability Engineer will monitor and maintain production systems 24/7 to ensure high availability and performance. They will also collaborate with R&D and support teams to troubleshoot complex infrastructure issues and document technical solutions.
What they look for
Requirements
Candidates must have experience with Linux systems, scripting languages like Python or Bash, and possess strong troubleshooting skills. An AI-driven mindset and the ability to work rotating shifts are also required for this role.
Benefits
Full description
Moon Active is a company driven by the mission to become a global leader in mobile gaming. Founded in 2011, our passion for creativity, cutting-edge technology, and delivering exceptional player experiences has resulted in games enjoyed by millions worldwide.
We're looking for a Site Reliability Engineer to join our talented community of professionals in our Warsaw office and contribute to the ongoing success of our leading mobile gaming products.
The Site Reliability department focuses on ensuring the continuous smooth operation of live production systems that serve millions of players globally. You will support production activities by troubleshooting, developing, maintaining, and documenting technical solutions related to Moon Active’s production infrastructure.
Responsibilities
- Ongoing monitoring and control of the availability of the different services of the production 24/7;
- Monitoring and detecting problems in the production environment, as well as 1st and 2nd tier infrastructure and application troubleshooting;
- Working within Operations with interfaces to Product, R&D and Support teams to escalate, troubleshoot and resolve complex issues;
- Ensuring proper documentation is provided for all supported SRE activities and standards.
Requirements
- AI driven mindset and experience in LLM tools usage;
- Experience with Linux based systems;
- Scripting experience (Python/Bash/etc);
- Basic monitoring experience;
- Good communications skills;
- Troubleshooting and problem-solving skills;
- A team player with the flexibility to work rotating shifts between 7:00 and 21:00
- Proficiency in English.
Advantages
- Be familiar with AWS/GCP (or other cloud providers);
- Be familiar with AI tools (Cursor, Claude Code, MCP etc.);
- Experience with monitoring tool and vendors such as Prometheus, Grafana, ELK, NewRelic, Signalfx, CloudWatch, DataDog, etc;
- Experience with SQL;
- Experience with IaC (Terraform);
- Experience with distributed systems, containers and Kubernetes.
We offer:
- Generous compensation with regular performance reviews;
- Paid vacation and sick leaves;
- Comprehensive medical insurance for you and your family member free of charge;
- Sports expenses reimbursement;
- Comfortable office in BC Gulliver;
- Daily lunches in the office and fully stocked kitchen with the greatest coffee;
- Newest technical equipment (macOS);
- Training & Development / Tuition reimbursement; online courses of your choice;
- Parental leave;
- Employee Referral Program with great bonuses;
- Regular team buildings and Company Happy Hours;
- Relocation bonus for nonlocal candidates;
- Reimbursement of car parking.
Similar roles
-
Senior Software Engineer - SRE & AIOps
ServiceNow Santa Clara, California, United States · $143K–$243K/yr
-
Senior Staff Software Engineer – SRE & AIOps
ServiceNow Santa Clara, California, United States · $191K–$334K/yr
-
Site Reliability Engineer Manager I
Stone - Linkedin Hamburg, Germany
-
DevOps / Site Reliability Engineer – Engineering
Practice By Numbers Kolkata, West Bengal, India
-
ASKUSR0145930 Site Reliability Engineer (SRE)
Essnova Solutions, Inc. Berkeley, California, United States · $166K/yr
-
Head of Site Reliability Engineering (SRE)
Computershare Bristol, England, United Kingdom