Site Reliability Engineer
Provident Pénzügyi Zrt Budapest, Central Hungary, Hungary
Financial Services · 1,001-5,000 employees
About the role
The Site Reliability Engineer will build, operate, and maintain AWS-based infrastructure while driving automation and operational stability. They will also collaborate with cross-functional teams to monitor system health, manage incidents, and support platform modernization initiatives.
What they look for
Requirements
Candidates must have 3–6 years of experience in infrastructure or cloud engineering with hands-on expertise in AWS environments. Proficiency in Infrastructure as Code, Windows Server administration, and scripting languages like Python or PowerShell is required.
Benefits
Full description
We are looking for a
Site Reliability Engineer
in its Global IT Team in Hungary.
International Personal Finance (IPF), the parent company of Provident, is looking for a Site Reliability Engineer to join our Global IT team in Hungary.
This role focuses on the reliable operation and continuous improvement of business‑critical platforms used across multiple European markets. You will work extensively with AWS‑based environments, driving automation, monitoring and operational stability, while addressing real production challenges.
You will collaborate closely with developers, architects and infrastructure specialists, contributing to the operation and evolution of our IT systems. The position offers a strong technical foundation and a clear development path, particularly for professionals moving from traditional infrastructure roles toward cloud‑ and automation‑focused engineering, as part of our ongoing platform modernisation initiatives.
What you will do
· Build, operate, and maintain AWS‑based infrastructure for both production and non‑production environments
· Implement Infrastructure as Code using CloudFormation to ensure consistent and reliable deployments
· Support application teams by maintaining reliable, scalable, and well‑monitored systems
· Monitor system health and key service metrics such as availability, latency, and error rates
· Identify repetitive operational tasks and automate them to minimise manual effort
· Implement and improve monitoring, alerting, and logging solutions
· Take part in incident response and on‑call rotations to help resolve production issues
· Contribute to post‑incident reviews and introduce preventative improvements
· Collaborate closely with development, security, and infrastructure teams
· Support cloud cost optimisation through monitoring and rightsizing of resources
· Operate AWS environments in line with internal cloud standards
Required skills & technologies
· 3–6 years of experience in infrastructure, cloud, or platform engineering roles
· Hands‑on experience operating AWS‑based environments, including:
o EC2, S3, RDS
o IAM, VPC, and networking concepts
· Strong experience with Infrastructure as Code, preferably using CloudFormation
· Proven Windows Server administration skills in production environments
· Experience supporting and troubleshooting production systems, including incident handling
· Solid understanding of system reliability, monitoring, alerting, and operational best practices
· Experience with monitoring and logging tools, such as Amazon CloudWatch
· Scripting and automation experience using Python, PowerShell, or similar languages
· Experience working in Agile and/or DevOps environments
· Ability to collaborate effectively with developers, architects, and infrastructure teams
· Good spoken and written communication skills in English
Nice to have skills
· Experience working with AWS‑based cloud environments, including multi‑account setups
· Exposure to automation and CI/CD pipelines (e.g. Azure DevOps)
· Experience or familiarity with containerized or serverless technologies (Kubernetes/EKS, AWS Lambda)
· Basic understanding of SRE and reliability concepts (monitoring, SLIs/SLOs, error budgets)
· Knowledge of AWS security best practices and general enterprise IT processes (e.g. ITIL)
· Experience working with distributed systems or microservices
· AWS certification or equivalent hands‑on experience
· Experience working in a multinational or international environment
What we offer
· Opportunity to work with modern AWS-based platforms across multiple European markets,
· Meaningful work – Contribute to shaping a sustainable future and making a real difference,
· Professional development and training opportunities,
· Performance-based annual bonus,
· Cafeteria and private health insurance,
· Corporate health programme,
· Laptop and mobile phone,
· Hybrid work model (home office available),
· International working environment,
· Participation in infrastructure modernisation and cloud transformation initiatives.
Similar roles
-
Site Reliability Engineer, Enterprise Technology Services
Apple Sunnyvale, California, United States
-
Site Reliability Engineer (SRE) - Principal Consultant
Capco New York, New York, United States · $146K–$183K/yr
-
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes
Software Mind Montreal, Quebec, Canada
-
Site Reliability Engineer - Datacenter
SpaceXAI Southaven, Mississippi, United States
-
Staff Site Reliability Engineer
Attentive United States · $180K–$240K/yr
-
Team Manager, SRE
Pythian Poland