Senior Site Reliability Engineer (f/m/d) - Germany (remote)
Cfactory Munich, Bavaria, Germany
IT Services and IT Consulting · 51-200 employees
Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
You will manage and secure the existing cplace Cloud 1.0 environment while actively shaping the architecture and operations of the Kubernetes-based cplace Cloud 2.0. Additionally, you will drive AI integration in operations and ensure technical compliance and reliability through infrastructure-as-code practices.
What they look for
Requirements
Candidates must have a degree in a STEM field and several years of experience as an SRE, DevOps, or platform engineer in business-critical environments. Strong hands-on expertise in Kubernetes, infrastructure-as-code, Linux, and software engineering in Go or Python is required.
Benefits
Full description
Your tasks
cplace is the platform for project and portfolio management that leading companies use to steer their most complex initiatives – grown in the DACH region and, with our launch in the US, on its way to becoming an international provider. As a Senior SRE in our Cloud Operations team, you will run our existing cplace Cloud 1.0 reliably and securely for customers in automotive, life sciences, manufacturing and retail – and actively shape the architecture and operating model of our Kubernetes-based cplace Cloud 2.0. AI is in cplace’s DNA: we use it intensively across all our work and expect you to apply it productively and critically and to help us take it further. • cplace Cloud 2.0: Co-building the Kubernetes platform – from cluster design, networking and storage to tenant isolation and scaling – plus planning and driving the migration of customer environments from Cloud 1.0
- Everything as code: Development of reusable Terraform modules, GitOps repositories and our own platform software (self-service portal, APIs, automation), including code reviews, automated tests and policy as code
- cplace Cloud 1.0: Operation and improvement of our environment of Linux servers, containers, SQL databases and Elasticsearch; lasting resolution of bugs, findings and capacity issues (incident and problem management, post-mortems); conversion of manual procedures into versioned, tested code (Ansible, Terraform, n8n)
- Improvement of monitoring, logging and alerting; ownership of backup & recovery, disaster recovery and business continuity, including regular testing
- Technical implementation of security and compliance requirements (e.g. SOC 2, ISO 27001, GDPR) and optimisation of cost and capacity
- Driving AI in operations, e.g. for incident and log analysis and agents for runbooks and routine tasks
- Close collaboration with product development for smooth releases, technical representation of the team in customer meetings, tenders and customer projects, and knowledge sharing through internal and external documentation and mentoring
- On-call duty in a fair rotation
What we expect
- Degree in computer science or a related STEM field, or comparable vocational training, plus several years (ideally 5+) of experience as an SRE, DevOps or platform engineer in business-critical production environments
- Solid hands-on experience with Kubernetes in production (operations, upgrades, troubleshooting, storage, networking) and with at least one cloud provider – AWS, GCP, Azure and/or Hetzner Cloud a strong plus
- Deep experience with infrastructure as code (Terraform, Ansible), CI/CD, GitOps and Git-based collaboration via pull requests and code reviews – with modular, tested code that stays maintainable for the team
- Strong Linux skills and experience with SQL databases (e.g. MariaDB), Elasticsearch/OpenSearch and observability stacks (e.g. Prometheus, Grafana, Loki/ELK)
- Solid software engineering skills, ideally in Go or Python – tools and automation with tests and clean structure rather than one-off scripts; confident Bash scripting a given
- Good understanding of cloud security (e.g. network segmentation, secrets management, WAF)
- Hands-on experience with AI tools in everyday engineering and a good sense of their strengths and limits
- Customer-focused, structured way of working, ability to explain technical topics clearly, and fluent German and English
This is what we offer
- Flexible work model and remote work option
- Room for creativity, co-design and further development
- Competitive salary, 30 days vacation & sabbatical option
- Wellpass, job bike & corporate benefits
- Modern and central offices in Munich, Hannover and Ludwigsburg
- Hardware of your choice
Be part of it
We are looking forward to meeting you. career@cplace.com
www.cplace.com | #bestteam collaboration Factory GmbH | Office München | Arnulfstraße 34 | 80335 München | Tel: 089/ 80 91 33 232 About us
cplace is a modern software platform for project and portfolio management (PPM). It revolutionizes and transforms the way people and organizations collaborate, enabling companies to efficiently manage complex projects. cplace combines the reliability of standard software with the flexibility of customized solutions. The adaptive platform with built-in AI provides a unified data foundation, promotes cross-site and cross-functional collaboration, and can be adapted in real time.
International industry leaders, including those in the automotive, aerospace, pharmaceutical and life sciences, and mechanical engineering industries, rely on cplace to develop innovative and complex products, utilize resources efficiently, improve decision-making processes, and successfully complete projects.
Behind the cplace brand is collaboration Factory GmbH, founded in 2014 and headquartered in Munich, Germany, with the subsidiary cplace, Inc. in North Carolina, US.
Similar roles
-
Staff Site Reliability Engineer
Okta Bengaluru, Karnataka, India
-
Lead Site Reliability Engineer (SRE) | Budapest
Deloitte Budapest, Central Hungary, Hungary
-
Site Reliability Engineer
Okta Bengaluru, Karnataka, India
-
Senior Site Reliability Engineer
Guidewire Software Bengaluru, Karnataka, India
-
Site Reliability Engineer, AVP
RBS Gurugram, Haryana, India
-
Senior Site reliability engineer
Capgemini Utrecht, Utrecht, Netherlands