Deutsche Telekom IT Solutions

Sovereign Engineering Platform SRE - T Cloud Public (REF5740Q)

Deutsche Telekom IT Solutions · Budapest, Central Hungary, Hungary

IT Services and IT Consulting · 5,001-10,000 employees

6 h ago
Senior (5-10 yrs) Full-time Hungary
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Design, build, and operate secure Kubernetes-based infrastructure for AI-assisted software development and engineering workflows. Implement GitOps, observability, and security governance to ensure reliable and audit-compliant platform operations.

What they look for

Kubernetes SRE Platform Engineering GitOps Terraform Ansible Helm Python CI/CD Observability Linux Security Infrastructure as Code GPU Vault Prometheus

Requirements

Requires 5+ years of experience in SRE, platform engineering, or DevOps with strong expertise in Kubernetes and Linux. Candidates must have hands-on experience with Infrastructure as Code, CI/CD pipelines, and secure configuration management in high-security environments.

Full description

Company Description

As Hungary’s most attractive employer in 2025 (according to Randstad’s representative survey), Deutsche Telekom IT Solutions is a subsidiary of the Deutsche Telekom Group. The company provides a wide portfolio of IT and telecommunications services with more than 5300 employees. We have hundreds of large customers, corporations in Germany and in other European countries.

DT-ITS recieved the Best in Educational Cooperation award from HIPA in 2019, acknowledged as the the Most Ethical Multinational Company in 2019. The company continuously develops its four sites in Budapest, Debrecen, Pécs and Szeged and is looking for skilled IT professionals to join its team.

Job Description

Mission 

Design, build, and operate the secure infrastructure foundation used by Meridian engineering teams for AI-assisted software development, model experimentation, repository analysis, CI/CD execution, and controlled handover work in isolated or sovereignty-sensitive environments 

Role focus 

This infrastructure and operations role centers on Kubernetes-based engineering platforms, GitOps, private registries, internal model endpoints, observability, access control, and reliable operations for AI-enabled SDLC workloads. The candidate should enable engineering velocity while preserving security, auditability, and operational discipline 

Key responsibilities 

  • Build and operate Kubernetes environments that host AI engineering tools, internal model gateways, retrieval components, workflow services, CI/CD runners, and documentation services 
  • Implement GitOps and Infrastructure as Code patterns for reproducible provisioning, configuration, policy enforcement, platform upgrades, and disaster recovery readiness 
  • Manage private registries, package mirrors, secrets, identity integration, network segmentation, storage classes, backup routines, and controlled connectivity models 
  • Provide observability for engineering workloads, including metrics, logs, traces, GPU and CPU utilization, service health, cost signals, and operational runbooks 
  • Work with software, security, and architecture teams to ensure the platform supports AI-assisted SDLC workflows without creating uncontrolled data exposure or audit gaps 

Examples of market tools, models, and platform components expected 

  • Platform tooling such as Kubernetes, Helm, Terraform, Ansible, ArgoCD, Crossplane, GitLab runners, Jenkins agents, private registries, and internal package mirrors. 
  • AI platform components such as vLLM, Ollama, OpenAI-compatible gateways, Qdrant or similar vector stores, Open WebUI, Continue-compatible endpoints, and workflow services. 
  • Observability and operations stacks such as Prometheus, Grafana, Loki, OpenTelemetry, ELK/OpenSearch, Alertmanager, SRE runbooks, and incident management tooling. 
  • Security and governance components such as Vault, Keycloak, network policies, RBAC, admission controls, image scanning, SBOM tooling, and audit logging. 
  • Infrastructure awareness covering GPU-backed nodes, CPU-only fallback, storage performance, network isolation, proxy patterns, on-premise environments, and dedicated landing zones. 

Qualifications

Candidate profile 

  • 5+ years in SRE, platform engineering, DevOps, cloud infrastructure, or operations roles with strong Kubernetes and Linux expertise. 
  • Proven experience building and operating production-grade engineering platforms with GitOps, Infrastructure as Code, observability, and operational runbooks. 
  • Hands-on skills in Terraform, Ansible, Helm, Python or shell scripting, CI/CD runners, private registries, and secure configuration management. 
  • Good understanding of networking, storage, secrets, access control, monitoring, backup, disaster recovery, and operational hardening in high-security environments. 
  • Comfortable supporting AI-enabled engineering workloads in sovereignty-driven contexts where isolation, controlled data handling, reliability, and auditability are mandatory. 

Additional Information

Please note: remote working is only possible from within Hungary due to European taxation regulations.

* Please be informed that our remote working possibility is only available within Hungary due to European taxation regulation.

  • Company: Deutsche Telekom TSI Hungary Kft.