Senior Platform Engineer - Cloud & Kubernetes
Weekday AI Abu Dhabi, Abu Dhabi, United Arab Emirates · ₹3M–₹4M/yr
Technology, Information and Internet · 11-50 employees
About the role
The role involves designing, deploying, and managing scalable cloud-native platforms on Microsoft Azure and Kubernetes. You will also be responsible for implementing infrastructure automation, observability, and disaster recovery solutions to ensure high availability and security.
What they look for
Requirements
Candidates must have 5+ years of experience in platform engineering, DevOps, or SRE roles with strong hands-on expertise in Azure and Kubernetes. Proficiency in Infrastructure-as-Code, CI/CD pipelines, and cloud governance is essential for this position.
Full description
𝗧𝗵𝗶𝘀 𝗿𝗼𝗹𝗲 𝗶𝘀 𝗳𝗼𝗿 𝗼𝗻𝗲 𝗼𝗳 𝘁𝗵𝗲 𝗪𝗲𝗲𝗸𝗱𝗮𝘆'𝘀 𝗰𝗹𝗶𝗲𝗻𝘁𝘀
𝗦𝗮𝗹𝗮𝗿𝘆 𝗿𝗮𝗻𝗴𝗲: 𝗥𝘀 𝟮𝟲𝟬𝟬𝟬𝟬𝟬 - 𝗥𝘀 𝟰𝟱𝟬𝟬𝟬𝟬𝟬 (𝗶𝗲 𝗜𝗡𝗥 𝟮𝟲-𝟰𝟱 𝗟𝗣𝗔)
Experience: 5+ yrs
Location: Abu Dhabi, United Arab Emirates
Job Type: Full-time
We are looking for an experienced Senior Platform Engineer to design, implement, secure, and operate scalable cloud-native platforms using Microsoft Azure and Kubernetes. The role focuses on building highly available, resilient, secure, and well-governed production environments while driving automation, observability, infrastructure-as-code, and platform engineering best practices.
The ideal candidate will have strong hands-on expertise across Azure, AKS, Kubernetes, Terraform, CI/CD, cloud governance, security, observability, and Disaster Recovery, along with the ability to provide technical leadership across complex platform environments.
Key Responsibilities
- Design, deploy, and manage Microsoft Azure infrastructure and AKS/Kubernetes platforms.
- Implement and maintain Disaster Recovery, backup, restore, and business continuity solutions.
- Build infrastructure automation using Terraform, Helm, Azure DevOps, and CI/CD pipelines.
- Manage Kubernetes networking, ingress, storage, namespaces, resource limits, security, and platform configurations.
- Implement observability solutions using Prometheus, Grafana, Loki, and related monitoring technologies.
- Manage TLS certificates, secrets, access controls, RBAC, and platform security mechanisms.
- Establish cloud and Kubernetes governance covering resource organisation, naming, tagging, security policies, and operational standards.
- Define Infrastructure-as-Code and CI/CD governance, including reusable Terraform modules, state management, code reviews, approval gates, environment promotion, and artifact management.
- Participate in architecture and technical design reviews for platforms, applications, integrations, and infrastructure changes.
- Drive security and compliance readiness through vulnerability remediation, access reviews, security baselines, audit controls, and policy enforcement.
- Define and monitor platform availability, SLIs, SLOs, capacity, performance, and operational health.
- Lead incident and problem management, including root-cause analysis, corrective actions, and prevention of recurring issues.
- Perform capacity planning, performance optimisation, and cloud cost optimisation / FinOps activities.
- Own DR testing, RTO/RPO validation, backup and recovery standards, and periodic recovery exercises.
- Plan and execute platform migrations, infrastructure upgrades, and cloud transformation initiatives.
- Evaluate emerging platform technologies, conduct POCs, and establish approved patterns for production adoption.
- Maintain technical documentation including HLDs, LLDs, architecture diagrams, SOPs, runbooks, troubleshooting guides, and DR procedures.
- Provide technical leadership, mentoring, and knowledge sharing across platform engineering teams.
What Makes You a Great Fit
- 5+ years of experience in platform engineering, cloud infrastructure, DevOps, SRE, or related roles.
- Strong hands-on expertise in Microsoft Azure and Kubernetes, particularly AKS.
- Strong experience with Docker, Helm, Terraform, and Azure DevOps / CI/CD.
- Good understanding of Azure networking, including VNets, NSGs, Private Endpoints, and Azure Firewall.
- Experience implementing observability using Prometheus, Grafana, Loki, or similar platforms.
- Strong knowledge of Linux, Bash scripting, Git, and GitOps practices.
- Proven experience with Disaster Recovery, backup/restore, RTO/RPO planning, and recovery testing.
- Strong understanding of Kubernetes security, governance, RBAC, policy enforcement, and Azure Policy.
- Experience establishing reusable Infrastructure-as-Code patterns and platform engineering standards.
- Strong understanding of CI/CD governance, release controls, environment promotion, and source-control standards.
- Experience working with databases such as PostgreSQL, MongoDB, MySQL, or Azure SQL.
- Strong knowledge of incident management, problem management, RCA, capacity planning, performance optimisation, and cloud cost management.
- Experience working in security-, compliance-, audit-, or governance-controlled environments.
- Ability to create and maintain HLDs, LLDs, architecture diagrams, operational runbooks, SOPs, and technical documentation.
- Strong experience participating in architecture and technical design reviews.
- Excellent troubleshooting, analytical, and problem-solving capabilities.
- Strong technical leadership, mentoring, communication, and cross-functional collaboration skills.
- Experience in banking or other regulated industries would be an advantage.
- Strong understanding of high availability, production operations, platform modernisation, and cloud transformation initiatives.
Similar roles
-
C005025 Kubernetes Platform Operations Maint Expert (NS) - TUE 6 Oct RELAUNCH
EMW, Inc. Brussels, Brussels-Capital, Belgium
-
Platform Engineer | Kubernetes, CI/CD & Observability
knowmad mood Madrid, Community of Madrid, Spain
-
Product Manager - Cloudera Kubernetes
Cloudera Prague, Prague, Czechia
-
DevOps Engineer - Linux/Kubernetes
Schuberg Philis Schiphol-Rijk, North Holland, Netherlands
-
Production Support Analyst Senior(5–8 years in DevOps cloud(AWS/Azure) and Kubernetes/EKS)
FIS pune, Maharashtra, India
-
Classified Cyber Systems Engineer – Kubernetes / Principal Classified Cyber Systems Engineer – Kubernetes
Northrop Grumman Huntsville, Alabama, United States · $96K–$179K/yr