Weekday AI

Staff DevOps Engineer

Weekday AI · Bengaluru, Karnataka, India · ₹4M–₹6M/yr

Technology, Information and Internet · 11-50 employees

7 h ago
Senior (5-10 yrs) Full-time India
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Staff DevOps Engineer will lead the design and implementation of scalable cloud-native infrastructure and CI/CD automation. They will also collaborate with cross-functional teams to ensure system reliability, security, and operational excellence.

What they look for

Kubernetes Terraform AWS GCP Azure CI/CD Jenkins GitHub Actions Infrastructure as Code Cloud Security Observability Containerization Site Reliability Engineering System Resilience Automation

Requirements

Candidates must have 8+ years of hands-on experience in DevOps or SRE roles with deep expertise in Kubernetes, Terraform, and cloud platforms. Strong communication skills and the ability to mentor engineering teams are essential for this senior-level position.

Full description

𝗧𝗵𝗶𝘀 𝗿𝗼𝗹𝗲 𝗶𝘀 𝗳𝗼𝗿 𝗼𝗻𝗲 𝗼𝗳 𝘁𝗵𝗲 𝗪𝗲𝗲𝗸𝗱𝗮𝘆'𝘀 𝗰𝗹𝗶𝗲𝗻𝘁𝘀

𝗦𝗮𝗹𝗮𝗿𝘆 𝗿𝗮𝗻𝗴𝗲: 𝗥𝘀 𝟰𝟮𝟬𝟬𝟬𝟬𝟬 - 𝗥𝘀 𝟱𝟲𝟬𝟬𝟬𝟬𝟬 (𝗶𝗲 𝗜𝗡𝗥 𝟰𝟮-𝟱𝟲 𝗟𝗣𝗔)

Experience: 8+ yrs

Location: Bengaluru, Karnataka, India

Job Type: Full-time

We are seeking an experienced Staff DevOps Engineer to lead the design, implementation, and continuous improvement of cloud-native infrastructure, CI/CD automation, and operational reliability practices. This role is ideal for a hands-on DevOps professional who combines deep technical expertise with the ability to influence engineering strategy, drive automation initiatives, and mentor high-performing teams.

As a senior member of the engineering organization, you will play a key role in building scalable deployment platforms, improving system resilience, strengthening cloud security practices, and enabling faster, safer software delivery across multiple product teams. You will collaborate closely with engineering, infrastructure, security, and platform teams to ensure that cloud environments are secure, observable, reliable, and optimized for rapid growth.

Key Responsibilities• Design, build, and maintain scalable CI/CD pipelines, deployment automation workflows, and infrastructure-as-code frameworks.

  • Drive end-to-end DevOps automation across build, test, release, and deployment processes to improve engineering productivity and delivery speed.
  • Implement and manage cloud-native infrastructure using Kubernetes, Terraform, and modern cloud platforms such as AWS, GCP, or Azure.
  • Partner with application, platform, and security teams to improve system reliability, scalability, and operational excellence.
  • Establish monitoring, logging, alerting, and observability standards to ensure proactive detection and rapid resolution of production issues.
  • Define and track operational KPIs such as deployment frequency, lead time for changes, infrastructure availability, incident reduction, and recovery time objectives.
  • Implement security best practices including IAM, RBAC, secrets management, encryption, vulnerability scanning, and audit logging across cloud environments.
  • Lead infrastructure capacity planning, performance optimization, disaster recovery readiness, and cost-efficiency initiatives.
  • Contribute to architecture reviews, technology roadmaps, and platform modernization efforts to support future scalability and reliability requirements.
  • Mentor DevOps and SRE engineers through code reviews, technical guidance, operational best practices, and knowledge-sharing initiatives.
  • Collaborate with senior leadership to communicate delivery progress, operational risks, infrastructure health, and strategic improvement opportunities.

What Makes You a Great Fit• 8+ years of hands-on experience in DevOps, Site Reliability Engineering (SRE), or cloud infrastructure engineering roles.

  • Strong expertise in Kubernetes, Terraform, Jenkins, GitHub Actions, CI/CD pipelines, and infrastructure automation.
  • Proven experience managing large-scale cloud-native environments on AWS, GCP, or Azure.
  • Deep understanding of containerization, orchestration, configuration management, and deployment automation.
  • Strong knowledge of monitoring, observability, incident response, and operational resilience practices.
  • Hands-on experience implementing cloud security controls, compliance standards, access governance, and secure infrastructure patterns.
  • Ability to troubleshoot complex distributed systems, identify root causes, and implement sustainable operational improvements.
  • Excellent communication and stakeholder management skills with the ability to collaborate effectively across engineering, infrastructure, security, and leadership teams.
  • Experience leading technical initiatives, influencing platform strategy, and mentoring engineering teams in a fast-paced product environment.
  • A proactive, ownership-driven mindset with a passion for automation, reliability, scalability, and continuous improvement in modern cloud-native software delivery ecosystems.