Senior Associate DevOps Engineer
National Payments Corporation of India (NPCI) · Hyderabad, Telangana, India
Financial Services · 1,001-5,000 employees
About the role
The DevOps Engineer will automate, deploy, monitor, and maintain critical applications and infrastructure while collaborating with cross-functional teams. Responsibilities include managing CI/CD pipelines, Kubernetes clusters, and ensuring high availability and security of production systems.
What they look for
Requirements
Candidates must have 3-8 years of experience in DevOps or SRE roles with hands-on expertise in Kubernetes, containerization, and infrastructure automation. Proficiency in Linux, scripting, and cloud platforms is required to support scalable and secure application delivery.
Full description
Profile: Devops Engineer Place of posting: Hyderabad Experience: 3+ years
About the Role We are looking for a DevOps Engineer who can help automate, deploy, monitor, and maintain critical applications and infrastructure. The ideal candidate should have hands-on experience with CI/CD pipelines, Kubernetes, cloud/on-prem infrastructure, automation, observability, and production support. The engineer will work closely with developers, QA teams, security teams, and infrastructure teams to ensure reliable, scalable, and secure application delivery.
Key Responsibilities Application Deployment & Automation Build and maintain CI/CD pipelines for application deployment. Automate build, testing, release, and deployment processes. Support blue-green, rolling, and canary deployments. Improve deployment speed and reliability through automation.
Kubernetes & Container Platform Deploy and manage containerized applications using Docker and Kubernetes. Troubleshoot application issues running on Kubernetes clusters. Manage Kubernetes resources such as Deployments, StatefulSets, Services, Ingresses, ConfigMaps, and Secrets. Implement Helm charts for application packaging and deployment.
Infrastructure as Code (IaC) Provision and manage infrastructure through code. Maintain reusable and version-controlled infrastructure templates. Automate environment creation and configuration.
Monitoring, Alerting & Logging Build dashboards and alerts to monitor applications and infrastructure. Analyze logs and system metrics to identify issues proactively. Support incident troubleshooting and root cause analysis. Improve platform observability and reliability.
Platform Operations & Reliability Ensure high availability and performance of production systems. Participate in on-call support and incident management activities. Perform capacity planning, performance tuning, and operational improvements. Implement backup, recovery, and disaster recovery procedures.
Security & Compliance Implement DevSecOps best practices. Manage secrets, certificates, access controls, and security policies. Support vulnerability remediation and patch management. Ensure adherence to enterprise security standards.
Collaboration Work closely with development teams to improve application reliability. Participate in architecture discussions and solution design. Document processes, runbooks, and operational procedures.
Required Technical Skills CI/CD & Source Control Git, GitLab CI/CD / Jenkins / GitHub Actions Pipeline automation Release management
Containers & Orchestration Docker Kubernetes Helm
Infrastructure Automation Terraform Ansible Infrastructure as Code (IaC)
Monitoring & Logging Prometheus Grafana OpenSearch / ELK Stack VictoriaMetrics Splunk (optional)
Operating Systems Linux (RHEL, Ubuntu) Shell/Bash scripting
Cloud & Platform Technologies AWS / Azure / GCP (any one preferred) Networking fundamentals Load Balancers DNS SSL/TLS Certificates
Scripting & Automation Python Bash API Automation
Preferred Skills (Good to Have) Kafka OpenSearch Apache Superset Apache Flink Rook-Ceph / Storage Platforms ArgoCD / GitOps Service Mesh (Istio) OpenTelemetry SRE Practices DevSecOps Tools Container Security Chaos Engineering Performance Testing
Soft Skills (Layman-Friendly) We are looking for someone who: Can solve production issues calmly and logically. Likes automating repetitive work instead of doing it manually. Can work with developers and explain technical problems clearly. Takes ownership of issues and follows them through to resolution. Understands how applications behave in production environments. Is eager to learn new tools and technologies. Has a proactive mindset and identifies problems before they impact users. Can create simple documentation and knowledge-sharing material.
Experience 3–8 years of experience in DevOps, Platform Engineering, Cloud Engineering, or SRE roles. Experience supporting business-critical or high-availability applications. Experience managing Kubernetes-based workloads in production environments.
AI Search Keywords CI/CD, GitLab, Jenkins, GitHub Actions, DevOps, Kubernetes, Docker, Helm, Terraform, Ansible, Linux, Bash, Python, Monitoring, Grafana, Prometheus, OpenSearch, ELK, VictoriaMetrics, Cloud, AWS, Azure, Automation, GitOps, ArgoCD, Kafka, DevSecOps, SRE, Incident Management, Production Support, Infrastructure as Code, Application Deployment, Observability, Containerization, Site Reliability Engineering.