DevOps Engineer – Azure | Terraform | Kubernetes 6916
Softgic · Mexico City, Mexico City, Mexico
IT Services and IT Consulting · 51-200 employees
About the role
Design, deploy, and manage scalable infrastructure on Microsoft Azure while optimizing CI/CD pipelines and ensuring system reliability. Collaborate with cross-functional teams to implement automation, security controls, and incident response strategies.
What they look for
Requirements
Requires 4–7 years of experience in DevOps or SRE roles with strong hands-on expertise in Azure, Terraform, and Kubernetes. Candidates must possess solid knowledge of Linux, scripting, and cloud infrastructure best practices.
Full description
We are looking for a skilled DevOps Engineer with strong hands-on experience in Microsoft Azure, Terraform, and Kubernetes (AKS) to design, build, and operate secure, scalable, and highly available cloud platforms.
The ideal candidate will play a key role in improving infrastructure automation and reliability, implementing Infrastructure as Code (IaC) practices, optimizing CI/CD pipelines, and ensuring the stability and performance of production environments.
We are looking for a professional with solid experience in production cloud environments, strong troubleshooting skills, and a good understanding of automation, monitoring, networking, and cloud security.
Responsibilities
- Design, deploy, and manage infrastructure on Microsoft Azure, ensuring scalability, availability, performance, and security.
- Manage Azure services such as Azure Kubernetes Service (AKS), Virtual Networks, Load Balancers, Storage Accounts, Azure Monitor, and Application Insights.
- Build, maintain, and optimize Terraform modules for automated infrastructure provisioning.
- Implement Infrastructure as Code (IaC) best practices and maintain consistency across Development, QA, and Production environments.
- Operate and support Kubernetes/AKS clusters across production and non-production environments.
- Manage container deployments, scaling, rolling upgrades, and resiliency strategies.
- Troubleshoot issues related to clusters, nodes, pods, networking, and cloud infrastructure.
- Design, implement, and maintain CI/CD pipelines for infrastructure and application deployments.
- Automate build, testing, deployment, release, and rollback processes.
- Integrate security controls, infrastructure validation, and quality gates into CI/CD pipelines.
- Implement monitoring, logging, and alerting solutions to ensure system availability and reliability.
- Participate in production incident response, Root Cause Analysis (RCA), and continuous improvement initiatives.
- Support Disaster Recovery (DR), backup, and high-availability strategies.
- Participate in on-call rotations when required.
- Collaborate closely with Software Engineering, QA, Security, and Architecture teams.
Requisitos
- 4–7 years of experience in DevOps, Platform Engineering, Site Reliability Engineering (SRE), or similar roles.
- Strong hands-on experience with Microsoft Azure in production environments.
- Proven expertise with Terraform for Infrastructure as Code (IaC).
- Solid hands-on experience with Kubernetes, particularly Azure Kubernetes Service (AKS).
- Experience designing, implementing, and maintaining CI/CD pipelines.
- Strong knowledge of Linux systems and administration.
- Solid understanding of networking, cloud security, and IAM fundamentals.
- Experience troubleshooting infrastructure and resolving incidents in production environments.
- Proficiency in scripting and automation using PowerShell, Bash, and/or Python.
- Experience with monitoring and observability tools, particularly Azure Monitor and Application Insights.
- Knowledge of high-availability architectures, backup strategies, and Disaster Recovery (DR).
- Understanding of cloud infrastructure best practices related to security, scalability, performance, and cost optimization.
- Bachelor’s degree in Computer Science, Systems Engineering, Software Engineering, or a related technical field, or equivalent practical experience.