Softgic

DevOps Engineer – Azure | Terraform | Kubernetes 6916

Softgic · Mexico City, Mexico City, Mexico

IT Services and IT Consulting · 51-200 employees

Yesterday
Senior (5-10 yrs) Full-time Mexico
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Design, deploy, and manage scalable infrastructure on Microsoft Azure while optimizing CI/CD pipelines and ensuring system reliability. Collaborate with cross-functional teams to implement automation, security controls, and incident response strategies.

What they look for

Microsoft Azure Terraform Kubernetes AKS Infrastructure as Code CI/CD Linux PowerShell Bash Python Azure Monitor Application Insights Networking Cloud Security IAM Disaster Recovery

Requirements

Requires 4–7 years of experience in DevOps or SRE roles with strong hands-on expertise in Azure, Terraform, and Kubernetes. Candidates must possess solid knowledge of Linux, scripting, and cloud infrastructure best practices.

Full description

We are looking for a skilled DevOps Engineer with strong hands-on experience in Microsoft Azure, Terraform, and Kubernetes (AKS) to design, build, and operate secure, scalable, and highly available cloud platforms.

The ideal candidate will play a key role in improving infrastructure automation and reliability, implementing Infrastructure as Code (IaC) practices, optimizing CI/CD pipelines, and ensuring the stability and performance of production environments.

We are looking for a professional with solid experience in production cloud environments, strong troubleshooting skills, and a good understanding of automation, monitoring, networking, and cloud security.

Responsibilities

  • Design, deploy, and manage infrastructure on Microsoft Azure, ensuring scalability, availability, performance, and security.
  • Manage Azure services such as Azure Kubernetes Service (AKS), Virtual Networks, Load Balancers, Storage Accounts, Azure Monitor, and Application Insights.
  • Build, maintain, and optimize Terraform modules for automated infrastructure provisioning.
  • Implement Infrastructure as Code (IaC) best practices and maintain consistency across Development, QA, and Production environments.
  • Operate and support Kubernetes/AKS clusters across production and non-production environments.
  • Manage container deployments, scaling, rolling upgrades, and resiliency strategies.
  • Troubleshoot issues related to clusters, nodes, pods, networking, and cloud infrastructure.
  • Design, implement, and maintain CI/CD pipelines for infrastructure and application deployments.
  • Automate build, testing, deployment, release, and rollback processes.
  • Integrate security controls, infrastructure validation, and quality gates into CI/CD pipelines.
  • Implement monitoring, logging, and alerting solutions to ensure system availability and reliability.
  • Participate in production incident response, Root Cause Analysis (RCA), and continuous improvement initiatives.
  • Support Disaster Recovery (DR), backup, and high-availability strategies.
  • Participate in on-call rotations when required.
  • Collaborate closely with Software Engineering, QA, Security, and Architecture teams.

Requisitos

  • 4–7 years of experience in DevOps, Platform Engineering, Site Reliability Engineering (SRE), or similar roles.
  • Strong hands-on experience with Microsoft Azure in production environments.
  • Proven expertise with Terraform for Infrastructure as Code (IaC).
  • Solid hands-on experience with Kubernetes, particularly Azure Kubernetes Service (AKS).
  • Experience designing, implementing, and maintaining CI/CD pipelines.
  • Strong knowledge of Linux systems and administration.
  • Solid understanding of networking, cloud security, and IAM fundamentals.
  • Experience troubleshooting infrastructure and resolving incidents in production environments.
  • Proficiency in scripting and automation using PowerShell, Bash, and/or Python.
  • Experience with monitoring and observability tools, particularly Azure Monitor and Application Insights.
  • Knowledge of high-availability architectures, backup strategies, and Disaster Recovery (DR).
  • Understanding of cloud infrastructure best practices related to security, scalability, performance, and cost optimization.
  • Bachelor’s degree in Computer Science, Systems Engineering, Software Engineering, or a related technical field, or equivalent practical experience.