SRE Architect- Security
Prodapt Chennai, Tamil Nadu, India
Technology, Information and Internet · 5,001-10,000 employees
About the role
The SRE Architect will design and govern highly reliable, scalable, and secure cloud platforms while embedding security into the entire software lifecycle. Responsibilities include defining SRE standards, implementing DevSecOps practices, and leading infrastructure automation and incident response strategies.
What they look for
Requirements
Candidates must have over 10 years of experience in SRE, DevOps, or Cloud Engineering with strong expertise in distributed systems and cloud-native technologies. Proficiency in infrastructure as code, container orchestration, and security frameworks is essential for this role.
Full description
Overview
We are looking for an experienced SRE Architect with a strong security focus to design, build, and govern highly reliable, scalable, resilient, and secure technology platforms. The role combines Site Reliability Engineering, Cloud Architecture, DevSecOps, and Cybersecurity, with responsibility for embedding security into every stage of the software and infrastructure lifecycle.
Responsibilities
- Key Responsibilities
SRE & Platform Architecture
- Define and implement enterprise-wide SRE architecture, standards, and best practices.
- Design highly available, scalable, fault-tolerant, and resilient cloud platforms.
- Establish SLIs, SLOs, SLAs, error budgets, and reliability engineering practices.
- Drive automation of infrastructure, deployment, monitoring, incident response, and operational processes.
- Design disaster recovery, business continuity, backup, and resilience strategies.
- Lead capacity planning, performance engineering, and reliability improvements.
Security & DevSecOps
- Integrate security controls into CI/CD pipelines, infrastructure, and application platforms.
- Implement DevSecOps practices including SAST, DAST, SCA, container security, secrets management, and IaC security.
- Define secure cloud architecture aligned with Zero Trust principles.
- Implement IAM, least-privilege access, encryption, key management, network segmentation, and workload security.
- Partner with security teams to identify and remediate vulnerabilities and security risks.
- Establish security and compliance guardrails across cloud and on-premise environments.
- Support security incident response, threat detection, and post-incident remediation.
Observability & Operations
- Architect centralized monitoring, logging, tracing, metrics, and security observability.
- Define proactive detection mechanisms for reliability and security incidents.
- Establish automated alerting, incident management, and remediation workflows.
- Lead root-cause analysis and drive permanent corrective actions.
- Develop reliability and security dashboards for engineering and leadership teams.
Cloud & Infrastructure
- Architect secure platforms across AWS, Azure, and/or Google Cloud.
- Implement infrastructure automation using Terraform, CloudFormation, Ansible, or equivalent technologies.
- Design and govern Kubernetes/container platforms and cloud-native workloads.
- Establish secure networking, service connectivity, API security, and platform controls.
Leadership & Governance
- Provide architectural leadership across SRE, DevOps, Platform Engineering, and Security teams.
- Define engineering standards, reference architectures, and operational/security policies.
- Mentor SRE, DevOps, and platform engineers.
- Work with engineering, security, product, and business stakeholders to balance reliability, security, performance, and cost.
- Evaluate emerging technologies and drive continuous improvement.
Requirements
- Required Skills & Experience
- 10+ years of experience in SRE, DevOps, Cloud, Infrastructure, or Platform Engineering, with significant architecture experience.
- Strong understanding of SRE principles, distributed systems, cloud architecture, and reliability engineering.
- Hands-on experience with AWS, Azure, or GCP.
- Strong Kubernetes and container-platform experience.
- Expertise in CI/CD and DevSecOps practices.
- Strong knowledge of Linux, networking, DNS, HTTP/TLS, databases, and distributed systems.
- Experience with Terraform or equivalent Infrastructure as Code technologies.
- Experience with observability platforms such as Prometheus, Grafana, OpenTelemetry, ELK, Splunk, or equivalent.
- Strong understanding of cloud and application security.
- Experience implementing IAM, secrets management, encryption, vulnerability management, and security controls.
- Strong scripting/programming skills in Python, Go, Bash, or similar.
- Experience designing disaster recovery and high-availability solutions.
Preferred Certifications
- AWS/Azure/GCP Solutions Architect or Security certification.
- Certified Kubernetes Administrator (CKA) or equivalent.
- CISSP, CCSP, or equivalent security certification.
- SRE/DevOps certifications are an advantage.
Key Success Metrics
- Improved platform availability and reliability.
- Reduction in MTTR and operational incidents.
- Improved security posture and vulnerability remediation times.
- Increased automation and deployment efficiency.
- Strong compliance with security and operational standards.
- Improved observability, resilience, and disaster recovery readiness.
Ideal Candidate
The ideal candidate is a security-minded SRE architect who can operate at both strategic and technical levels—designing resilient platforms, embedding security by design, automating operations, and influencing engineering teams toward a culture of reliability, security, and continuous improvement.
Similar roles
-
Senior Staff Software Engineer – SRE & AIOps
ServiceNow Santa Clara, California, United States · $191K–$334K/yr
-
Site Reliability Expert
Valtech Montreal, Quebec, Canada · CA$120K–CA$170K/yr
-
Senior Site Reliability Engineer
Planet Canada · $143K–$203K/yr
-
Staff SRE Software Engineer, Google Home
Google San Francisco, California, United States · $207K–$300K/yr
-
Site Reliability Engineer III- Network
JPMorgan Chase & Co. Hyderabad, Telangana, India
-
Senior Site Reliability Engineer I
Braze San Francisco, California, United States · $129K–$232K/yr