Site Reliability Engineer (II)
NCR Corporation Hyderabad, Telangana, India
Software Development · 10,001+ employees
About the role
Design, build, and maintain scalable cloud platform solutions while managing CI/CD pipelines and infrastructure-as-code. Drive operational excellence through proactive monitoring, incident response, and the deployment of AI/ML systems at scale.
What they look for
Requirements
Requires at least 3 years of experience with CI/CD tooling and 2 years in AI/ML, SRE, or DevOps roles. Candidates must possess strong scripting skills and proficiency in containerization technologies like Docker and Kubernetes.
Full description
About NCR VOYIX
NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.
Role Overview
The Principal Cloud Engineer plays a key role in designing, implementing, and operating cloud platform capabilities that support reliable, secure, and efficient software delivery at scale. This role requires strong technical depth, a high degree of ownership, and the ability to influence platform direction while partnering closely with application teams and cross-functional stakeholders.
Key Responsibilities
- Design, build, and maintain scalable, secure, and highly available cloud platform solutions.
- Own and improve CI/CD pipelines, deployment automation, and infrastructure-as-code to enable consistent and reliable delivery.
- Drive operational excellence through proactive monitoring, alerting, incident response, and continuous reliability improvements.
- Lead analysis and resolution of complex platform and production issues, including participating in on-call and after-hours support when required.
- Design, build, and deploy machine learning models and AI systems at scale.
- Develop pipelines for data ingestion, training, validation, and inference.
- Develop CI/CD pipelines for AI/ML and platform services.
- Build and maintain cloud-native infrastructure (AWS/Azure/GCP) using IaC tools (Terraform, ARM, etc.)
- Contribute to architectural decisions related to cloud infrastructure, networking, security, and system reliability.
- Partner with application, security, and compliance teams to establish and enforce platform standards and best practices.
- Identify and address technical debt, operational risks, and scalability concerns within the platform.
- Mentor and support other engineers by sharing knowledge, reviewing designs and code, and promoting DevOps best practices.
- Continuously assess and improve platform tooling, performance, cost efficiency, and security posture.
- Document platform architecture, operational procedures, and standards to support consistency and knowledge sharing.
Required Qualifications
- 3+ yrs of hands-on experience with CI/CD tooling, automation, and deployment strategies.
- 2+ yrs of experience in AI/ML engineering, SRE, or DevOps roles.
- Solid understanding of containerization and orchestration technologies (e.g., Docker, Kubernetes).
- Experience with observability, monitoring, and incident management practices.
- Strong scripting or programming skills (e.g., Python, Bash, PowerShell, or similar).
- Demonstrated ability to operate production systems with reliability and security in mind.
- Strong communication skills and ability to collaborate across teams.
Preferred Qualifications
- Experience designing shared platform services or internal developer platforms.
- Familiarity with security best practices, identity management, and compliance controls in cloud environments.
- Experience optimizing cloud cost, performance, and resource utilization.
Engineer Expectations
- Demonstrates consistent ownership of systems and outcomes.
- Takes initiative to identify and solve problems beyond immediate assignments.
- Balances short‑term delivery needs with long‑term platform health.
Offers of employment are conditional upon passage of screening criteria applicable to the job
EEO Statement
Integrated into our shared values is NCR Voyix’s commitment to equal employment opportunity. All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law. NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential. We believe in understanding and respecting differences among all people. Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.
Statement to Third Party Agencies To ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes
“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.com email domain.”
Similar roles
-
Staff Site Reliability Engineer - Federal
ServiceNow San Diego, California, United States · $150K–$262K/yr
-
Senior Site Reliability Engineer
Synapse Health $134K–$184K/yr
-
Senior Site Reliability Engineer
Mirantis Hyderabad, Telangana, India
-
Site Reliability Engineering Lead
Experian Hyderabad, Telangana, India
-
Senior Site Reliability Engineer
Fivetran Novi Sad, Vojvodina, Serbia
-
Senior Site Reliability Engineer
Veeam Software pune, Maharashtra, India