Open Text Corporation

Lead Site Reliability Administrator

Open Text Corporation Mississauga, Ontario, Canada · CA$118K–CA$177K/yr

Software Development · 10,001+ employees

7 h ago
sre Senior (5-10 yrs) Full-time Canada
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Site Reliability Engineer will architect and manage scalable cloud infrastructure while ensuring high availability and performance. They will also lead incident response, perform root cause analysis, and collaborate with development teams to optimize CI/CD pipelines.

What they look for

Kubernetes Azure Helm Terraform PowerShell Python Bash Go CI/CD Octopus Deploy Jenkins GitLab CI GitHub Actions Observability Infrastructure as code Site reliability engineering

Requirements

Candidates must hold a bachelor's degree in Computer Science, Engineering, or a related field. Proficiency in scripting languages and experience with Azure infrastructure, Kubernetes, and observability tools are required.

Benefits

Health benefits Paid time off Vacation entitlement Variable compensation Commission opportunities

Full description

OPENTEXT - THE INFORMATION COMPANY

OpenText is a global leader in information management, where innovation, creativity, and collaboration are the key components of our corporate culture. As a member of our team, you will have the opportunity to partner with the most highly regarded companies in the world, tackle complex issues, and contribute to projects that shape the future of digital transformation.

 

AI-First. Future-Driven. Human-Centered.

At OpenText, AI is at the heart of everything we do—powering innovation, transforming work, and empowering digital knowledge workers. We are hiring talent AI can't replace to help us shape the future of information management. Join us.

 

About Us: Opentext is a leading innovator in cloud solutions, dedicated to providing robust and scalable infrastructure to support our clients' needs. We are seeking a talented Site Reliability Engineer (SRE) to join our dynamic team and help us maintain and improve our cloud operations.

 

Your Impact:

As a Site Reliability Engineer (SRE) at OpenText, you will be responsible for ensuring the reliability, scalability, and performance of our cloud infrastructure. You will work closely with development and operations teams to design, implement, and maintain systems that are resilient and efficient.

 

What The Role Offers:

 

  • Architect and manage Kubernetes clusters supporting pods across multi-region deployments to ensure high availability and scalability.
  • Design and maintain Azure -based infrastructure with a focus on security, performance, and cost optimization.
  • Develop and maintain Helm charts for consistent and automated application deployment.
  • Use Terraform to provision and manage infrastructure as code across multiple environments.
  • Monitor and optimize Windows and Linux-based systems, ensuring performance, reliability, and compliance with operational standards.
  • Collaborate with development teams to ensure seamless CI/CD integration and deployment of microservices.
  • Implement and maintain CI/CD pipelines using Octopus Deploy, Jenkins, GitLab CI, and GitHub Actions.
  • Lead root cause analysis and resolution of infrastructure and application performance issues.
  • Participate in on-call rotations and lead incident response for critical systems, ensuring rapid recovery and postmortem analysis.
  • Leverage generative AI tools to accelerate scripting, documentation, troubleshooting, and automation tasks.
  • Explore and contribute to AI-driven observability, alerting, and self-healing strategies for cloud infrastructure and applications, including anomaly detection and automated remediation.

 

What You Need To Succeed

 

  • Bachelor’s degree in Computer Science, Engineering, or a related field.
  • Azure DevOps certification or equivalent certifications.
  • Familiarity with observability tools such as Data Dog, Zabbix, Grafana, ELK stack, or AI-enhanced monitoring platforms.
  • Proficiency in scripting languages such as PowerShell, Python, Bash, or Go.
  • Experience presenting reliability strategies to leadership and driving cross-functional initiatives.
  • Exposure to AI/ML-based monitoring, anomaly detection, or automated remediation systems.
  • Interest in contributing to internal AI adoption strategies for cloud operations.

 

One Last Thing   OpenText is more than a corporation—it’s a global community built on trust, character, and purpose. Here, we act ethically, care deeply about people, and always put our clients first. We help teams succeed through collaboration, tackle challenges with resilience, and innovate with intention.

 

OpenText's commitment to diversity and inclusion surpasses legal requirements, evident in our Equal Employment Opportunity Statement of Policy which promotes a respectful and empowering environment for employees of all backgrounds, culture, national origin, race, color, gender, gender identification, sexual orientation, family status, age, veteran status, disability, religion, or other basis protected by applicable laws.

 

If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please submit a ticket at Ask HR. Our proactive approach fosters collaboration, innovation, and personal growth, enriching OpenText's vibrant workplace.

 

Compensation: At OpenText, we offer a thoughtfully designed benefits package that supports your physical, emotional, and financial wellbeing. As you move through the hiring process, we’re happy to provide more details about our compensation programs, including variable and commission compensation opportunities for eligible roles, vacation entitlement, and paid time off.

Salary Range: $ 118,000 - $177,000 CAD per annum ; Depending on the candidate’s education, experience, skills, geographical location, and alignment with internal equity and external market, actual salary may vary and be higher or lower than the range posted.

AI Usage Disclosure: As part of our commitment to transparency, we use artificial intelligence (AI) tools to assist in various stages of our recruitment process, including resume screening, candidate matching, interview scheduling, and communications. These tools are designed to improve efficiency, reduce bias, and enhance candidate experience. All decisions regarding hiring are made by qualified human professionals, and we continuously monitor our AI systems to ensure fairness and compliance with applicable regulations.

Similar roles