Sr. Site Reliability Administrator

OpenTextRichmond Hill, Waterloo, Mississauga, OntarioOn-siteFull-timeSenior, 5–8 yearsListed 2 hours ago

Apply now

About this role

OPENTEXT - THE INFORMATION COMPANY

OpenText is a global leader in information management, where innovation, creativity, and collaboration are the key components of our corporate culture. As a member of our team, you will have the opportunity to partner with the most highly regarded companies in the world, tackle complex issues, and contribute to projects that shape the future of digital transformation.

AI-First. Future-Driven. Human-Centered.

At OpenText, AI is at the heart of everything we do—powering innovation, transforming work, and empowering digital knowledge workers. We are hiring talent AI can't replace to help us shape the future of information management. Join us.

YOUR IMPACT

The role Cloud Application Engineer/Site Reliability Engineer is to build solutions to enhance availability, performance, and stability of OpenText services as well as automating away repetitive work as part of a cloud dev ops organization. This role would be a great fit for someone with creative and innovative problem-solving skills. You will develop and implement solutions that operate at scale. Our teams are empowered and expected to improve our products to truly deliver a reliable experience to customers.

WHAT THE ROLE OFFERS

- Collaborate with Agile squads, developers, business partners, and sustain teams to define technical requirements and enhance operational readiness through effective logging, monitoring, and metrics solutions.
- Design and implement proactive monitoring, alerting, and observability capabilities to reduce incidents and improve system reliability.
- Provide advanced production support, troubleshooting, and incident management while meeting established Service Level Agreements (SLAs).
- Take ownership of the incident resolution process, including root cause analysis (RCA), SWAT investigations, and preventive action planning.
- Partner with development teams to drive system stability, defect remediation, production readiness, and smooth transitions into sustainment.
- Develop, maintain, and execute operational runbooks, support procedures, and best-practice patterns for production environments.
- Work with business and IT stakeholders to create real-time monitoring, alerting, and KPI dashboards based on business transaction tracking.
- Lead performance analysis and provide technical guidance on system optimization, issue resolution, and application reliability improvements.
- Collaborate with application owners and cross-functional teams to mitigate risks, remediate audit findings, validate deployments, and ensure operational compliance.
- Support a 24x7x365 environment through on-call rotations, shift coverage, knowledge-sharing initiatives, team backup responsibilities, and continuous training activities.

WHAT YOU NEED TO SUCCEED

- Strong expertise in Linux systems administration, scripting, and troubleshooting complex production environments using languages such as Shell, Python, Perl, or JavaScript.
- Hands-on experience with cloud platforms (AWS, Azure, or Google Cloud) and modern platform technologies including Kubernetes, Cloud Foundry, BOSH, Docker, and other containerization solutions.
- Solid understanding of microservices architecture, RESTful APIs, API gateways (e.g., Apigee), and authentication standards such as OAuth 2.0.
- Experience designing and maintaining CI/CD and automation pipelines using tools such as Ansible, Rundeck, Argo CD, or similar DevOps technologies.
- Strong knowledge of middleware and application platforms including Apache, Tomcat, Spring Framework, and Java-based enterprise applications.
- Experience supporting distributed systems, high-volume web applications, message brokers (Kafka, RabbitMQ), search platforms (Elasticsearch, Solr), and both relational and NoSQL databases.
- Deep expertise in observability, application performance monitoring, centralized logging, and monitoring tools such as Dynatrace, New Relic, AppDynamics, Zabbix, Checkmk, Graylog, and Kibana.
- Proven ability to diagnose, troubleshoot, and resolve complex application, infrastructure, and network issues while applying security and operational best practices.
- Demonstrated leadership in driving scalable technical solutions, managing competing priorities, and working effectively both independently and within cross-functional teams.
- Strong analytical, organizational, and problem-solving skills, with a passion for understanding system internals, improving reliability, and supporting ITIL-based operational excellence.

One Last Thing 
OpenText is more than just a corporation, it's a global community where trust is foundational, the bar is raised, and outcomes are owned. 
Join us on our mission to drive positive change through privacy, technology, and collaboration. At OpenText, we don't just have a culture; we have character. Choose us because you want to be part of a company that embraces innovation and empowers its employees to make a difference.

OpenText's commitment to diversity and inclusion surpasses legal requirements, evident in our Equal Employment Opportunity Statement of Policy which promotes a respectful and empowering environment for employees of all backgrounds, culture, national origin, race, color, gender, gender identification, sexual orientation, family status, age, veteran status, disability, religion, or other basis protected by applicable laws.

If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please submit a ticket at Ask HR . Our proactive approach fosters collaboration, innovation, and personal growth, enriching OpenText's vibrant workplace.

Compensation: At OpenText, we offer a thoughtfully designed benefits package that supports your physical, emotional, and financial wellbeing. As you move through the hiring process, we’re happy to provide more details about our compensation programs, including variable and commission compensation opportunities for eligible roles, vacation entitlement, and paid time off.

Salary Range: $92,320 -138,480; Depending on the candidate’s education, experience, skills, geographical location, and alignment with internal equity and external market, actual salary may vary and be higher or lower than the range posted.

AI Usage Disclosure: As part of our commitment to transparency, we use artificial intelligence (AI) tools to assist in various stages of our recruitment process, including resume screening, candidate matching, interview scheduling, and communications. These tools are designed to improve efficiency, reduce bias, and enhance candidate experience. All decisions regarding hiring are made by qualified human professionals, and we continuously monitor our AI systems to ensure fairness and compliance with applicable regulations.