OpenText

Sr. Site Reliability Administrator

OpenText • $92K — $138K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of hands-on experience in Linux systems administration and troubleshooting complex production environments.
  • Proficient in scripting languages such as Shell, Python, Perl, or JavaScript.
  • Experience with major cloud platforms like AWS, Azure, or Google Cloud.
  • Deep knowledge of modern DevOps tools and technologies, including Kubernetes, Docker, and CI/CD pipelines.
  • Strong understanding of microservices architecture and API management practices.
  • Familiarity with middleware technologies like Apache and Tomcat, as well as enterprise applications.
  • Expertise in observability tools for application performance monitoring and logging.

Responsibilities

  • Collaborate with cross-functional teams to define operational requirements and improve service readiness.
  • Design proactive monitoring and alerting systems to enhance reliability and reduce incidents.
  • Provide advanced troubleshooting support while adhering to established SLAs.
  • Drive incident resolution processes, including root cause analysis and preventive measures.
  • Maintain operational documentation and best practices for production environments.
  • Develop real-time monitoring dashboards that track critical business transactions.
  • Lead performance optimizations and ensure compliance with operational standards.

Benefits

  • Comprehensive benefits package supporting physical, emotional, and financial wellbeing.
  • Flexible vacation entitlement and paid time off policies.
  • Opportunities for continuous learning and professional development.
Full Job Description
YOUR IMPACT

The role Cloud Application Engineer/Site Reliability Engineer is to build solutions to enhance availability, performance, and stability of OpenText services as well as automating away repetitive work as part of a cloud dev ops organization. This role would be a great fit for someone with creative and innovative problem-solving skills. You will develop and implement solutions that operate at scale. Our teams are empowered and expected to improve our products to truly deliver a reliable experience to customers.

WHAT THE ROLE OFFERS
  • Collaborate with Agile squads, developers, business partners, and sustain teams to define technical requirements and enhance operational readiness through effective logging, monitoring, and metrics solutions.
  • Design and implement proactive monitoring, alerting, and observability capabilities to reduce incidents and improve system reliability.
  • Provide advanced production support, troubleshooting, and incident management while meeting established Service Level Agreements (SLAs).
  • Take ownership of the incident resolution process, including root cause analysis (RCA), SWAT investigations, and preventive action planning.
  • Partner with development teams to drive system stability, defect remediation, production readiness, and smooth transitions into sustainment.
  • Develop, maintain, and execute operational runbooks, support procedures, and best-practice patterns for production environments.
  • Work with business and IT stakeholders to create real-time monitoring, alerting, and KPI dashboards based on business transaction tracking.
  • Lead performance analysis and provide technical guidance on system optimization, issue resolution, and application reliability improvements.
  • Collaborate with application owners and cross-functional teams to mitigate risks, remediate audit findings, validate deployments, and ensure operational compliance.
  • Support a 24x7x365 environment through on-call rotations, shift coverage, knowledge-sharing initiatives, team backup responsibilities, and continuous training activities.

WHAT YOU NEED TO SUCCEED
  • Strong expertise in Linux systems administration, scripting, and troubleshooting complex production environments using languages such as Shell, Python, Perl, or JavaScript.
  • Hands-on experience with cloud platforms (AWS, Azure, or Google Cloud) and modern platform technologies including Kubernetes, Cloud Foundry, BOSH, Docker, and other containerization solutions.
  • Solid understanding of microservices architecture, RESTful APIs, API gateways (e.g., Apigee), and authentication standards such as OAuth 2.0.
  • Experience designing and maintaining CI/CD and automation pipelines using tools such as Ansible, Rundeck, Argo CD, or similar DevOps technologies.
  • Strong knowledge of middleware and application platforms including Apache, Tomcat, Spring Framework, and Java-based enterprise applications.
  • Experience supporting distributed systems, high-volume web applications, message brokers (Kafka, RabbitMQ), search platforms (Elasticsearch, Solr), and both relational and NoSQL databases.
  • Deep expertise in observability, application performance monitoring, centralized logging, and monitoring tools such as Dynatrace, New Relic, AppDynamics, Zabbix, Checkmk, Graylog, and Kibana.
  • Proven ability to diagnose, troubleshoot, and resolve complex application, infrastructure, and network issues while applying security and operational best practices.
  • Demonstrated leadership in driving scalable technical solutions, managing competing priorities, and working effectively both independently and within cross-functional teams.
  • Strong analytical, organizational, and problem-solving skills, with a passion for understanding system internals, improving reliability, and supporting ITIL-based operational excellence.

Compensation: At OpenText, we offer a thoughtfully designed benefits package that supports your physical, emotional, and financial wellbeing. As you move through the hiring process, we're happy to provide more details about our compensation programs, including variable and commission compensation opportunities for eligible roles, vacation entitlement, and paid time off.

Salary Range: $92,320 -138,480; Depending on the candidate's education, experience, skills, geographical location, and alignment with internal equity and external market, actual salary may vary and be higher or lower than the range posted.

About OpenText

OpenText Corporation is a Canadian company that develops and sells enterprise information management software. The company is headquartered in Waterloo, Ontario, Canada. OpenText software applications manage content or unstructured data for large companies, government agencies, and professional service firms. OpenText aims to help customers digitize their operations and optimize their information management processes. The company has made several acquisitions over the years to expand its product offerings and customer base. OpenText has a global presence with offices in North America, Europe, Asia, and Australia.
Learn more about OpenText
Size
14,300 employees
Market Cap
$7.8 billion
Industry
Founded
1991
5 Year Trend
+8.8%
NASDAQ

Similar Jobs

More Jobs at OpenText

More Information Technology Jobs

Find similar Sr. Site Reliability Administrator jobs: