OpenText

Sr. Site Reliability Administrator

OpenText • $93K — $138K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field, or equivalent experience.
  • 4+ years in IT supporting large-scale enterprise systems.
  • 2+ years operating or supporting distributed data platforms like Kafka, Elasticsearch, and others.
  • 2+ years with automation tools like Terraform and Ansible.
  • Strong Linux systems administration knowledge.
  • Experience with public cloud infrastructures such as AWS, Azure, or GCP.
  • Excellent troubleshooting skills for resolving complex technical issues.
  • Outstanding written and verbal communication skills.

Responsibilities

  • Operate, maintain, and scale distributed data services such as Kafka and Elasticsearch.
  • Build and enhance infrastructure across on-premises and public cloud.
  • Develop and maintain Infrastructure-as-Code (IaC) using Terraform and Ansible.
  • Apply patches and ensure compliance and security of systems.
  • Participate in design, deployment, and monitoring of data platforms.
  • Support incident response and participate in on-call rotations.
  • Assist with capacity planning, performance tuning, and health assessments.

Benefits

  • Comprehensive benefits package that supports physical, emotional, and financial wellbeing.
  • Opportunities for learning and professional development through knowledge-sharing and training activities.
Full Job Description
YOUR IMPACT

As a Senior Site Reliability Engineer, you will play a key hands-on role within a globally distributed SRE team responsible for the reliability, performance, and stability of the data services that power our customer-facing SaaS products. You will work on well-defined components of our distributed systems stack - such as Kafka, Elasticsearch, Cassandra, Solr, Redis, and OpenSearch - across both on-premises and public cloud environments (AWS, Azure, GCP). This role is ideal for engineers who enjoy solving challenging operational problems, contributing to automation and reliability improvements, and collaborating across teams to support high-quality services. You'll deepen your expertise in distributed systems while helping the team execute on operational excellence.

WHAT THE ROLE OFFERS
  • Operate, maintain, and scale distributed data services including Kafka, Elasticsearch, Cassandra, Solr, Redis, and OpenSearch.
  • Build, enhance, and support infrastructure across on-prem and public cloud environments (AWS, Azure, GCP).
  • Develop and maintain Infrastructure-as-Code (IaC) using Terraform and Ansible.
  • Apply patches, perform routine maintenance, and ensure systems remain secure and compliant with internal standards.
  • Participate in the design, deployment, and monitoring of data platforms in collaboration with SRE and engineering teams.
  • Support incident response activities and participate in the on-call rotation for critical services.
  • Assist with capacity planning, performance tuning, and health assessments of data services.
  • Create and maintain documentation, including operational procedures, change plans, and incident reports.
  • Contribute to automation and reliability initiatives that improve service performance and reduce manual work.
  • Support service requests and help ensure SLA/OLA commitments are met.
  • Participate in team knowledge-sharing and training activities.
  • May require shift work and participation in a 24x7 on-call rotation.

WHAT YOU NEED TO SUCCEED
  • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field - or equivalent practical experience.
  • 4+ years of experience in Information Technology supporting large-scale enterprise systems.
  • 2+ years operating or supporting distributed data platforms (e.g., Kafka, Elasticsearch, Cassandra, Solr, Redis, OpenSearch).
  • 2+ years working with automation and configuration tools such as Terraform and Ansible.
  • Strong knowledge of Linux systems administration.
  • Experience working with public cloud infrastructure (AWS, Azure, or GCP).
  • Solid troubleshooting skills and ability to resolve complex technical problems.
  • Excellent written and verbal communication skills.
  • Self-driven, detail-oriented, and able to manage multiple tasks in a fast-moving environment.
  • Familiarity with ITIL processes; certification is a plus.
  • Experience with observability tools (Prometheus, Zabbix, Grafana, New Relic, etc.) is a plus.

Compensation: At OpenText, we offer a thoughtfully designed benefits package that supports your physical, emotional, and financial wellbeing. As you move through the hiring process, we're happy to provide more details about our compensation programs, including variable and commission compensation opportunities for eligible roles, vacation entitlement, and paid time off.

Salary Range:$93,320-$138,480; Depending on the candidate's education, experience, skills, geographical location, and alignment with internal equity and external market, actual salary may vary and be higher or lower than the range posted.

About OpenText

OpenText Corporation is a Canadian company that develops and sells enterprise information management software. The company is headquartered in Waterloo, Ontario, Canada. OpenText software applications manage content or unstructured data for large companies, government agencies, and professional service firms. OpenText aims to help customers digitize their operations and optimize their information management processes. The company has made several acquisitions over the years to expand its product offerings and customer base. OpenText has a global presence with offices in North America, Europe, Asia, and Australia.
Learn more about OpenText
Size
14,300 employees
Market Cap
$7.8 billion
Industry
Founded
1991
5 Year Trend
+8.8%
NASDAQ

Similar Jobs

More Jobs at OpenText

More Information Technology Jobs

Find similar Sr. Site Reliability Administrator jobs: