Aviva

Senior Manager, Site Reliability & Infrastructure Engineering

Aviva$125K — $175K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 10+ years of technology experience in Site Reliability Engineering, Cloud, Infrastructure Engineering, or related fields.
  • 5+ years of experience leading and managing technical teams.
  • Experience with enterprise observability tools like Dynatrace or Grafana.
  • Hands-on experience with Rubrik or similar enterprise backup and recovery platforms.
  • Knowledge of AWS or comparable cloud services and their operational models.
  • Familiarity with disaster recovery and cyber recovery frameworks.
  • Bachelor's degree in Computer Science, Engineering, or related fields.

Responsibilities

  • Establish and mature Site Reliability Engineering practices across applications and platforms.
  • Partner with teams to incorporate reliability requirements into all development stages.
  • Drive improvements in incident response and problem management processes.
  • Enhance Infrastructure Operations through automation and self-healing techniques.
  • Leverage observability insights to boost infrastructure performance and availability.
  • Maintain and evolve technology resilience capabilities across various platforms.
  • Lead and coach engineering teams in reliability and recovery practices.

Benefits

  • Comprehensive rewards package including annual bonus and retirement savings plan.
  • Provision for professional development and career advancement opportunities.
  • Competitive vacation policy with the option to buy additional days off.
  • Wellness programs to support physical and mental health.
  • Hybrid work model providing flexibility in work arrangements.
Full Job Description
The Senior Manager, Site Reliability & Infrastructure Engineering will lead Aviva Canada's evolution toward an engineering-led reliability and resilience capability across critical applications, platforms and services. This role will work across on-premises, AWS, SaaS/vendor-hosted and hybrid application environments, partnering with Engineering Operations, Cloud & Platform Engineering, Application Engineering, Cybersecurity, Operational Resilience, Business Continuity, Enterprise Architecture, Vendor Management and key partners including AWS, DXC, CGI, Snowflake and other SaaS providers.

The focus is to shift Aviva from periodic recovery and resilience testing to proactive, measurable reliability engineering. This uses observability platforms like Dynatrace, incident response tools such as PagerDuty or equivalent, and modern recovery platforms like Rubrik to improve service availability, operational insight, recovery readiness, and customer outcomes.

What you'll do

Site Reliability Engineering
  • Establish and mature SRE practices across critical applications and platforms, including SLIs, SLOs, SLAs, service health indicators, post-incident reviews and continuous reliability improvement.
  • Partner with application, platform, infrastructure and vendor teams to embed reliability requirements into design, development, operational acceptance, release and production support processes.
  • Drive improvements in incident response, problem management, root cause analysis, alert quality, critical issue workflows, service health reporting and reduction of repeat incidents.
  • Improve Infrastructure Ops via automation, self-healing patterns, runbook improvements, AI-assisted operations and repeatable engineering practices.
  • Use observability insights to improve infrastructure availability, performance, capacity planning, & operational readiness.


Technology Resilience, Backup & Cyber Recovery
  • Maintain and evolve Aviva Canada's technology resilience capability across on-premises platforms, AWS, critical applications, integration services and SaaS/vendor-hosted platforms.
  • Ensure resilience planning validates end-to-end recoverability, including application, data, integration, identity, network, platform, cloud and vendor dependencies.
  • Support annual BCP, DR and technology resilience testing, including RTO/RPO validation, dependency mapping, recovery sequencing, test evidence, gap management and remediation tracking.
  • Own and mature backup and cyber recovery practices, including hands-on use of Rubrik or a comparable enterprise data protection platform for backup policy management, immutability, restore validation, coordinating recovery procedures, reporting, evidence capture and operational support for critical workloads.
  • Partner with Cybersecurity on ransomware and destructive cyber event readiness, including clean recovery, backup integrity validation, isolated recovery environments, cyber recovery vault operations and restoration of critical services.
  • Ensure resilience and recovery requirements are embedded into architecture, cloud migration, vendor onboarding, operational readiness, change delivery and service governance.


Leadership
  • Lead, manage and coach group of engineers working across reliability, observability, production support, technology resilience and recovery practices.
  • Build a clear operating model for SRE and technology resilience ownership across application teams, infrastructure teams, cloud/platform teams, cybersecurity, operational resilience and third-party providers.
  • Define reliability and resilience reporting for senior leaders, including service health, SLO performance, incident trends, MTTR, alert quality, recovery readiness, open risks and remediation status.
  • Partner with Operational Resilience, Business Continuity, Technology Risk and Cybersecurity teams to ensure evidence, controls and remediation plans are audit-ready and aligned to regulatory expectations.
  • Influence engineering and operations teams to adopt reliability-by-design, automation-first and evidence-based ways of working.
  • Serve as a designated point of contact for material risks relating to service reliability, monitoring gaps, application recoverability, backup coverage, cyber recovery readiness, RTO/RPO gaps and vendor resilience ambiguity.
  • Partner with key infrastructure managed service providers to continuously improve service quality, strengthen operational performance, hold vendors accountable for meeting agreed service levels, outcomes, remediation commitments and continuous improvement targets.


What you'll bring

  • 10+ years of technology experience across Site Reliability Engineering, Infrastructure Engineering, Platform Engineering, Cloud, Technology resilience, Disaster recovery, Cyber recovery or related disciplines.
  • 5+ years of experience leading & managing technical teams, with the ability to mentor engineers, set direction, manage priorities and high visible projects.
  • Experience with enterprise observability tooling such as Dynatrace, Grafana, Datadog or equivalent platforms.
  • Hands-on experience operating Rubrik or a comparable enterprise backup and cyber recovery platform, including backup policy configuration, immutable backup concepts, restore testing, recovery workflow execution, access controls, reporting, evidence capture and support for critical workloads.
  • Experience supporting distributed, business-critical applications across hybrid environments, including on-premises infrastructure, cloud platforms, APIs, middleware, databases, containers and SaaS/vendor-hosted systems.
  • Working knowledge of AWS or comparable cloud platforms, including cloud resilience, infrastructure as code, monitoring, logging, backup/recovery, identity, network dependencies and landing zone operating models.
  • Demonstrated experience with incident, problem, change, and release management practices across complex and/or highly regulated environments.
  • Knowledge of cyber recovery concepts, including ransomware recovery, clean-room recovery, isolated recovery environments, backup integrity validation and cyber incident recovery planning.
  • Strong written and verbal communication skills, including demonstrating proficiency in producing clear executive reporting, operational dashboards and audit-ready evidence.
  • Bachelor's degree in Computer Science, Engineering, Information Systems or equivalent.
  • P&C insurance domain experience, Guidewire/Snowflake exposure, AWS/Rubrik/Dynatrace certifications would be considered assets.


What you'll get

  • The salary band for this position ranges from $125,000 - $175,000. Please note that individual salary is determined by factors such as job-related knowledge, skills and experience, as well as internal equity.
  • Compelling rewards package including base compensation, eligibility for annual bonus, retirement savings, share plan, health benefits, personal wellness, and volunteer opportunities.
  • Outstanding Career Development opportunities.
  • We'll support your professional development education.
  • Competitive vacation package with the option to purchase 5 extra days off per year.
  • Employee driven programs focused on gender, LGBTQ+, origins, diversity, and inclusion.
  • Corporate wellness programs to support our employees' physical and mental health.
  • Employee discount on home and auto insurance (where applicable).
  • Hybrid flexible work model.


This job advertisement is for a new vacancy which has been posted both internally & externally.

Aviva Canada may use AI (Artificial Intelligence) tools to assist us throughout the recruitment process to screen, assess or select applicants for a position.

#LI-PS1
#LI-Hybrid

About Aviva

Natural health retailer and resource. Aviva promotes a holistic approach to your health and well-being through proper nutrition and a healthy, natural lifestyle. Aviva operates a retail store in Winnipeg, MB Canada and serves customers across Canada and around the world through www.avivahealth.com. Everything for healthy living.

Aviva Careers

Joining Aviva presents a prime opportunity to be part of a team that values leadership, innovation, and diversity in the workplace. Aviva, a leading insurance and financial services company, offers a variety of job opportunities that cater to a range of skills and professional interests, making it an ideal place for career growth and development.

Explore Job Opportunities

Aviva is actively hiring and continuously seeks passionate, creative, and solution-driven team players. Search open positions that match individual skills and interests in areas ranging from finance to technology. Each position at Aviva offers the chance to work with a global team that drives industry-leading solutions.

Experience Professional Growth

Aviva is committed to the professional growth of its employees. The company offers unmatched training, development, and certification support, allowing every team member to reach their full potential. Leadership and diversity training are integral parts of Aviva's commitment to fostering an inclusive environment.

Internship Programs

Start a career at Aviva with an internship that offers real-world experience and networking opportunities. Aviva's internships provide a robust platform for learning and applying practical skills in a dynamic setting, preparing interns for full-time positions within the company.

Benefits and Culture

Aviva is renowned for its employee benefits and positive company culture. The company prioritizes the well-being and satisfaction of its team, offering comprehensive benefits that enhance both personal and professional life. Aviva's culture is built on a foundation of respect for diversity and an ongoing commitment to inclusion.

Innovation and Networking

At Aviva, innovation is at the core of all operations. Employees are encouraged to lead transformation projects and introduce new strategies that drive growth and efficiency. Networking within Aviva provides connections to industry leaders and peers who are equally eager to share insights and collaborate on projects.

Career Advancement

Aviva believes in nurturing talent and providing employees with opportunities for advancement. Whether through promotions, cross-departmental moves, or leadership roles, team members have numerous paths to elevate their careers.

Join the Aviva Team

Interested candidates are encouraged to apply. Prepare a resume, practice for the interview, and explore the exciting employment and career opportunities at Aviva. Personalize job alerts to stay informed about the latest news and positions that align with career preferences and skills.

Stay Connected with Aviva Careers

Keep up to date with career tips, insider perspectives, and industry-leading insights—all from the professionals who are part of Aviva. Explore what it means to work at a company where innovation, leadership, and diversity are valued above all.

READ CAREERS BLOG

Job Alert Emails

Subscribe to receive job alerts, latest news, and insider tips tailored to preferences. Discover the rewarding opportunities that await at Aviva, a company committed to the professional and personal growth of its employees.
Learn more about Aviva
Size
28,596 employees
Industry
NASDAQ

Similar Jobs

More Jobs at Aviva

More Information Technology Jobs

Find similar Senior Manager, Site Reliability & Infrastructure Engineering jobs: