Job DescriptionWe are looking for an individual who exemplifies the attributes of a leader, mentor and decision-maker. We are currently building out our SRE team, with the goal being to provide expertise and tooling to manage the health, security, and availability of their applications in production. We work with other teams to provide guidance throughout the lifecycle of building, deploying, and operating the application.
What Will You Do? - Run the production environment by monitoring availability and taking a holistic view of system health
- Build tools to manage platform infrastructure and applications
- Debug production issues across services and levels of the stack and provide primary operational support and engineering for multiple large distributed software applications
- Help adopt and drive the tool creation for application health monitoring and alerting.
- Improve reliability, quality, and time-to-market of our suite of software solutions
- Measure and optimize system performance, with an eye toward pushing our capabilities forward, getting ahead of application team needs, and innovating to continually improve
- Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding.
- Participate in system design consulting, platform management, and capacity planning.
- Create sustainable systems and services through automation and uplifts.
- Balance feature development speed and reliability with well-defined service level objectives.
What Do You Need To Succeed? Must have: - Overall, 4-6 years of support experience in Openshift, Azure & Kubernetes.
- 4-5 years of experience as an SRE supporting multiple applications and very strong programming skills in Java/SpringBoot/Python.
- Oracle and SQL database operational experience in the cloud/on-premise and writing/understanding database queries (SQL and/or No-SQL) and Object Oriented design and development
- Exposure to OCP, and GitHub is desirable
- Having a good overall understanding of networking-related areas like certificates, load balancers etc.
- Monitoring using Splunk, Dynatrace, RUM, Grafana & other related tools
- Experience with the operational aspects of software systems such as monitoring, centralized logging, and alerting.
- Experience in micro-services, public cloud (Azure preferred) & container technologies and working knowledge of Mainframes & JCL is nice to have
Nice-to-have:- Knowledge of public cloud (Microsoft Azure and AWS) and private cloud (OpenShift) platforms and development of applications in multi-cloud, hybrid environments
- Knowledge of containers and orchestration (e.g: Docker, Kubernetes)
What's in it for you?We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.
A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
Leaders who support your development through coaching and managing opportunities
Ability to make a difference and lasting impact
Work in a dynamic, collaborative, progressive, and high-performing team
A world-class training program in financial services
Flexible work/life balance options
Opportunities to do challenging work
#LI
Job SkillsAgile Methodology, Group Problem Solving, IT Systems Integration, Organizational Leadership, Product Services, Software Development Life Cycle (SDLC), System Applications, System Integration Testing (SIT), Systems Software
Additional Job DetailsAddress:RBC WATERPARK PLACE, 88 QUEENS QUAY W:TORONTO
City:Toronto
Country:Canada
Work hours/week:37.5
Employment Type:Full time
Platform:TECHNOLOGY AND OPERATIONS
Job Type:Regular
Pay Type:Salaried
Posted Date:2026-09-10
Application Deadline:2026-10-10
Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above