Rightsline is hiring a DevOps Manager, based in Toronto, to lead the team that owns our cloud platform, delivery pipelines, and production reliability across North America and Europe. You will manage and develop a team of five - a Senior DevOps Engineer, a DevOps Engineer, and two Senior Systems Engineers - and hold end-to-end accountability for the infrastructure our product runs on. You will report to the Senior DevOps Architect and partner closely with Engineering, QA, Security and Compliance, and Support.
What You Will OwnTeam Leadership & Accountability- Manage, coach, and develop a team of four infrastructure engineers, including senior individual contributors - people with deep expertise you will lead through judgment and clarity rather than positional authority.
- Establish unambiguous ownership: every system, environment, pipeline, and recurring operational duty has a named owner and a documented backup.
- Run the performance cycle - goals, regular one-on-ones, written feedback, and development plans calibrated against our engineering level framework - and address both strong performance and gaps directly and fairly.
- Hold the team to its commitments, and build the visibility that makes commitments legible: a single prioritized queue of work, current ticket hygiene, and clear closure on action items.
- Recruit, onboard, and retain, building enough bench depth that no system depends on one person being available.
Operational Rigor & Change Management- Own the change management process for infrastructure and environments end to end: proposal, peer review, scheduled execution, verification, and a durable record of what changed and why.
- Define and hold the line on what "done" means for environment work - the change applied consistently across every environment and region it belongs in, code and state in agreement, monitoring updated, documentation updated, and the ticket closed.
- Drive infrastructure work through version control and code review rather than console changes, and systematically close configuration drift between environments and regions.
- Make documentation a deliverable: architecture diagrams, runbooks, escalation paths, and environment inventories that an on-call engineer can rely on at three in the morning.
- Own incident practice - on-call rotation, severity definitions, blameless postmortems, and remediation items that are tracked to closure.
Cloud Platform & Delivery- Own our AWS accounts outright - account structure and organization, IAM and access boundaries, service quotas, networking, billing and tagging, and the guardrails that keep each account consistent with the next.
- Own everything running in those accounts across both regions - containerized services, load balancing, relational databases, search, messaging, and serverless components - including capacity, performance, patching, and security posture.
- Own the server estate on both Linux and Windows: build standards, configuration management, patching cadence, hardening, and lifecycle, so neither platform becomes the one nobody wants to touch.
- Own infrastructure as code in Terraform: module design, environment composition, state management, and the review standards that keep it maintainable as the fleet grows.
- Own the CI/CD estate - a templated pipeline-per-component-per-environment fleet across twelve environments and two regions - along with the release promotion model and production approval gates.
- Own observability and log management so that alerts are actionable, noise is low, and the data needed to investigate an issue is still there when the question gets asked.
- Partner with Engineering to reduce deployment risk and shorten lead time to production, and with QA on environment availability, parity, and refresh.
Reliability, Disaster Recovery & Compliance- Own availability and performance objectives for production in both regions, and the monitoring that demonstrates they are being met.
- Run our recurring disaster recovery test cycle end to end - including evidence against our 15-minute RPO and 4-hour partial / 24-hour full RTO objectives - and drive every finding to closure before the next cycle.
- Produce infrastructure evidence for SOC 2 and PCI DSS as a by-product of how the team already works: access reviews, change records, backup and restore verification, and vulnerability management.
- Partner with Security on hardening, secrets management, remediation timelines, and audit readiness, and keep our EU data residency commitments intact.
Stakeholder & Cost Management- Act as the escalation point for infrastructure issues affecting customers, communicating status credibly to engineering leadership and, when needed, to customers.
- Own cloud spend and infrastructure tooling costs, with a clear view of cost per environment and a plan for where it should go.
- Coordinate effectively with engineering colleagues in the United States and India, including planning work and releases across time zones.
What You Will Bring- 7+ years in DevOps, SRE, or infrastructure engineering, including 2+ years directly managing a DevOps or infrastructure team, with hiring, performance, and development responsibility. Management experience specifically in this space is required - leading infrastructure engineers is a different job from leading application developers.
- Experience leading senior engineers and architects - specialists who know their domains better than you do - and a track record of earning their trust.
- A demonstrated record of introducing process rigor to a technically strong team without slowing it down: change management, documentation standards, and on-call and incident practice that the team genuinely adopted and sustained.
- Comfort holding people accountable - setting expectations explicitly, following up consistently, and having the direct conversation early rather than late.
- Deep hands-on AWS experience, including account-level ownership: organizations and account structure, IAM, networking and load balancing, container orchestration, relational databases, and multi-region architecture.
- Working fluency across both Linux and Windows Server - you can hold a credible technical conversation about either, and you understand what it takes to keep a mixed estate patched, hardened, and consistent.
- Terraform at scale - modules, environment composition, state management, and the review standards that keep infrastructure code healthy.
- Ownership of production CI/CD (AWS CodePipeline and CodeBuild, Spinnaker, GitHub Actions, or comparable), including release promotion and approval gates.
- Experience operating under an audit framework such as SOC 2, PCI DSS, or ISO 27001, and owning disaster recovery testing rather than just participating in it.
- Strong written communication - you document decisions, and your updates are clear to people who were not in the room.
- Bachelor's degree in Computer Science or a related field, or equivalent practical experience.
Preferred Qualifications:- Multi-tenant SaaS at scale, particularly with EU data residency requirements.
- Depth in observability tooling such as New Relic, Sumo Logic, Datadog, or Grafana, including cost-aware log retention design.
- Search infrastructure (Solr or Elasticsearch) and .NET application hosting.
- Windows-side depth - Active Directory, IIS, Group Policy - alongside Linux administration.
- FinOps or cloud cost optimization experience with measurable results.
- Kubernetes, and infrastructure-code quality tooling such as Terratest, Checkov, or policy-as-code.
- AWS certification (Solutions Architect Professional or DevOps Engineer Professional).
- Experience integrating AI-assisted development tooling into delivery pipelines and the guardrails that requires.
Working Arrangements- Based in Toronto, Canada, in a hybrid arrangement - our Toronto office is available to you, with the flexibility to work from home.
- Must be legally authorized to work in Canada.
- Regular overlap with engineering colleagues in the United States and India.
- Participation in escalation coverage, and occasional planned off-hours work for releases, migrations, and disaster recovery testing.