Four Seasons Hotels & Resorts

Senior DevOps Engineer - Digital Platforms

Four Seasons Hotels & Resorts • $90K — $125K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • University degree or college diploma in Computer Science, Software Engineering, or related discipline; relevant experience considered.
  • 7+ years in DevOps, Site Reliability Engineering, or related fields with demonstrated technical leadership.
  • Expert knowledge in DevOps, SRE, platform engineering, and software delivery practices.
  • Advanced experience with Azure and/or AWS infrastructure and managed services.
  • Strong Infrastructure as Code expertise with Terraform or CloudFormation.
  • Advanced CI/CD experience with platforms like Azure DevOps or Jenkins.
  • Familiar with AI-assisted engineering and safe usage of coding assistants.

Responsibilities

  • Serve as the senior technical authority for DevOps and reliability engineering, influencing architecture and operational decisions.
  • Define and evolve a multi-year DevOps and SRE roadmap aligned with business and technology priorities.
  • Establish reusable engineering patterns and standards across various digital properties.
  • Mentor engineers and build organizational capability through documentation and knowledge sharing.
  • Design and optimize CI/CD pipelines incorporating quality, security, and performance checks.
  • Lead performance engineering and capacity planning for global digital platforms, addressing systemic issues.
  • Embed security, privacy, and accessibility requirements into operational practices.

Benefits

  • Flexible work environment fostering innovation and collaboration.
  • Access to continuous learning and professional development opportunities.
  • Participation in a global team with exposure to multinational projects.
  • Diverse and inclusive workplace culture promoting respect and empowerment.
Full Job Description
Senior DevOps Engineer - Digital Platforms

The Senior DevOps Engineer - Digital Platforms is a senior, hands-on technical role accountable for advancing the reliability, scalability, security, performance, and operational maturity of Four Seasons' guest-facing digital platforms. The role supports a diverse portfolio that includes FourSeasons.com, the Four Seasons mobile app, the corporate website, Press Room, and the cloud services and integrations that enable these experiences.

This individual provides deep technical leadership, advances DevOps and Site Reliability Engineering practices, solves complex cross-platform problems, and establishes reusable capabilities that improve deployment safety, observability, resilience, recovery, and engineering efficiency. They influence architecture and delivery decisions in partnership with internal and external technology teams.

The successful candidate combines strong cloud and automation expertise with a practical understanding of high-traffic, global digital experiences and the importance of protecting guest trust, revenue, brand reputation, and operational continuity. They demonstrate advanced AI aptitude and a strategic mindset, using AI to challenge legacy processes, improve engineering effectiveness, enable responsible adoption, and deliver measurable business and operational outcomes.

What You'll Be Doing:

Technical Leadership & DevOps Strategy
  • Serve as the senior technical authority for DevOps and reliability engineering across front-end digital platforms, influencing architecture and operational decisions across teams to ensure balanced reliability, delivery speed, security, maintainability, cost, and guest experience.
  • Define and evolve a multi-year DevOps and SRE roadmap aligned with platform modernization, business priorities, the digital product roadmap, and enterprise technology standards.
  • Establish reusable engineering patterns, guardrails, and standards across web, mobile, eCommerce, and related digital properties.
  • Mentor engineers and analysts and build organizational capability through documentation, coaching, knowledge sharing, and communities of practice.

Cloud Infrastructure, Platform Engineering & Automation
  • Participate in the design, implementation, and continuous improvement of our cloud infrastructure and shared platform capabilities across Azure and AWS environments supporting digital experiences.
  • Build and maintain Infrastructure as Code using Terraform, CloudFormation, or comparable tools, with version control, peer review, automated validation, and repeatable environment provisioning.
  • Engineer secure, scalable patterns for containers, serverless services, networking, content delivery, configuration, and identity management.
  • Automate recurring operational tasks, environment configuration, release validation, recovery procedures, and platform lifecycle activities to reduce manual work and operational risk.
  • Partner with architecture, security, infrastructure, and application teams to ensure platform designs meet enterprise standards and remain supportable throughout their lifecycle.
  • Actively develops team capability by mentoring colleagues, sharing expertise, promoting engineering best practices, and fostering a culture of collaboration, learning, and continuous improvement.

CI/CD, Release Engineering & Developer Experience
  • Design and optimize CI/CD pipelines for web, mobile, API, CMS, and supporting services using modern source control, build, test, security scanning, artifact management, deployment, and rollback practices.
  • Implement deployment patterns such as blue-green, canary, feature flags, progressive delivery, and automated rollback where appropriate to reduce guest and business impact.
  • Embed automated quality, security, accessibility, performance, and compliance checks into delivery pipelines in partnership with QA, Security, Privacy, and Accessibility teams.
  • Improve developer experience by creating self-service tooling, standard pipeline templates, environment automation, and clear engineering documentation.
  • Use delivery metrics and failure patterns to improve deployment frequency, lead time, change failure rate, recovery time, and release predictability.
  • Leverage AI-enabled tools and automation appropriately to improve the quality, efficiency and effectiveness of work, while validating outputs and remaining accountable for results.
  • Lead responsible adoption by setting expectations, enabling experimentation, building capability and ensuring appropriate human oversight.

Observability, Guest Journey Monitoring & Reliability
  • Define and implement end-to-end observability across applications, cloud infrastructure, APIs, integrations, content delivery networks, third-party services, and critical guest and partner journeys.
  • Build actionable dashboards and alerts using technologies such as Elastic/Kibana, Azure Monitor, Application Insights, CloudWatch, OpenTelemetry, BrowserStack, and comparable tools.
  • Establish service-level indicators, service-level objectives, error budgets, health checks, synthetic monitoring, real user monitoring, and reliability reporting for critical services and journeys.
  • Correlate technical telemetry, digital analytics, and business signals to identify emerging issues and reduce detection and recovery time.
  • Apply anomaly detection, intelligent alerting, and AI-assisted operational analysis where appropriate, with appropriate validation, security, and human review.
  • Measures business benefit to ensure the application of AI results in net positive enterprise outcomes.

Performance, Scalability & Resilience Engineering
  • Lead performance engineering and capacity planning for global, high-traffic digital platforms, including page speed, API latency, mobile performance, CDN behaviour, caching, throughput, and infrastructure utilization.
  • Partner with engineering and QA teams to define performance budgets, load-testing strategies, resilience tests, and launch criteria for major releases and platform changes.
  • Identify systemic bottlenecks and single points of failure, implementing and validating improvements, including scalability, redundancy and business-continuity capabilities, that enable critical services to scale for peak demand and recover predictably from failures.
  • Balance performance, availability, and cloud cost through evidence-based optimization and capacity decisions.

Production Operations, Incident & Problem Management
  • Provide senior technical leadership during complex production incidents affecting guest, hotel, corporate, or partner experiences, coordinating diagnosis and recovery across internal teams and vendors.
  • undefined
  • Develop and improve runbooks, escalation paths, support procedures, on-call readiness, vendor handoffs, and operational acceptance criteria.
  • undefined
  • Participate in after-hours support for critical incidents and major releases when required.

Security, Privacy, Accessibility & Compliance
  • Embed security, privacy, PCI DSS, accessibility, vulnerability-management, and auditability requirements into infrastructure, CI/CD, monitoring, and operational practices.
  • Partner with Security and engineering teams on vulnerability remediation, secure configuration, patching, secrets management, certificate management, access controls, and supply-chain security.
  • Support WCAG and AODA-aligned digital delivery by enabling automated accessibility testing and reliable governance evidence within release processes.
  • Maintain the controls, documentation, and evidence required to support global regulatory obligations and Four Seasons architecture, security, change, accessibility, and compliance reviews.

Cross-Functional Delivery & Vendor Leadership
  • Collaborate across Digital Product, Brand, Content, eCommerce, Engineering, QA, Analytics, Architecture, Security, and enterprise platform teams.
  • Translate business priorities and guest experience needs into clear reliability, automation, performance, and operational requirements.
  • Provide technical oversight for external engineering, cloud, CDN, monitoring, and SaaS partners, including design review, delivery quality, operational readiness, and issue escalation.
  • Communicate complex technical risks, options, and recommendations clearly to technical and non-technical stakeholders, including senior leadership.
  • Support investment planning and technology evaluation through technical assessments, risk analysis, estimates, and measurable success criteria.


What You Bring:
  • University degree or college diploma in Computer Science, Software Engineering, Information Technology, or a related discipline. Equivalent relevant experience will be considered.
  • 7+ years of progressive experience in DevOps, Site Reliability Engineering, cloud engineering, platform engineering, or production operations. Demonstrated experience operating enterprise-scale, customer-facing web and mobile platforms in cloud-first environments, including technical leadership across multiple teams and complex delivery ecosystems.
  • Azure, AWS, Kubernetes, Terraform, DevOps, SRE, security, or IT service management certifications are assets.
  • Expert knowledge of DevOps, SRE, platform engineering, cloud operations, production readiness, and software delivery practices.
  • Advanced hands-on experience with Azure and/or AWS infrastructure and managed cloud services.
  • Strong Infrastructure as Code experience with Terraform, CloudFormation, or similar tooling.
  • Advanced CI/CD experience with Azure DevOps, GitHub Actions, Jenkins, GitLab CI, or comparable platforms.
  • Experience with container platforms, including Kubernetes and Docker, and associated configuration and secrets management.
  • Advanced observability skills spanning logs, metrics, traces, synthetics, real user monitoring, mobile telemetry, alerting, and dashboard engineering.
  • Experience with Elastic, Azure Monitor, CloudWatch, Open Telemetry, Prometheus, Grafana, or comparable tools.
  • Strong knowledge of global web and mobile architectures, CDNs, DNS, caching, APIs, microservices, CMS platforms, SaaS integrations, and identity flows.
  • Experience with performance engineering, load testing, capacity planning, resilience testing, disaster recovery, and cost optimization.
  • Scripting and automation skills using Python, PowerShell, Bash, TypeScript, or comparable languages.
  • Strong understanding of secure software delivery, vulnerability management, PCI DSS, privacy, WCAG/AODA accessibility, and change governance.
  • Working knowledge of AI-assisted engineering and operations, including safe use of coding assistants, anomaly detection, incident analysis, and automated remediation.
  • Effectively manages multiple initiatives and competing priorities, organizes complex work into achievable outcomes, flags risks and dependencies early, and consistently delivers high-quality results in fast-paced environment.


Who You Are:
  • Ownership and accountability: Takes responsibility for platform outcomes and follows complex issues through to durable resolution.
  • Communication: Explains complex technical issues, trade-offs, and risks clearly to technical and non-technical audiences.
  • Collaboration and influence: Builds trust across teams and aligns diverse stakeholders around practical engineering decisions.
  • Discernment: Balances reliability, speed, security, cost, compliance, and guest experience in ambiguous situations.
  • Continuous improvement: Uses evidence and feedback to improve engineering standards, operating practices, and team capability.
  • Global mindset: Designs and operates services for an international business with global users, partners, and time zones.
  • Demonstrates curiosity and continuous learning, adapting skills and ways of working as AI capabilities and organizational practices evolve.
  • Applies sound judgment, protects information and considers the impact of AI-enabled work on people and outcomes.
  • Demonstrates a commitment to continuous learning, proactively developing new technical and professional capabilities, remaining current with emerging technologies, industry trends, and applying new knowledge to improve business and operational outcomes.


Salary Range: $90,000

About Four Seasons Hotels & Resorts

Four Seasons Hotels and Resorts is a luxury hotel and resort company that was founded in 1960 by Isadore Sharp. The company is headquartered in Toronto, Canada, and operates more than 100 hotels and resorts in over 40 countries. Four Seasons is known for its high-end amenities and personalized service, and has been recognized with numerous awards and accolades. The company is privately held and has been owned by Cascade Investment, a company controlled by Bill Gates, since 2007.
Learn more about Four Seasons Hotels & Resorts
Size
45,000 employees
Industry
Founded
1993

Similar Jobs

More Jobs at Four Seasons Hotels & Resorts

More Information Technology Jobs

Find similar Senior DevOps Engineer - Digital Platforms jobs: