What success looks like in this role: U.S. Citizenship and Residency RequiredPosition Summary Serves as a senior technical leader responsible for the strategy, architecture, design, implementation, and governance of enterprise Disaster Recovery (DR) capabilities across infrastructure, applications, cloud platforms, and data environments.
Partners with business and technology leadership to ensure resilience of mission-critical systems through development, testing, and execution of Disaster Recovery Plans (DRPs) aligned to recovery objectives (RTO/RPO). Provides leadership and coordination across internal teams and third-party providers to ensure rapid restoration of services during disaster events.
As leader of the Disaster Recovery service, this role owns end-to-end DR strategy and execution, including setup of the service and coordination of quarterly DR testing, which may require periodic travel.
Core Responsibilities
Architecture Strategy & Design- Defines enterprise Disaster Recovery architecture, standards, and frameworks aligned with ITIL/ITSM Service Continuity Management practices, building on responsibilities such as:
- "Develops resilient architectures across cloud, infrastructure, and applications, including failover, redundancy, and replication strategies."
- Translates business requirements and Business Impact Analysis (BIA) into DR technical solutions and architecture designs.
Disaster Recovery Planning & Governance - Leads development, maintenance, and governance of Disaster Recovery Plans (DRPs), including application-specific recovery strategies.
- Ensures DRPs align with defined Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO).
- Establishes governance processes to maintain DRP accuracy through change management and annual review cycles.
- Conducts gap analysis and recommends improvements to DR processes, technologies, and services.
Cross-Functional Leadership & Coordination - Provides leadership and oversight for Disaster Recovery activities across internal teams and third-party providers.
- Facilitates DR planning and execution coordination across business, infrastructure, and application teams.
- Acts as the primary technical authority for DR architecture and recovery strategy decisions.
Disaster Recovery Readiness - Defines infrastructure, network, and data protection requirements necessary for DR execution.
- Designs and governs backup, replication, and recovery strategies to ensure system recoverability and data integrity.
- Ensures DR environments are maintained in alignment with production configurations.
DR Service Operations - Reviews and onboards new or modified solutions into the Disaster Recovery service.
- Monitors and evaluates DR infrastructure capacity across on-premises and cloud environments.
- Identifies risks and communicates capacity or recovery concerns to stakeholders.
Testing & Validation - Leads planning, execution, and reporting of Disaster Recovery testing.
- Coordinates participation across internal teams and third-party providers during DR tests.
- Analyzes results, identifies gaps, and drives remediation actions.
Disaster Recovery Execution - Provides technical leadership during disaster events and ensures execution of DRPs.
- Coordinates failover to DR sites and restoration of mission-critical systems.
- Leads post-incident analysis and continuous improvement initiatives.
You will be successful in this role if you have:Education - Required: Bachelor's degree in Computer Science, Information Systems, or related field
- Preferred: Master's degree
Experience - 12+ years of experience in IT architecture, infrastructure, or enterprise systems
- Demonstrated experience designing and implementing enterprise Disaster Recovery solutions
- Experience with cloud platforms (Oracle OCI, Azure, hybrid environments)
This role may require access to export-controlled commodities and technology. Therefore, to conform to U.S. export control regulations, applicant should be eligible for any required authorizations from the U.S. Government.