Senior Program Manager, Disaster Recovery
Lead Twilio's enterprise Disaster Recovery program by developing plans, conducting tests and audits, integrating risks, and serving as the technical authority during live disruptions to ensure resilient recovery of critical systems. Requires CBCP certification, cloud infrastructure expertise, and 7+ years of DR program management experience.
About the job
Responsibilities
- Lead the annual plan development, review, and executive sign-off lifecycle for critical product and non-product systems.
- Actively integrate disaster recovery vulnerabilities, Service Readiness Framework (SRF) findings, and exercise gaps into existing risk frameworks to quantify, score, and systematically drive down operational risk across the company.
- Author clear program updates, annual compliance metrics, and formal responses to internal and external audit findings.
- Partner across the organization to support the execution of unified Gamedays and quarterly technical tests to ensure critical applications fail over predictably.
- Benchmark and audit actual technical execution against strict Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO).
- Convert static, document-heavy continuity playbooks into clean, runnable-style procedures to optimize recovery coordination.
- Review and audit the independent business continuity and disaster recovery plans of high-exposure third-party SaaS vendors to proactively isolate and eliminate downstream single points of failure (SPOFs).
- Verify fallback processing environments, ensuring that automated remediation workflows are fully provisioned and mapped before disruptions occur.
- Serve as the technical disaster recovery authority during active live disruptions, seamlessly synchronizing recovery activities with existing tactical response teams.
Qualifications
Required:
- Active Certified Business Continuity Professional (CBCP) designation.
- Practical familiarity with cloud infrastructure platforms (AWS console operations), containerization ecosystems, database synchronization architectures, and centralized configuration directories.
- Proven experience leading technical response or recovery workflows during high-severity live outages alongside incident management teams.
- Exceptional verbal and written communications skills; ability to distill highly dense technical single points of failure (SPOFs) into clear, actionable financial and engineering risk metrics for senior security leadership.
Desired:
- 7+ years of dedicated technical Disaster Recovery program management experience in a highly scalable, cloud-native enterprise environment.
Skills
Disaster Recovery, Business Continuity, Cbcp, AWS, Cloud Infrastructure, Containerization, Database Synchronization, Incident Management, Rto, Rpo, Risk Management, Gamedays, Audit
Similar jobs
Technical Program Management jobsLeads complex global enablement programs for Sales Engineering, translating business priorities into measurable roadmaps, training, and workflow change. Requires 6+ years of related experience, strong senior stakeholder influence, data and reporting expertise, and PROSCI/ADKAR or equivalent change-management certification.
Leads complex global enablement and change programs for Global Support Engineering, translating business priorities into measurable roadmaps, training, and workflow improvements. The role requires senior stakeholder influence, data analysis, executive communication, and PROSCI/ADKAR expertise.
Leads complex end-to-end enterprise SaaS deployments, managing scope, timelines, risks, integrations, training, and cross-functional stakeholders through launch and adoption. Requires 8+ years of project management experience, enterprise implementation expertise, and strong executive communication skills.
Leads complex global People Operations programs spanning geographic expansion, entity formation, M&A integrations, and operational infrastructure changes. Requires 7+ years of cross-functional program leadership, strong stakeholder management, executive communication, and practical AI fluency.
Coordinates day-to-day execution for engineering teams delivering U.S. government and space mission programs. The role manages tasking, dependencies, milestones, blockers, and cross-functional communication, requiring at least 3 years of relevant experience and security-clearance eligibility.