Infrastructure Disaster Recovery & Resiliency Lead
Corebridge FinancialAbout the role
AVP, Infrastructure Disaster Recovery & Resiliency Lead
Who are we?
At Corebridge Financial, Action is Everything. We are a new company, but not a new business. Formerly AIG Life & Retirement, we are one of the largest and most established providers of retirement solutions and insurance products in the United States, with a long and proven track record of serving our clients. Every day, we proudly partner with financial professionals and institutions to make it possible for more people to take action in their financial lives, for today and tomorrow.
Get to know the business
At Corebridge Financial, our Information Technology team helps equip colleagues and customers with the latest tools designed for process efficiency and operational excellence. We also play a critical role in the protection of internal and external infrastructure to protect from security risks, while building strategies and driving innovation.
About the role
The Infrastructure Disaster Recovery & Resiliency Lead (DR & Resiliency Lead) implements the technical capabilities/solutions and process framework congruent to Corebridge Financials’ direction and goals to improve IT resiliency and operational maturity.
The DR & Resiliency Lead works closely with business resiliency, enterprise architecture, infrastructure, and application support teams to review IT resiliency requirements, solution adoption, resiliency plans, and recovery test cycles.
A successful candidate will:
- Understand various IT systems’ business requirements, overall operational environment, and facilities availability
- Work with engineering and architecture teams to validate that designs are sufficient to meet business Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO) requirements.
- Review and validate that the system visibility with proper tooling is in place to monitor/govern availability and resiliency.
- Review and validate that designs and implementations are updated as changes occur in the environment.
- Drive regular recovery and availability testing to verify that the solutions are performant and fit for purpose.
- Present findings, improvements, and opportunities to the IT Executive leadership team, Risk Management, Audit, and the Business teams.
This role will report directly to the Head of Infrastructure Operations - Service Delivery.
This position is remote-work capable; however, Corebridge Financial is looking for candidates who can work within the Eastern and/or Central time zones.
Responsibilities
- Ownership and accountability of IT resiliency operational services.
- Define IT resiliency and recovery capability operational model for maintaining services availability in disaster recovery situations and in support of continuous IT Modernization. Jointly work with enterprise/application/infrastructure teams to manage and track the implementation of required system resiliency designs, configurations, and visibility.
- Execute system recovery strategies for on-premises, co-hosted, Cloud (AWS and Azure) and SaaS applications and systems.
- Collaborate with IT teams to develop operational recovery plans and test schedules
- Coordinate recovery plan testing with IT teams and consolidate the test results
- Review results with IT teams and identify areas of improvement based on test results.
- Partner with IT and Risk Management teams in the implementation of proactive strategies and process enforcements to improve IT resiliency capabilities.
- Monitor and validate that IT resiliency solutions designs & implementations are aligned with Corebridge Financial Business and IT objectives with required resource capacity, failover and recovery flow and data recovery capabilities to meet business Recovery Time Objectives (RTO), and Recovery Point Objectives (RPO), and business service levels.
- Collaborate with internal and external partner teams to ensure appropriate resources are lined up and IT resiliency capabilities are planned and budgeted to meet resiliency and recovery requirements.
- Promote standard operating procedure documentation, knowledge base, and restoration checklist to improve disaster recovery technical documentation and recovery system architecture quality.
- Support continual process improvement initiatives and lead projects related to IT resiliency, operational efficiency, and team effectiveness.
Skill and Experience
- Bachelor's degree in Computer Science or related discipline, or equivalent work experience. Minimum 10+ years of Technical & management experience.
- Proven experience in architectural and solution capabilities to implement IT resiliency infrastructure/application designs and process framework t
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s