Executive Director Enterprise Software Testing Strategy & Solutions
DTCCAbout the role
Are you ready to make an impact at DTCC?
Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation in the financial markets. We are committed to helping our employees grow and succeed. We believe that you have the skills and drive to make a real impact. We foster a thriving internal community and are committed to creating a workplace that looks like the world that we serve.
The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure of the global capital markets. The team delivers high-quality information through activities that include development of essential, building infrastructure capabilities to meet client needs and implementing data standards and governance.
Pay and Benefits:
- Competitive compensation, including base pay and annual incentive
- Comprehensive health and life insurance and well-being benefits, based on location
- Pension / Retirement benefits
- Paid Time Off and Personal/Family Care, and other leaves of absence when needed to support your physical, financial, and emotional well-being.
- DTCC offers a flexible/hybrid model of 3 days onsite and 2 days remote (onsite Tuesdays, Wednesdays and a third day unique to each team or employee).
The Impact you will have in this role:
We are seeking an innovative, senior-level Executive Director to lead and shape our enterprise-wide IT infrastructure and resiliency testing strategy. This role will ensure seamless alignment with organizational objectives, regulatory standards, and risk management frameworks. The successful candidate will provide strategic oversight for testing methodologies spanning private cloud, network infrastructure, cloud platforms, and hybrid environments, driving excellence and resilience across our technology landscape.
Your Primary Responsibilities:
- Resilience and Reliability Leadership: Architect and execute a comprehensive resilience testing strategy across all layers of the technology stack, including infrastructure, platforms, and application across all hosting environments (on-premises data centers, private clouds, or public cloud platforms). Oversee the adoption and integration of reliability practices such as chaos engineering, automated failover, and service-level objective (SLO) monitoring to ensure robust resilience and reliability throughout the entire technology landscape.
- Automation & Tooling: Lead the implementation of advanced automation frameworks and resilience testing tools, embedding them within CI/CD pipelines and cloud deployment workflows.
- Metrics & Analytics: Define, track, and report on resilience metrics (e.g., recovery objectives, system availability, fault tolerance) to drive continuous improvement and proactive risk mitigation.
- Innovation & Emerging Technologies: Pilot and scale the use of AI/ML-driven tools for predictive resilience analytics, anomaly detection, and automated vulnerability assessment across both private and public cloud.
- Best Practices & Governance: Establish and promote enterprise-wide best practices for resilience testing, leveraging open-source frameworks, commercial offerings and cloud provider solutions.
- Collaboration & Change Leadership: Partner with senior leaders in Application Delivery, Architecture, Platform Engineering, and Security to embed resilience thinking into the software development lifecycle and foster a culture of reliability.
Talent Development: Build and lead a high-performing, diverse, and globally distributed team of resilience and quality engineering experts.
**NOTE: The Primary Responsibilities of this role are not limited to the details above. **
Key Skills
- Deep expertise in software engineering, infrastructure, and architecture, with proven leadership of quality engineering or similar functions—ideally within financial services or other highly regulated industries.
- Deep expertise in resilience engineering, infrastructure and reliability testing, for both on-premises and cloud-native architectures.
- Proven experience with chaos engineering, and automated resilience validation in large-scale organizations.
- Strong background in infrastructure-as-code (IaC), automation, and container orchestration (e.g., Kubernetes).
- Skilled in leveraging AI/ML for predictive analytics and anomaly detection related to system resilience.
- Familiarity with security, compliance, a
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s