SRE Manager, Engineering (GMT, Remote)
CrowdStrikeAbout the role
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We’re also a mission-driven company. We cultivate a culture that gives every CrowdStriker both the flexibility and autonomy to own their careers. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.
About the Role:
As an SRE Manager for CrowdStrike’s Technical Operations team, you will lead a team of engineers in a production environment with tens-of-thousands of bare metal and virtual compute nodes. You will manage a team responsible for day-to-day operations, infrastructure management, and operational excellence. This involves a live-site first mentality for real time response to issues, monitoring of a complex network of assets, provisioning new resources, new data center assets, and the systems administration of core functionality for engineering.
Please note that this role will require you to work GMT (Hawaiian) hours.
What You'll Do:
Team Leadership
Manage and mentor a team of SRE engineers
Conduct performance reviews, career development planning, and hiring
Foster a culture of reliability, automation, and continuous improvement
Coordinate cross-functional collaboration with Engineering, Security, and Operations teams
Technical Operations
Oversee 24/7 monitoring and incident response for production systems
Drive SLI/SLO definition and monitoring across services
Lead post-incident reviews and implement preventive measures
Ensure compliance with security and regulatory requirements
Strategic Planning
Develop and execute reliability roadmaps aligned with business objectives
Capacity planning and infrastructure scaling strategies
Technology evaluation and adoption decisions
Budget planning and resource allocation
Process & Automation
Champion infrastructure-as-code and automation initiatives
Establish and improve operational procedures and runbooks
Drive adoption of observability and monitoring
Implement table-top exercises and disaster recovery testing
What You'll Need:
Bachelor's degree in Computer Science or related field, or equivalent work experience.
7+ years of engineering experience, preferably in a production role
2+ years of hands-on management experience leading engineering teams.
Demonstrated success in working across organizational boundaries to drive complex
technical initiatives
Solid design and problem solving skills with demonstrated passion for engineering
excellence, quality, security and performance
Strong cross-group collaboration and interpersonal communication skills working with
variety of roles including engineering, product management, project management, etc
Experience leading teams working with one or more of the following technologies: Linux, VMWare, FreeBSD, Storage Area Networks
Experience leading distributed teams in a remote-first environment
Proficiency in hybrid/on-prem cloud environments
Deep understanding of distributed systems and reliability engineering principles
Bonus Points:
Leadership Competencies:
Strategic thinking and ability to trans
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s