Sr. Manager, Site Reliability Engineer (SRE)
Planet FitnessAbout the role
About Us
Founded in 1992 in Dover, NH, Planet Fitness is one of the largest and fastest-growing franchisors and operators of fitness centers in the world by number of members and locations. As of March 31, 2026, Planet Fitness had approximately 21.5 million members and 2,909 clubs in all 50 states, the District of Columbia, Puerto Rico, Canada, Panama, Mexico, Australia and Spain. The Company’s mission is to enhance people’s lives by providing a high-quality fitness experience in a welcoming, non-intimidating environment, which we call the Judgement Free Zone®. Approximately 90% of Planet Fitness clubs are owned and operated by independent business owners.
At Planet Fitness, our unique mission has always been to enhance people’s lives by providing a high-quality fitness experience in a welcoming, non-intimidating environment. And we’re proud of the amazing Planet Fitness team that supports our clubs and team members. They are comprised of dynamic, dedicated, and talented individuals who represent our values of integrity, transparency, passion, respect, and excellence (while having fun!) in everything they do.
Joining the PF family means being part of a company that cares about bettering the health and wellbeing of our communities. It means being a part of a supportive, engaging workforce with an inclusive culture that values diversity and creates an environment where everyone can feel they belong. It means encouraging professional growth and development. It means making true, lasting connections with your co-workers with celebrations, team building activities and engaging corporate events! It means creating a positive impact in our local communities through our Judgement Free Generation® philanthropic initiative. It means being part of a brand that you can be proud of!
For the past 30 years, we’ve helped millions of people in their fitness journey and revolutionized the industry along the way. And we’re just getting started!
Overview
The Sr. Manager, Site Reliability Engineering (SRE) leads the strategy, execution, and continuous improvement of reliability, availability, and performance across Planet Fitness’s retail technology ecosystem. This role is responsible for ensuring that both digital platforms and in-club systems run reliably, efficiently, and at scale. This position will lead teams responsible for incident management, observability, platform reliability, and end-to-end technology support. This role operates at the intersection of software engineering and infrastructure, driving automation, reducing toil, and embedding reliability into the software development lifecycle.
This role follows a hybrid schedule and requires regular, in-person work at our Hampton, NH or future Boston, MA office. Our hybrid model is M/T/W in office and TH/F are optional work-from-home. Candidates must reside within commuting distance of either office. Fully remote work is not available for this role.
Responsibilities
Reliability & Performance Engineering
- Define and own Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets across critical systems.
- Drive reliability engineering practices across the software lifecycle, embedding resilience into system design.
- Lead performance engineering, capacity planning, and scalability strategies aligned to retail demand cycles.
- Implement and mature incident management processes, including post-incident root cause analysis (RCA) and continuous improvement loops.
- Design and implement cross-functional support and escalation process for IT platforms to align with business and technical goals; including stakeholder alignment across engineering, operations, and support teams.
- Act as incident commander for high-severity issues, ensuring rapid resolution and clear stakeholder communication.
- Establish production readiness and operational acceptance criteria for new platforms and services.
Platform & Automation
- Champion infrastructure as code (IaC), automation, and self-healing systems to reduce manual intervention and operational toil.
- Partner with platform engineering to build scalable, resilient cloud-native architectures.
- Drive adoption of CI/CD pipelines, safe deployment strategies (e.g., canary, blue/green), and automated rollback mechanisms.
Observability & Monitoring
- Implement comprehensive observability (metrics, logs, traces) to provide actionable insights into system health.
- Standardize monitoring frameworks and alerting strategies aligned to business-critical services.
- Enable real-time visibility into customer experience across digital and physical retail channels.
Cross-Functional Collaboration
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s