Senior Site Reliability Engineering Manager
DexcomAbout the role
About Dexcom
Founded in 1999, Dexcom, Inc. (NASDAQ: DXCM), develops and markets Continuous Glucose Monitoring (CGM) systems for ambulatory use by people with diabetes and by healthcare providers for the treatment of people with diabetes. The company is the leader in transforming diabetes care and management by providing CGM technology to help patients and healthcare professionals better manage diabetes. Since the company’s inception, Dexcom has focused on better outcomes for patients, caregivers, and clinicians by delivering solutions that are best in class - while empowering the community to take control of diabetes. Dexcom reported full-year 2022 revenues of $2.9B, a growth of 18% over 2021. Headquartered in San Diego, California, with additional offices in the Americas, Europe, and Asia Pacific, the company employs over 8,000 people worldwide.
Meet the team:
Dexcom’s Site Reliability Engineering (SRE) team exists to empower our SW Dev Teams to engineer highly reliable systems through which people take control of their health. We love hands on experimentation and continuous learning.
Where you come in:
We’re looking for an experienced, driven Senior Manager of Site Reliability Engineering, someone who's hands-on and eager to lead. Your role will be pivotal in spearheading the creation and upkeep of cutting-edge, dependable systems. As the manager of our SRE team, you'll oversee a group of skilled professionals, including SREs and DBAs, and collaborate closely with our cloud and application engineering teams. Together, we'll ensure our customers enjoy seamless experiences powered by resilient and streamlined systems.
Key Responsibilities:
1. Hands-On Engineering and Innovation:
Remain hands-on with code, contributing directly to critical system components and automation tools.
Innovate and experiment with new approaches to improve system reliability, performance, and efficiency.
Stay at the forefront of SRE practices, continuously integrating new ideas and technologies into our systems.
2. Mentorship and Team Development:
Mentor and guide team members, fostering a culture of excellence and continuous learning.
Share expertise and insights, helping to elevate the technical capabilities of the entire SRE team.
Lead by example, demonstrating best practices in coding, system design, and operational excellence
3. Execution and Team Management:
Define and communicate clear goals and objectives, aligning team efforts with organizational priorities.
Implement efficient workflows and processes to maximize team productivity without compromising quality.
Foster an environment of accountability and transparency, where team members are empowered to take initiative and ownership of their work.
Regularly review team performance and progress, providing constructive feedback and making strategic adjustments as needed.
Promote collaboration and effective communication within the team and across departments to ensure smooth execution of projects.
4. Strategic Incident Management and Prevention:
Oversee critical incident management processes, ensuring rapid and effective resolution.
Analyze system performance and incidents to identify trends and areas for improvement.
Develop and implement proactive strategies to prevent system failures and enhance overall reliability.
Drive a blameless culture of continuous learning and growth.
What makes you successful:
Extensive experience as a Site Reliability Engineer or in a similar role, demonstrating strong technical leadership and the ability to effectively manage and execute complex projects.
Expertise in the latest SRE tools and practices, such as Kubernetes and gitops driven pipelines, with proficiency in implementing these efficiently within teams.
Exceptional problem-solving skills for designing and implementing complex systems, focused on streamlined execution and team synergy.
Proficiency in programming and automation, maintaining a hands-on approach that ensures operational excellence and timely project delivery.
Outstanding communication and mentorship skills, adept at leading and inspiring teams, fostering a collaborative and results-driven work environment.
Preferred bonus skills include GitHub Actions, ArgoCD, Helm, Crossplane, Datadog, Google Cloud Platform, Go, Python, and Functional Programming, enriching team versatility and problem-solving capacity.
What you’ll get:
A front row seat to life changing CGM technology. Learn about our brave #dex
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s