Jobs and Careers
TA

Senior Site Reliability Engineer

Tactile Medical
Minneapolis, United Statesfull_timeVerifiedPosted 22 Apr 2026
💰 $131,040/yr($93,600/yr$131,040/yr)

About the role

At Tactile Medical, we specialize in developing at-home therapy devices to treat lymphedema, chronic venous insufficiency and respiratory illnesses.

The Senior Site Reliability Engineer (SRE) is responsible for ensuring reliability, observability, and operational excellence across Tactile Medical’s digital products and internal platforms. This includes the digital therapy ecosystem (mobile apps, React portals, clinician tools), the .NET API layer, CosmosDB backed data platforms, WooCommerce commerce components, Azure Service Bus integrations, and the cloud infrastructure that supports regulated medical device workflows. This role sits at the intersection of DevOps, cloud operations, compliance, and product support, ensuring that production systems meet uptime expectations and regulatory requirements while enabling rapid iteration for the Digital Solutions and Software Engineering teams. The SRE will help establish and mature the operational reliability strategy — including incident management, performance monitoring, infrastructure automation, and continuous improvement — with a specific focus on supporting a regulated medical device + digital health environment. The systems managed by this role directly impact patients and device connectivity, making reliability and quality essential to business continuity and patient outcomes.

Accountabilities & Responsibilities
Production Ownership & Incident Response:

  • Serve as the operational owner for the production environment supporting Tactile’s digital solutions.
  • Lead incident response processes, coordinating with Digital, IT, Marketing, Operations, and Product Support teams.
  • Participate in on-call rotation and oversee escalation pathways for Tier 2 & 3 technical support.
  • Ensure post incident documentation aligns with regulated quality expectations (e.g., CAPA inputs, RCA documentation in accordance with ISO 13485 / QMS processes).

Observability, Monitoring & Data Quality:

  • Build and maintain end to end observability across: Native and hybrid mobile applications, Patient, partner and internal portals, Device connectivity & data ingestion services, Payment and WooCommerce commerce flows, .NET backend services and Azure integrations
  • Build dashboards and alerts in Datadog, Azure Monitor (or preferred tools) to detect anomalies.
  • Conduct database level investigations for usage analytics, reliability metrics, and management level reporting (patient usage trends, connectivity patterns, error rates).

Automation, Infrastructure & Cloud Operations:

  • Lead infrastructure automation using Terraform and Azure DevOps
  • Automate monitoring configuration, system audits, log standards, and compliance-related reporting.
  • Collaborate with IT Security and Compliance to maintain operational readiness for HIPAA and internal QMS audits.
  • Support the transition of legacy components toward more scalable and modern cloud patterns where needed.

Reliability Engineering & Development Partnership:

  • Define and maintain SLOs, SLIs, and reliability metrics that balance innovation velocity with platform stability.
  • Work closely with developers to embed reliability into CI/CD, code quality, test coverage, and deployment patterns.
  • Lead post incident reviews and manage the continuous reliability improvement backlog.
  • Offer guidance on resilient architecture decisions, retry patterns, failure modes, and performant API design.

Cross-Functional Collaboration:

  • Act as a reliability subject matter expert across Digital Solutions, Product, Engineering, Product Support, and Security.
  • Ensure production change control aligns with quality and regulatory expectations.
  • Support compliance documentation for software releases, infrastructure changes, and security controls.

 

Qualifications

Education & Experience

Required:

  • Bachelor’s degree in Computer Science, Information Technology, or related field.
  • Master’s degree or certifications (e.g., Azure, Kubernetes, SRE) are a plus.
  • 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure roles.
  • Proven experience in regulated industries (healthcare, finance, etc.) is highly

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Tactile Medical

View company profile →