Jobs and Careers
FU
Senior SRE Engineer
Function HealthUS - Remote, United StatesRemotefull_timeVerifiedPosted 9 Dec 2025
About the role
About Us:Function was founded with a singular focus: empower you to live 100 healthy years. We’re doing that by using the best available technology to make sure people don't suffer or die a preventable death. Function has been recognized as one of Fast Company’s Most Innovative Companies of 2024, and is venture-backed by Andreessen Horowitz (a16z). Hundreds of thousands of members have joined Function to take control of their health. We are growing our team and seeking out world-class talent that deeply believes in our mission to positively impact global health, has a relentless bias toward action and a growth mindset. Function fosters a collaborative and dynamic environment, where every day we are building the future.
Role Overview:The ideal SRE/DevOps engineer is a technically strong, security‑minded problem‑solver who operates production systems with a calm, data‑driven approach, proactively improves tooling, and communicates effectively across teams while continually leveling up their own skills and those of the organization.
Primary Responsibilities:
Automate Software Delivery
Operate Production Infrastructure
Observe, Troubleshoot & Remediate
Optimize Performance & Cost
Continuously Improve Tooling & Process
Collaborate & Coach
Role Overview:The ideal SRE/DevOps engineer is a technically strong, security‑minded problem‑solver who operates production systems with a calm, data‑driven approach, proactively improves tooling, and communicates effectively across teams while continually leveling up their own skills and those of the organization.
Primary Responsibilities:
Automate Software Delivery
- Build and maintain robust CI/CD pipelines (e.g., GitHub Actions, Jenkins, Argo CD) that integrate automated testing, security scanning, and one‑click rollback to accelerate safe releases.
Operate Production Infrastructure
- Provision, configure, and manage secure, highly‑available cloud (AWS, GCP) and on‑prem environments with Infrastructure as Code (Terraform).
Observe, Troubleshoot & Remediate
- Instrument systems with metrics, logs, and traces (Prometheus, Grafana, Datadog, OpenTelemetry); own the on‑call rotation, rapidly diagnose incidents, and drive blameless post‑mortems.
Optimize Performance & Cost
- Continuously assess latency, capacity, and cloud spend; tune applications and scale containerized workloads (Docker, Kubernetes) to meet SLAs while controlling costs.
Continuously Improve Tooling & Process
- Research, evaluate, and standardize new tools or practices that boost reliability, security, or developer velocity; automate toil wherever it appears.
Collaborate & Coach
- Partner with software, QA, and security teams to embed DevOps/SRE best practices; create clear documentation and share operational knowledge.
Must‑Have Skills & Experience:
- 6-8 years in production SRE/DevOps or related software‑engineering roles, including participation in an on‑call rotation
- Deep hands‑on expertise with at least one major provider (AWS, GCP, or Azure), covering networking, IAM, and managed services
- Proficient with Terraform (preferred) or similar IaC tooling; experienced in module design, remote state, and policy‑as‑code
- Proven ability to design declarative pipelines and automate build, test, deploy, and rollback workflows
- Strong knowledge of Docker image design and Kubernetes operations (Helm, controllers, service meshes)
- Practical use of monitoring, logging, and tracing stacks; comfortable leading incident bridges and post‑incident analysis
- Fluency with Git workflows, code review culture, and clear written/verbal communication
- Proficiency in at least one language (Python, Go, or Bash) to automate tasks and build small services
- Self‑starter who is proactive, curious, and relentlessly focused on eliminating manual toil
Nice-to-Have Qualifications:
- Security or compliance experience (SOC 2, ISO 27001, FedRAMP)
- Database operations at scale (PostgreSQL, MySQL, Redis, or MongoDB)
- Observability platform tuning (Grafana Loki, Elastic, Honeycomb)
- Relevant certifications (CKA/CKAD, AWS SA‑Pro, Terraform Associate)
To be a strong fit you also need:
- Bias Toward Acti
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s