Jobs and Careers
PR

Senior Site Reliability Engineer

Prove
United States, United Statesfull_timeVerifiedPosted 10 Oct 2025
💰 $180,000/yr($165,000/yr$180,000/yr)

About the role

About Prove 

As the world moves to a mobile-first economy, businesses need to modernize how they acquire, engage with and enable consumers. Prove’s phone-centric identity tokenization and passive cryptographic authentication solutions reduce friction, enhance security and privacy across all digital channels, and accelerate revenues while reducing operating expenses and fraud losses. Over 1,000 enterprise customers use Prove’s platform to process 20 billion customer requests annually across industries, including banking, lending, healthcare, gaming, crypto, e-commerce, marketplaces, and payments. For the latest updates from Prove, follow us on LinkedIn.

Prove is driving the future of digital identity. We are looking for Provers who know how to make an impact. We’re talking self-starting professionals who thrive in a fast-paced environment, process information quickly, and make intelligent decisions. The work is challenging and requires not only smart but natural curiosity and tenacity. Teamwork is also important to us – we work together and play together.   

Prove has big plans, and we’re excited about the future. If this sounds like the place for you – come join our team! 

Title: Senior Site Reliability Engineer

Department: Platform Engineering 

Reports To: Manager, Site Reliability

FLSA Status: Exempt

Location: Seattle, WA

 

Position Overview

We are seeking an experienced Senior Site Reliability Engineer to join our Platform Engineering team. In this role, you will be instrumental in designing, implementing, maintaining and deploying highly available complex, scalable and reliable systems leveraging automation, effective monitoring and infrastructure-as code. Working closely with our application engineering teams to ensure our services meet the highest standards of reliability, performance, and security.

 

Key Responsibilities

 

Observability Leadership

  • Design and implement comprehensive observability solutions across our infrastructure and within applications
  • Lead the initiative to establish a companywide instrumentation standard based in Opentelemetry wide events.
  • Build advanced monitoring dashboards that provide real-time visibility into system health and performance
  •  Establish metrics, logging, and tracing systems that enable quick identification and resolution of issues
  •  Create alerting thresholds and automated responses based on service level objectives (SLOs)
  •  Drive a culture of observability throughout the engineering organization

 

Container Orchestration

  •   Lead Kubernetes cluster management, optimization, and scaling initiatives
  •   Design and implement infrastructure-as-code deployments for container-based applications
  •   Optimize container resource allocation and utilization
  •   Build automated deployment pipelines that ensure consistent, reliable releases
  •   Establish best practices for containerization and orchestration across teams

 

Infrastructure Management

  •  Design, build, and maintain scalable cloud infrastructure on AWS
  •  Implement infrastructure-as-code using tools such as Terraform
  •  Automate routine operational tasks to reduce toil and improve efficiency
  •  Ensure infrastructure security compliance and implement least-privilege access controls
  •  Optimize cloud resource utilization and costs

 

Incident Response

  •   Integrate observability-driven alerts with our Incident Management systems
  •   Lead incident response efforts during service disruptions
  •   Conduct thorough post-incident reviews and implement preventative measures
  •   Use observability data to perform root cause analysis and system improvements
  •   Document incidents, responses, and lessons learned to build organizational knowledge

 

Performance Optimization

  •   Identify and resolve performance bottlenecks across the technology stack
  •   Conduct capacity planning and scaling exercises to meet future demands
  •   Implement auto-scaling solutions based on performance metrics
  •   Optimize database performance and query efficiency
  •   Design and implement application stress testing methods and systems

 

Required Qualifications

  • 5+ years of experience in Site Reliability Engineering, DevOps, or similar roles. Software Engineering roles with a stron

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Prove

View company profile →