Senior Site Reliability Engineer
HiveWatchAbout the role
<div class="content-intro"><p><strong>About Us:</strong></p> <p>HiveWatch is a tech-forward, inclusive organization fostering the evolution of the physical security industry. We are a diverse team of forward thinkers who empower each other to find creative and collaborative solutions in an industry ripe for modernization. We are passionate about the problems we’re solving for our customers and equally passionate about the company we’re building. </p> <p>HiveWatch is here to help security teams pivot from chasing threats to preventing them. We protect organizations, people, and property through the intelligent orchestration of physical security programs. With better communication, more insights, and less “noise”, we are modernizing what it means for businesses and their employees to truly feel safe.</p></div><p><strong>POSITION OVERVIEW:</strong></p> <p>HiveWatch is seeking a Senior Site Reliability Engineer to join our Platform Team, where you'll build and operate mission-critical edge infrastructure that connects our SaaS platform to customer systems. You'll help ensure exceptional performance, reliability, and observability across our distributed environment.</p> <p><strong>WHAT YOU'LL DO</strong>:</p> <ul> <li>Improve the reliability of mission-critical systems including production monitoring, alerting, and capacity planning</li> <li>Debug and resolve complex production issues across the full stack, from infrastructure to application code</li> <li>Participate in a regular on-call rotation to provide 24/7 coverage for critical systems</li> <li>Perform root cause analysis requiring deep code-level investigation and implement preventive measures</li> <li>Build automation and tooling to reduce operational toil and improve system reliability</li> <li>Maintain CI/CD pipelines, observability infrastructure, and database performance optimization</li> <li>Increase the resiliency, scalability, and maintainability of production environments</li> <li>Maintain on-call procedures and disaster recovery processes</li> <li>Contribute to on-call runbooks, postmortems, and reliability best practices</li> </ul> <h3> </h3> <p><strong>OUR TECH STACK:</strong></p> <ul> <li>Languages: Kotlin, Rust, TypeScript, and Python</li> <li>Deployments: GitHub Actions, Terraform, Terragrunt, and Helm</li> <li>Infrastructure: AWS (Kinesis, Serverless, RDS, EKS), Kubernetes, Docker, Postgres, IoT Edge, Red Hat Enterprise Linux, Rocky Linux</li> </ul> <p><strong>PREFERRED QUALIFICATIONS</strong>:</p> <ul> <li>Strong experience with AWS architecture and services</li> <li
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s