Jobs and Careers
SU
Senior Site Reliability Engineer
Supernova TechnologyUnited Statesfull_timeVerifiedPosted 20 Aug 2025
💰 $170,000/yr($130,000/yr – $170,000/yr)
About the role
About UsFounded in 2014, we offer the industry’s first and only cloud-based, fully-customizable, end-to-end software solution to automate securities-based lending from origination through the life of the loan. By combining thought leadership in suitability and risk management with industry-leading education and the latest technology, Supernova enables advisors to deliver holistic, goals-based advice and to help their clients achieve financial wellness. We partner with the industry’s largest banks, most prominent insurance companies and leading online brokerages to democratize access to securities-based lending and better the entire financial ecosystem.
Why Join Supernova?At Supernova Technology, we believe that the best results come from a team that is passionate, driven, and supported in all aspects of their professional lives. Here, you’ll work alongside talented and innovative individuals who are committed to driving the future of securities-based lending technology. We foster a culture of collaboration, continuous learning, and growth, where each person’s contributions make a real impact.
Job DescriptionThe Senior Site Reliability Engineer will own the reliability, scalability, and performance of our production systems. This role bridges engineering, platform, and security teams to ensure infrastructure meets strict uptime, compliance, and client experience requirements. This position will lead the design and implementation of observability tools, incident response processes, and resilience strategies, shifting the organization from reactive to proactive reliability practice
Why Join Supernova?At Supernova Technology, we believe that the best results come from a team that is passionate, driven, and supported in all aspects of their professional lives. Here, you’ll work alongside talented and innovative individuals who are committed to driving the future of securities-based lending technology. We foster a culture of collaboration, continuous learning, and growth, where each person’s contributions make a real impact.
Job DescriptionThe Senior Site Reliability Engineer will own the reliability, scalability, and performance of our production systems. This role bridges engineering, platform, and security teams to ensure infrastructure meets strict uptime, compliance, and client experience requirements. This position will lead the design and implementation of observability tools, incident response processes, and resilience strategies, shifting the organization from reactive to proactive reliability practice
RESPONSIBILITIES:
- Ensure systems meet high-availability targets through well-defined SLAs, SLOs, and SLIs.
- Own and optimize the monitoring, logging, and alerting stack to ensure actionable alerts.
- Lead incident response and postmortem processes, driving remediation and prevention.
- Plan capacity and optimize performance to address bottlenecks before they impact customers.
- Automate operational tasks to reduce manual intervention.
- Collaborate with DevOps to improve CI/CD reliability and with Platform Engineering to ensure infrastructure scalability.
- Implement reliability controls required for SOC 2 and other regulatory standards.
QUALIFICATIONS:
- 5-8 years in SRE, operations, or performance engineering roles.
- Bachelor's Degree in Computer Science or related fields
- Advanced expertise with monitoring and alerting tools.
- Proficiency in at least one programming or scripting language such as Python, Go, or Bash.
- Strong background in AWS cloud environments.
- Experience with container orchestration using AWS ECS.
- Proven track record in leading high-severity incident response calmly and effectively.
- Familiarity with ITIL, postmortem processes, and change management controls.
- Demonstrated ability to work cross-functionally with development, platform, and security teams.
- Reliability-focused mindset with an emphasis on uptime and recovery speed.
- Analytical problem-solving skills supported by metrics and data.
- Calm and effective performance in high-pressure situations.
- Technical depth to diagnose and resolve complex system issues.
- Proactive leadership in anticipating and addressing reliability risks.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s