Jobs and Careers
ST

Senior Site Reliability Engineer

Striveworks
United Statesfull_timeVerifiedPosted 25 Feb 2025
💰 $190,000/yr($150,000/yr$190,000/yr)

About the role

The Role

As a Senior Site Reliability Engineer (SRE) at Striveworks, you will be challenged—and trusted—on day one to take ownership of specific product deployments by maintaining, optimizing, and enhancing our on-premises and cloud computing environments. You will play a crucial role in the successful deployment of our software solutions to clients. You will be responsible for executing technical aspects of implementation projects and for ensuring the seamless integration, customization, and configuration of our software. Your expertise will play a critical role for the company as we deploy new instances of Striveworks’ machine learning operations (MLOps) capabilities to customer infrastructure.

You are right for this opportunity if you value and possess technical expertise and you enjoy pushing the boundaries of your capabilities. You will be responsible for maintaining Striveworks’ software deployments using Infrastructure-as-Code (IaC) methodologies. 

Your day-to-day will include:

  • Automating IaC to manage virtual machines and deploy containers, services, and other infrastructure; leaning on expertise to deploy custom Kubernetes clusters in AWS, Azure, GCP, on-premises, or hybrid cloud environments
  • Working with platform developers, DevOps, and customer-facing teams to define requirements and build solutions for customer use cases of the platform
  • Software deployments to commercial and, later, unclassified, CUI, Secret, and Top Secret Department of Defense (DoD) networks
  • Incident response and initial triage of critical system faults

The Senior SRE works on the DevOps team and acts as a liaison between DevOps, platform developers, and customer-facing teams, taking on operational tasks to ensure the efficient functioning of Striveworks’ solutions. The Senior SRE monitors, automates, and improves software reliability, performance, and availability for various projects. They work alongside a team of software engineers and data scientists to help them deploy and operate their work as functional products, learning from them so that building effective AI solutions becomes second nature. They may provide guidance and leadership to junior SRE team members.

You will directly contribute to the success of mission-critical systems within national security and commercial clients. You will be expected to wear multiple hats and to step into vacuums where improvements are needed, and you will be given the breadth to explore new technologies and solutions. 

The anticipated base pay range for this position is $150,000 to $190,000/year. Striveworks’ total compensation package includes a competitive base salary and annual performance-based equity grants. 

This position offers a hybrid/on-site work environment at our office in northwest Austin with up to 20% travel.

The Right Fit

During our hiring process, we spend a lot of time discussing shared values. 

Why? We passionately believe that fostering an environment where people can self-actualize and pursue greatness is the best way to achieve our individual and collective goals. 

What does this mean for you? We want to create an environment where you can thrive and achieve your goals, where you know the team shares your goals, and where you make and accept decisions for the team with humility. At Striveworks, we want your say/do ratio to be 1, and we want you to know that being part of a top-tier team means that there is no smartest person in the room. If that makes sense, we’re already on the same page.  

Here’s what we’re looking for:

  • 6+ years of direct, hands-on experience in:
    • Microservice deployment in Kubernetes
    • Diagnosing and resolving issues within containerized environments
    • Helm Chart and Kustomizations development/deployment
    • Python and Bash programming
    • Automation and IaC (e.g., Terraform, Ansible)
    • Cloud infrastructure (e.g., AWS, Azure, GCP, or OpenStack)
    • Managing and troubleshooting Linux systems (e.g., RHEL, Ubuntu, CentOS)
  • The ability to work cross-functionally to define requirements and build solutions for customer use cases of the platform
  • The ability to respond professionally and competently to incident reports and triage critical system faults
  • US person (Permanent Resident or US Citizen), or otherwise able to obtain a Top Secret security clearance
  • Willingness and ability to obtain and maintain a Top Secret security clearance

The Wish List

We are very interested in candidates who possess the above qualifications, and we appreciate and consider the addition of:

  • Active Top Secret security clearance and intimate familiarity with DOD networking, tools, infrastructure, security require

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Striveworks

View company profile →