Jobs and Careers
CO

Senior Site Reliability Engineer | Remote US

Coalfire
United States, United StatesRemotefull_timeVerifiedPosted 21 Mar 2024
💰 $135,000/yr($78,000/yr$135,000/yr)

About the role

About Coalfire
Coalfire is on a mission to make the world a safer place by solving our clients’ toughest cybersecurity challenges. We work at the cutting edge of technology to advise, assess, automate, and ultimately help companies navigate the ever-changing cybersecurity landscape. We are headquartered in Denver, Colorado with offices across the U.S. and U.K., and we support clients around the world.  But that’s not who we are – that’s just what we do.  We are thought leaders, consultants, and cybersecurity experts, but above all else, we are a team of passionate problem-solvers who are hungry to learn, grow, and make a difference.    And we’re growing fast.  We’re looking for a Site Reliability Engineer to support our Cloud Services team. This can be a remote position (must be located in the United States). Position Summary
As a Senior Site Reliability Engineer at Coalfire within our Cloud Services (CMS) group, you will be a self-starter, passionate about cloud technology, and thrive on problem solving. You will work within major public clouds, utilizing automation and your technical abilities to operate the most cutting-edge offerings from Cloud Service Providers (CSPs). This role directly supports leading cloud software companies to provide seamless reliability and scalability of their SaaS product to the largest enterprises and government agencies around the world.

What You'll Do

  • Become a member of a highly collaborative engineering team offering a unique blend of Cloud Infrastructure Administration, Site Reliability Engineering, Security Operations, and Vulnerability Management across multiple clients.
  • Coordinate with client product teams, engineering team members, and other stakeholders to monitor and maintain a secure and resilient cloud-hosted infrastructure to established SLAs in both production and non-production environments.
  • Be a subject matter expert on innovating and implementing using automated orchestration and configuration management techniques. Deeply understand the design, deployment, and management of secure and compliant enterprise servers, network infrastructure, boundary protection, and cloud architectures using Infrastructure-as-Code.
  • Create, maintain, and peer review automated orchestration and configuration management codebases, as well as Infrastructure-as-Code codebases. Maintain IaC tooling and versioning within Client environments.
  • Define processes, implement, and upgrade client environments with CI/CD infrastructure code, and provide and facilitate internal feedback to development teams for environment requirements and necessary alterations.
  • Own clients across AWS, Azure and GCP, serving as a SME and optimizing their unique native services in client environments.
  • Configure and tune cloud-based tools, manage cost, security, and compliance for the client’s environments.
  • Identify repetitive tasks/areas of improvement and develop technical solutions to automate repeatable tasks as well as enhancements to CMS offerings.
  • Respond to environment-specific alerts, and review dashboards via analytics tools such as Splunk and Elastic Stack.
  • Work closely with client DevOps and product teams to provide 24x7x365 support to environments through Client ticketing systems.
  • Serve as a leader for definition, testing, and validation of incident response and disaster recovery documentation and exercises.
  • Participate in on-call rotations as needed to support Client critical events that may lay outside of business hours.
  • Serve as a SME for testing and data reviews to evaluate the effectiveness of current security and operational measures, in addition to remediating deviations from current security and operational measures.
  • Maintain and author detailed diagrams representative of the Client’s cloud architecture.
  • Create, maintain, and peer review standard operating procedures, operational runbooks, technical documents, and troubleshooting guidelines.

What You'll Bring

  • BS or above in related Information Technology field or equivalent combination of education and experience
  • 5+ years experience in 24x7x365 production operations
  • 5+ years experience supporting cloud operations and automation in AWS, Azure or GCP (and aligned certifications)
  • 5+ years experience with Infrastructure-as-Code and orchestration/automation tools such as Terraform and Ansible
  • Expert-level experience with IaaS platform capabilities and services (cloud certifications expected)
  • Strong experience working within an automated CI/CD pipeline for release development, testing, remediation, and deployment
  • Strong experience working within container orchestration solutions such as Kubernetes, Docker, EKS and/or ECS
  • Strong experience within ticketing tool solutions such as Jira and ServiceNow
  • Experience using env

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Coalfire

View company profile →