Jobs and Careers
EL

Platform SRE - Site Reliability Engineer

Elasticsearch
United States, United Statesfull_timeVerifiedPosted 1 Aug 2023
💰 $200,560/yr($126,800/yr$200,560/yr)

About the role

Elastic is an open source search company that powers enterprise search, observability, and security solutions built on one technology stack that can be deployed anywhere. From finding documents to monitoring infrastructure to hunting for threats, Elastic makes data usable in real time and at scale. Thousands of organizations worldwide, including Barclays, Cisco, eBay, Fairfax, ING, Goldman Sachs, Microsoft, The Mayo Clinic, NASA, The New York Times, Wikipedia, and Verizon, use Elastic to power mission-critical systems. Founded in 2012, Elastic is a distributed company with Elasticians around the globe. Learn more at elastic.co

Thanks to our ongoing expansion we have the opportunity to grow our Site Reliability team. We're a part of the Elastic Cloud engineering team with a focus on solving Cloud operations problems and keeping the SaaS online, who aren’t afraid to get our hands dirty. We are the first line of consumers for Elastic's products and our experience helps influence the direction of the stack. While most organizations may have a single or a handful of Elastic Stack deployments, here you’ll be responsible for identifying, troubleshooting and reporting platform problems to product engineers (or fixing the code yourself) in order to ensure that the thousands of Elasticsearch clusters we manage are providing a stable and reliable service. We’re looking for people who are just as passionate about troubleshooting issues with distributed systems as they are to automate, code and collaborate to solve problems.

What You Will Be Doing:
  • You will report and solve problems within the Elastic Cloud infrastructure services and collaborate on issues with product engineers
  • You will participate in SRE software engineering, writing code for the continuing reduction of human intervention in operational tasks and automation of processes
  • You will monitor the Elastic Cloud platform and Cloud infrastructure, responding to incidents, correcting and improving systems to prevent incidents and planning capacity
  • You will manage Cloud provider infrastructure, system deployments and product releases
  • You will be involved in resolving Elastic Cloud customer support issues
  • You will demonstrate and promote best practices for teams using Cloud platforms
  • You will participate in 24x365 on-call schedules
What You Bring Along:
  • You are either an experienced sysadmin with professional skills in Linux, preferably on distributed systems at scale, and a demonstrable interest in using software engineering to solve operational problems; or a software engineer with real interest, and ideally some experience, in Linux systems, networking, monitoring and automation.
  • You have experience with Java and preferably a DevOps background
  • You have at least three years of experience using a public Cloud; AWS, GCP, Azure, Softlayer or OpenStack
  • You are comfortable writing software to automate API-driven tasks at scale. SRE use Python and Go regularly but are also encouraged to contribute to the product codebase in Java, Scala, and Python.
  • You have used Ansible, Puppet, Chef or another config management suite, know where it's broken, and open to trying new alternatives
Bonus Points:
  • Healthy knowledge of Linux (have compiled your own kernel at some point, know how to trace syscalls, understand TCP, care about the difference between sysvinit/runit/systemd, etc.)
  • Relentless desire to automate and build software tools
  • Desire to represent work in git, driven by a GitHub workflow through issues and pull requests
  • Love open source development, and have contributed to some project somewhere (doesn't have to be ours), whether through mailing lists, patches, documentation, etc.
  • Enjoy working remotely and the communication it requires
  • Love a diverse environment, working with people all over the world

#LI-CB1


Compensation for this role is in the form of base salary.  This role does not have a variable compensation component.  

The typical starting salary range for new hires in this role is listed below.  In select locations (including Seattle WA, Los Angeles CA, the San Francisco Bay Area CA, and the New York City Metro Area), an alternate range may apply as specified below. 

These ranges represent the lowest to highest salary we reasonably and in good faith believe we would pay for this role at the time of this posting.  We may ultimately pay more or less than the posted range, and the ranges may be modified in the future.  

An employee's position within the salary range will be based on several factors including, but not limited to, relevant education, qualifications, certifications, experience, skills, geographic location, performance, and business or organizational needs.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Elasticsearch

View company profile →