Staff Site Reliability Engineer, Remote, 4 day week (Swing shift)
ServiceNowAbout the role
Company Description
At ServiceNow, our technology makes the world work for everyone, and our people make it possible. We move fast because the world can’t wait, and we innovate in ways no one else can for our customers and communities. By joining ServiceNow, you are part of an ambitious team of change makers who have a restless curiosity and a drive for ingenuity. We know that your best work happens when you live your best life and share your unique talents, so we do everything we can to make that possible. We dream big together, supporting each other to make our individual and collective dreams come true. The future is ours, and it starts with you.
With more than 7,700+ customers, we serve approximately 85% of the Fortune 500®, and we're proud to be one of FORTUNE 100 Best Companies to Work For® and World's Most Admired Companies™.
Learn more on Life at Now blog and hear from our employees about their experiences working at ServiceNow.
Unsure if you meet all the qualifications of a job description but are deeply excited about the role? We still encourage you to apply! At ServiceNow, we are committed to creating an inclusive environment where all voices are heard, valued, and respected. We welcome all candidates, including individuals from non-traditional, varied backgrounds, that might not come from a typical path connected to this role. We believe skills and experience are transferrable, and the desire to dream big makes for great candidates.
Job Description
The Site Reliability Engineering team is a group of highly technical engineers who are tasked with maintaining and developing the reliability, scalability and performance of the ServiceNow platform and infrastructure. The SRE is empowered to drive technical resolutions across the technology stack from application through to hardware and all stops in between. The ultimate goal of the SRE is to never have to escalate an issue to an engineering or development team and to completely own the resolution of incidents. They are also tasked with driving forward the operability of the platform to drive down incident numbers and to reduce MTTR. To accomplish this the team combines Software Development, Networking and Systems Engineering expertise with a strong desire to be challenged by problems of scale and complexity and to make services better for our customers.
What you get to do in this role:
As an Engineer in the SRE team you will:
- Provide relief and sustainable resolution to issues within our infrastructure.
- Use your experience in software development, systems engineering, and networking to proactively prevent repeatable issues.
- Drive initiatives with partner teams to improve the reliability and performance of the infrastructure through improved system design.
- Drive a culture of intolerance to manual activity which results in a highly automated environment delivering scalable solutions
- Drive monitoring and automation initiatives
**Please note this is a Swing shift role with a Wed-Sat working week, and includes a shift allowance to compensate**
Qualifications
To be successful in this role you have:
- Deep knowledge of Linux systems
- Experience working with relational database: MySQL, MariaDB or PostgresSQL.
- Experience working with systems at scale - supporting critical services with focus on automation, observability, availability, and performance.
- Experience with Kubernetes to orchestrate the deployment, scaling, and management of containers.
- Experience Coding in various languages; preferrably Python, JavaScript, and Ruby
- Networking skills, IP addressing and routing.
- Team-first attitude and an uncompromising attention to detail.
- An eye for proactively anticipating potential issues, expertise in performing root cause analysis, and a mindset focused on building effective solutions to prevent recurrence.
- Good collaboration and communication skills
Good to have:
- Expertise in Observability and Monitoring of applications, services, and networks at scale.
- Experience with DevOps automation, CI/CD pipeline and agile methodologies such as Gitlab CI-CD.
- Experience writing test specifications and understand the fundamentals of test automation.
- Experience working with Cloud technologies such as Azure and AWS.
- Experience in configuration management of infrastructure using Ansible or puppet.
- Experience developing on the ServiceNow Platform
Additional Information
ServiceNow
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s