Jobs and Careers
OR

Senior Site Reliability Developer

Oracle
United States, United Statesfull_timeVerifiedPosted 12 Feb 2024
💰 $158,200/yr($79,000/yr$158,200/yr)

About the role

Customers rely on Oracle Cloud Infrastructure (OCI) to power their business as they tackle some of the world’s biggest challenges. We’re looking for Senior Site Reliability Developers/Engineers who would be responsible for Advanced Operations (AO) and critical issues of production environments, including systems and databases, supporting critical business operations. Will perform administration and analysis for multiple production environments and recommend new and novel solutions to improve availability, performance, and supportability. This is an opportunity to bring a combination of deep technical knowledge with administration/analysis knowledge of Oracle's Cloud Infrastructure to provide critical issue support to a wide range of complex production environment problems related to immense growth, scaling, using the cloud, extremely high performance, and high availability requirements. 

What you’ll do

You will be working in Advanced Operations (AO) on the Site Reliability Development/Engineering (SRD/SRE) team for US Gov Operations with shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioural characteristics of production services. Responsible for the design and delivery of the critically important stack, with focus on security, resiliency, scale, and performance. Authority for end-to-end performance and operability. Partner with development teams and Point Operations (PO) in defining and implementing improvements in service architecture. Articulate technical characteristics of services and technology areas and guide Development Teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. Understand and communicate the scale, capacity, security, performance attributes, and requirements of the service and technology stack. Demonstrate clear understanding of automation and orchestration principles. You will act as ultimate point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs). This role will allow you to apply a deep understanding of service topology and their dependencies required to solve issues and define mitigations. Understand and explain the effect of product architecture decisions on distributed systems. Professional curiosity and a desire to a develop deep understanding of services and technologies. 

 

 

Requirements:

  • U.S. Citizenship
  • Bachelor’s Degree in Computer Science or other STEM related fields, and/or equivalent professional experience and demonstration of consistent career growth. (M.S. an advantage)
  • Experience with Linux (or any UNIX OS) System Administration, Networking, Storage, Compute and Virtualization. (Certifications in CCNA networking, Virtualization, or UNIX OSs an advantage)
  • Strong familiarity of cloud concepts, platforms, distributed systems, and networking.
  • Strong Cybersecurity awareness/experience (e.g. CompTIA Security+, CISSP, all advantages)
  • Experience in participating and leading incident bridges. 
  • Customer obsession, passion for delighting customers. 
  • Experience in cloud technical support, operations, NOC or similar is preferred, but not required. 
  • Demonstrable ability to quickly learn new technical domains and then train others. 

 

What we’ll offer you

  • A competitive salary with exciting benefits
  • Flexible and remote working so you can do your best work
  • Learning and development opportunities to advance your career
  • An Employee Assistance Program to support your mental health
  • Employee resource groups that champion our diverse communities
  • Core benefits such as medical, life insurance, and access to retirement planning
  • An inclusive culture that celebrates what makes you unique

 

 

Shift Format:

This role is a Monday through Friday core hours role; will involve working on-call shift rotations for escalated incident management 10-15%(max) (i.e. on-call support on ~50 days) of a calendar year based on a 40-hour work week, including nights, weekends and public holidays.

The role involves working 5 days a week (40 hours) and one additional on-call time on an every-two-months basis. You will be assigned to a shift rota at most once every two months.

Overtime is not generally expected in this role, there may be extreme exceptions for on-call or overtime duties rarely and exclusively based on business needs.

Disclaimer:

Certain US customer or client-facing roles may be required to comply with applicable require

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Oracle

View company profile →