Principal Site Reliability Engineer, Product Software
EquinixAbout the role
Who are we?
Equinix is the world’s digital infrastructure company®, operating over 250 data centers across the globe. Digital leaders harness Equinix's trusted platform to bring together and interconnect foundational infrastructure at software speed. Equinix enables organizations to access all the right places, partners and possibilities to scale with agility, speed the launch of digital services, deliver world-class experiences and multiply their value, while supporting their sustainability goals.
Our culture is based on collaboration and the growth and development of our teams. We hire hardworking people who thrive on solving challenging problems and give them opportunities to hone new skills and try new approaches, as we grow our product portfolio with new software and network architecture solutions. We embrace diversity in thought and contribution and are committed to providing an equitable work environment that is foundational to our core values as a company and is vital to our success.
Principal Site Reliability Engineer, Product SoftwareEquinix is the world’s digital infrastructure company, operating 250 data centers across the globe and providing interconnections to all the key clouds and networks. Businesses need one place to simplify and bring together fragmented, complex infrastructure that spans private and public cloud environments. Our global platform allows customers to place infrastructure wherever they need it and connect it to everything they need to succeed.
At Equinix, we help the world’s digital leaders scale with agility, speed the launch of digital services, deliver world-class experiences, and transform people’s lives. Our culture is based on collaboration and the growth and development of our teams.
We hire hardworking people who thrive on solving challenging problems and give them opportunities to hone new skills, and try new approaches, as we grow our product portfolio with new software and network architecture solutions. We embrace diversity in thought and contribution and are committed to providing an equitable work environment. that is foundational to our core values as a company and is vital to our success
Job Summary
Equinix is the world's digital infrastructure company. Much of the internet that you know flows through our rapid interconnected network and data centers. Equinix Internet Exchange is the world's largest Internet Exchange while Equinix Fabric has the largest market share in Software Defined Interconnection. Our team is building Equinix's next generation networking and Multi Cloud Networking platforms which will enable all Equinix's Digital Interconnection services.
We're looking for a Principal SRE who will be responsible for building easy to operate, reliable and high preforming applications and networking software for our next generation networking platform.
As a Principal SRE, you will drive strategy for building highly scalable and reliable systems. You will be a stakeholder in software and network design and engineering, providing technical leadership in all aspects of running services across critical and pivotal initiatives. You will focus on scaling, modernizing, and automating Equinix’s Multi Cloud Networking service and next generation networking architectures. You will have the flexibility to work on different projects to utilize new technologies as we continue to explore new areas and drive innovation across the infrastructure space. This makes you a key contributor in building and maintaining the components that drive our customers’ experience.
This role presents a great opportunity for a passionate senior technical leader to make a tangible impact to the future-state technology stack of Equinix’ global networks and infrastructure. If you believe in the power of technology to change the world, and if you value creativity, work comfortably in a fast-paced environment and are eager to learn new technologies, we would like to talk to you.
Responsibilities
Design and manage networking software services in a highly concurrent, scalable, distributed transactional systems
Provide leadership in designing and improving incident management processes, defining oncall schedules and responsibilities, leading the team in troubleshooting strategies and fixing production issues in quick turnaround time
Create software systems or lead the team in using systems such as CICD pipelines to help deliver code from commit to deployment in minutes
Develop an observability strategy for the services
Work with the team for constant improvement of system performance and scale using system profiling tools and stress testing techniques
Provide thought leadership and significant technical contributions to help develop target Multi
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s