Senior Site Reliability Engineer
PayPalAbout the role
The Company
PayPal has been revolutionizing commerce globally for more than 25 years. Creating innovative experiences that make moving money, selling, and shopping simple, personalized, and secure, PayPal empowers consumers and businesses in approximately 200 markets to join and thrive in the global economy.
We operate a global, two-sided network at scale that connects hundreds of millions of merchants and consumers. We help merchants and consumers connect, transact, and complete payments, whether they are online or in person. PayPal is more than a connection to third-party payment networks. We provide proprietary payment solutions accepted by merchants that enable the completion of payments on our platform on behalf of our customers.
We offer our customers the flexibility to use their accounts to purchase and receive payments for goods and services, as well as the ability to transfer and withdraw funds. We enable consumers to exchange funds more safely with merchants using a variety of funding sources, which may include a bank account, a PayPal or Venmo account balance, PayPal and Venmo branded credit products, a credit card, a debit card, certain cryptocurrencies, or other stored value products such as gift cards, and eligible credit card rewards. Our PayPal, Venmo, and Xoom products also make it safer and simpler for friends and family to transfer funds to each other. We offer merchants an end-to-end payments solution that provides authorization and settlement capabilities, as well as instant access to funds and payouts. We also help merchants connect with their customers, process exchanges and returns, and manage risk. We enable consumers to engage in cross-border shopping and merchants to extend their global reach while reducing the complexity and friction involved in enabling cross-border trade.
Our beliefs are the foundation for how we conduct business every day. We live each day guided by our core values of Inclusion, Innovation, Collaboration, and Wellness. Together, our values ensure that we work together as one global team with our customers at the center of everything we do – and they push us to ensure we take care of ourselves, each other, and our communities.
Job Summary:
The Cloud Infrastructure and Site Reliability Engineering team at PayPal is responsible for ensuring high availability, performance and scalability of critical systems powering PayPal Shopping/Honey’s business. We collaborate with cross-functional teams, automate operational processes, and drive reliability best practices to improve end-to-end system architecture and reduce incidents.This job influences process quality and effectiveness while overseeing team performance. They determine appropriate actions in complex situations, engage in incident response, and contribute to developing scalable software systems, ensuring the reliability of digital services through collaboration and expertise.
Job Description:
Essential Responsibilities:
- Take ownership of system performance monitoring, identify inefficiencies, and lead initiatives to improve the overall availability and reliability of digital platforms and applications.
- Lead and manage the response to complex, high-priority incidents, ensuring prompt resolution and a thorough root cause analysis to prevent future occurrences.
- Design and implement advanced automation frameworks to improve operational efficiency, streamline processes, and reduce human error.
- Lead reliability-focused initiatives, ensuring systems are highly available, resilient, and scalable, and promote best practices across engineering teams.
- Enhance the monitoring infrastructure by identifying key metrics, optimizing alerting, and improving system observability to ensure the reliability of large-scale systems.
- Forecast resource requirements and lead capacity planning activities to ensure systems can scale effectively to meet growing user demand.
- Ensure robust disaster recovery strategies are in place and conduct regular testing to ensure systems can recover quickly from failures.
- Partner with engineering and product teams to identify opportunities for improving system architecture, focusing on scalability, reliability, and fault tolerance.
- Provide mentorship and technical guidance to junior site reliability engineers, fostering skill development and knowledge sharing.
- Drive continuous improvement across operational workflows, identifying areas for optimization, cost reduction, and performance enhancement.
Expected Qualifications:
- 3+ years relevant experience and a Bachelor’s degree OR Any equivalent combination of education and experience.
Additional Responsibilities & Preferred Qualifications:
Y
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s