Senior Site Reliability Engineer - Database Specialty
WorldpayAbout the role
Job Description
Are you ready to write your next chapter?
Make your mark at one of the biggest names in payments. With proven technology, we process the largest volume of payments in the world, driving the global economy every day. When you join Worldpay, you join a global community of experts and changemakers, working to reinvent an industry by constantly evolving how we work and making the way millions of people pay easier, every day.
What makes a Worldpayer? It’s simple: Think, Act, Win. We stay curious, always asking the right questions to be better every day, finding creative solutions to simplify the complex. We’re dynamic, every Worldpayer is empowered to make the right decisions for their customers. And we’re determined, always staying open – winning and failing as one.
We’re looking for a Senior Site Reliability Engineer to join our ever-evolving Platforms team to help us unleash the potential of every business.
Are you ready to make your mark? Then you sound like a Worldpayer.
About the team
The Payrix team is a passionate, global team of payments and software experts who provide vertical software companies with an all-in-one platform and a white-glove approach, to capitalize on the opportunities within embedded payments for growth, innovation, and transformation. Our clients are leading vertical software providers that want to make their products stickier and grow revenue by offering payment solution to their end-customers.
Our SRE team places a high value on individuals who demonstrate intellectual curiosity and openness. We engage in collaboration across the organization, empowering our engineers to build reliable software while fostering a blameless and inclusive culture.
What you’ll own
The SRE team is dedicated to achieving operational excellence ensuring that we deliver an exceptional customer experience. In this role, you will bring database expertise to the SRE and Infrastructure teams as well as our engineering teams. You will work on helping build secure, scalable and highly available platform, collaborating with our product engineering teams to ensure alignment in reducing database related incidents, enhancing platform resilience, scalability, and advancing our incident response practices.
Work on database reliability and performance aspects from within the SRE team.
Analyze solutions and implement best practices for our database clusters (PostgreSQL).
Provide database expertise to engineering teams (for example through reviews of database migrations, queries and performance optimizations).
Work with peer SREs to roll out changes to our production environment and help mitigate database-related production incidents.
Work on observability of relevant database metrics and help achieve our database objectives.
Work on automation of database infrastructure and help build self-service tools.
Guiding product engineering teams on how to achieve operational excellence for new product and feature launches.
Collaborate with our platform and operations teams to build, maintain, performance tune and optimize our cloud infrastructure.
Debugging difficult technical challenges and making systems and products both work better.
Educate, mentor and hold the engineering teams accountable to improve the reliability of our systems and make reliability a core value of the engineering culture.
Participate in an on-call rotation to provide timely troubleshooting and resolution of urgent issues.
What you bring
An operational mindset and a drive to achieve operational excellence with at least 7+ years of experience in database administration (SQL and NoSQL).
Strong experience building and maintaining data services on RDBMS and distributed databases.
Familiarity with design and maintaining highly available cloud-based database systems in AWS or another cloud platform.
You have strong understanding of database internals, query optimization, and indexing strategies.
Experience with high availability solutions like logical replication, and clustering.
Ability to perform capacity planning and ensure an architecture is scalable to support fluctuating volumes
You have solid knowledge of SQL, PL/pgSQL and internals of PostgreSQL.
You have experience working in a distributed production environment.
You have experience with infrastructure automation and configuration management (at least one of Ansible, Terraform, Pulumi, etc)
Understanding of implementing solutions to reduce service disruptions and improving MTTD
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s