Jobs and Careers
MA

Senior Software Engineer, Site Reliability Engineering

Mastercard
O'Fallon, United Statesfull_timeVerifiedPosted 20 May 2025
💰 $115,000/yr

About the role

Our Purpose

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

Title and Summary

Senior Software Engineer, Site Reliability Engineering

Senior Software Engineer, Site Reliability Engineering

The Mastercard Digital Enablement Services (MDES) team is looking for a strong, innovative Site Reliability Engineer to contribute to Mastercard's next generation of Digital Payment products that change how our customers choose to pay. We are looking for Senior Engineers who can bring unique perspectives and innovative ideas to all areas of DevOps and are interested in continuing to improve our platform through the ever-changing technology landscape.
In this role, our mission is to bridge gaps between Software Engineering, Operations, Internal and External Partners by strengthening relations, safely progressing engineering feature delivery and commitments. Through this role we shorten feedback loops, add collaborative focus on lowering operational overhead, provide exemplary capacity management, increasing availability & resiliency, and reducing latency to the MDES platform.
An ideal candidate is someone who works well with a cohesive team-oriented group of engineers, has a passion for developing, improving, and implementing great software, enjoys solving problems in a challenging environment, and has the desire to take their career to the next level.
If any of these opportunities excite you, we would love to talk to you!

Role
• Be part of a team of site reliability engineers supporting services before they go live through activities such as system design consulting, performance engineering, tuning, chaos testing, capacity planning and launch reviews.
• Increase, maintain and communicate service metrics once live by measuring and monitoring availability, latency, performance and overall system health.
• Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability, velocity and recommend performance tuning enhancements.
• Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
• Review production incidents to identify and drive solutions to prevent reoccurrence and minimize customer impact.
• Manage individual project priorities, deadlines, and deliverables.
• Create and maintain technology roadmaps.
• Look at all tasks with an eye for automation; then work to automate them.
• Build, manage and maintain robust dashboards reflecting system health.
• Apply expert technical capabilities across discipline(s) to troubleshoot and solve problems.

All About You
Education (preferred):
Bachelor’s Degree in Computer Science, Computer Systems, Information Technology or related. Equivalent experience is acceptable.

Desirable Knowledge/Experience

Minimum:
• Solid knowledge and experience with web applications and distributed systems infrastructure.
• Solid Knowledge and understanding of Software Engineering Concepts and Methodologies
• Excellent verbal and written communication adjustable to a diverse audience with various levels of technical and business acumen.
• Experience with monitoring and alerting tools like Dynatrace, Splunk, Prometheus.
• Interest and ability to learn new coding languages like Java and Python, frameworks like Spring, and paradigms as needed.
• IT experience including demonstrating thought-leadership and relationship building across large-scale organizations
• Knowledge and understanding of Service Level Objectives, Observability, Golden Signals and Availability calculations.
• Experience with static analysis tools to improve software quality
• Knowledge of CI/CD platforms (Jenkins, Bamboo, Concourse, XLR, etc)

Preferred:
• Experience with Java, Python, Scala, or other Object-oriented programming languages
• Experience with Git, BitBucket, Stash or other version control systems
• Experience with Maven and/or Gradle
• Experience building pipelines in Jenkins, Bamboo, Concourse or XLR
• Experience building test suites in JMeter, LoadRunner, Gatling and/or Blazemeter
• Experience with performance tuning of cloud-native applications
• Experience working across teams to troubleshoot complex issues and providing guidance

Skills/Abilities:
• Disp

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Mastercard

View company profile →