Staff Site Reliability Engineer - Event Management and Service Assurance Administration
VisaAbout the role
Company Description
Visa is a world leader in digital payments, facilitating more than 215 billion payments transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories each year. Our mission is to connect the world through the most innovative, convenient, reliable and secure payments network, enabling individuals, businesses and economies to thrive.
When you join Visa, you join a culture of purpose and belonging – where your growth is priority, your identity is embraced, and the work you do matters. We believe that economies that include everyone everywhere, uplift everyone everywhere. Your work will have a direct impact on billions of people around the world – helping unlock financial access to enable the future of money movement.
Join Visa: A Network Working for Everyone.
Job Description
Essential Functions
Responsible for day-to-day support, security, maintenance and availability of the event management tools, based upon business requirements while adhering to tight operations, security and procedural models.
Responsible for complex integrations, advanced maintenance, high availability, disaster recovery, use case development, customer support, training and documentation of our Tools environments.
Function as the Event Management resource on key infrastructure projects and initiatives.
Demonstrate a technical acumen to support complex Tooling Solutions, under minimal supervision.
Operate in complex, highly-secure, highly-available, operations-centric datacenter environments and interact with other technology domain experts as required, to build and maintain the System Management solutions in those datacenter environments.
Ability to build documentation for tooling infrastructure.
Work with other Engineering disciplines to develop Reference Architectures and technology standards based on Infrastructure Building Blocks that can be offered as services to our internal customers.
Participate in On-call support rotation.
Maintaining and managing Event Management Systems
Performing Changes
Upgrading the Event Management system so its compliant with Vendor support
Applying any TSR , Qualys security patches
Performing yearly DR activities
Reporting
User Management.
Developing and testing new integrations.
Evaluation of new Tools for AI/ML.
Work with Vendors for solutioning and providing additional functionalities that will help in increasing the availability and Performance Monitoring of VISA applications.
This is a hybrid position. Hybrid employees can alternate time between both remote and office. Employees in hybrid roles are expected to work from the office 2-3 set days a week (determined by leadership/site), with a general guidepost of being in the office 50% or more of the time based on business needs.
Qualifications
Basic Qualifications
• 5 or more years of relevant work experience with a Bachelors Degree or at least 2 years of work experience with an Advanced degree (e.g. Masters, MBA, JD, MD) or 0 years of work experience with a PhD OR 8+ years of relevant work experience.
Preferred Qualifications
• 6 or more years of work experience with a Bachelors Degree or 4 or more years of relevant experience with an Advanced Degree (e.g. Masters, MBA, JD, MD) or up to 3 years of relevant experience with a PhD.
• At least 5 or more years of relevant experience working with Event Management Tools like Netcool/Moogsoft/BigPanda/DataDog/ITOM etc.
• Experience in monitoring tools like Catchpoint/Thousand Eyes/SCOM.
• 5 years plus in a successful development environment regularly releasing well-tested solutions with complex integrations to meet application and service objectives.
• Working knowledge on integration of Netcool (Omnibus/WebGui and Impact) with integrations.
o Ticketing tools like Service Now/Remedy etc.
o Notification tools like MIR3/Everbridge/PagerDuty etc
• 4-7 years' experience required with scripting & UI languages such as PowerShell, PERL, PYTHON, etc.
• Scripting proficiency in both Windows and UNIX environments
• Competent with Linux, Unix and Windows operating systems commands, performance and debugging, and an understanding of web services.
• A strong foundation in the fundamentals of computer science, complex application architectures and complex systems.
• Hands-on experience with NOI features like IBM Event Analytics
• Knowledge of virtualization technologies and products from VMware, Citrix or Microsoft
• Hands-on experience with Service Now.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s