Production Reliability Engineer
Early WarningAbout the role
At Early Warning, we’ve powered and protected the U.S. financial system for over thirty years with cutting-edge solutions like Zelle®, Paze℠, and so much more. As a trusted name in payments, we partner with thousands of institutions to increase access to financial services and protect transactions for hundreds of millions of consumers and small businesses.
Positions located in Scottsdale, San Francisco, Chicago, or New York follow a hybrid work model to allow for a more collaborative working environment.
Candidates responding to this posting must independently possess the eligibility to work in the United States, for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship.
Overall Purpose
This position is responsible for stability, performance, and growth of key business platforms.
Essential Functions
Work closely with Production Reliability Engineer – Agile Release Train, Product Owners, Architecture, Security, Engineering, and other teams to collaborate on requirements, priorities, etc. Provide input and guidance to platform projects to ensure reliability, scalability, current functionality, capacity, and performance requirements are not adversely impacted.
Documents, manages, and supports code deployments, hardware upgrades, patching, certificate renewals into the Customer Acceptance Testing (CAT), Disaster Recovery (DR) and Production (PRD) environments. Perform the appropriate functional, regression, performance and other assisted testing.
Work closely with Product and Customer Enablement Team to maintain up to date topology diagrams and business workflow diagrams for supported platforms.
Review, provide clarification, update, and manage Business Review Documents (BRD) and Production Readiness Checklists (PRC). Provide status report updates as requested by Project Managers and company leaders. Prioritize, manage, and keep current on tasks, stories, and assignments within team Kanban board.
Identify, create, and publish customer communications regarding maintenance windows, code deployments, Disaster Recovery Exercises or any other potential customer impacting events within Customer Acceptance Testing (CAT), Disaster Recovery (DR) and Production (PRD) environments.
Identify areas where efficiencies and/or automation could: improve processes, reduce, or eliminate manual processes, reduce risk, and provide a better user experience. Work with others to architect the solution, create the epics or stories required, prioritize, test, implement and measure the improvements.
Support the company commitment to risk management and protecting our integrity and confidentiality of systems and data. This is not limited to but can include: identifying, creating, documenting, managing, and supporting Self-Identified-Issues (SII’s)
Provide support on testing, timelines, and requirements for bank on-boarding
Review changes in User Acceptance Testing (UAT) and participate in discussions for scheduling what changes will be deployed and into which environment
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s