Senior Specialty Systems Operations Engineer
Wells FargoAbout the role
About this role:
Wells Fargo is seeking a highly motivated and experienced Systems Engineer to join our Systems Operations team. In this role, you will be a key contributor to ensuring the stability, reliability, and performance of our critical systems and infrastructure. You will leverage your technical expertise to lead and participate in application upgrades, vulnerability remediation, and automation initiatives, while collaborating with cross-functional teams to resolve technical issues and drive continuous improvement. The ideal candidate will possess a strong understanding of application monitoring, cloud deployment, CI/CD pipelines, and automation tools, with a passion for applying Site Reliability Engineering (SRE) principles to optimize system performance and availability.
In this role, you will:
Lead or participate in managing all installed systems and infrastructure, moderately complex application upgrades and vulnerability remediation efforts, within the Systems Operations functional area
Lead team to meet moderately complex technical deliverables while leveraging solid understanding of technical process controls or standards
Act independently as a liaison for the line of business in support of daily inquiries, problem and incident management, project delivery, and escalations by following established guidebooks
Liaise with functional or operational managers to understand their current and future information needs and develop plans and schedules for integrating these needs into existing operations
Recommend solutions to resolve technical issues and achieve highest levels of systems and infrastructure availability by automating platform activities to lower human intervention time on related tasks
Review and analyze moderately complex operational support systems, application software, and system management tools to ensure the highest levels of systems and infrastructure availability
Facilitate discussions on preventative action, root cause analysis and resolutions with service management
Collaborate and consult with peers, mid-level managers, vendors and other technical personnel to resolve technical issues and achieve highest levels of systems and infrastructure availability and reliability
Provide training and mentoring to lower experienced team members on guidebook changes and lead team to meet technical deliverables, while leveraging solid understanding of technical process controls or standards
Collaborate with development teams to update business continuity plans and provide input on development of new and existing support guidebooks
Required Qualifications:
4+ years of systems engineering or technology architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
Desired Qualifications:
Advance understanding of application monitoring stack (Logs, Metrics, Events, Traces, Alerts) and ability to visualize and setup end to end observability (Infra and App components)
Strong experience in using industry standard monitoring tools (AppDynamics, Splunk, ELK, APM, Grafana, Prometheus
Experience in deploying the application to cloud platforms
Experience in using CI/CD tools like Jenkins, uDeploy, Gradle, Groovy and Maven
Experience in CM tools like Ansible and Puppet
Proficient in one of the programming Languages (Java and Python)
Knowledge of Web services
Experience in working Agile methodology
Proficient in multiple infrastructure technologies
Linux experience
Database experience - hands on in Oracle SQL, Pl/SQL, Mongo DB
DevOps experience
Very good experience on Autosys
Knowledge of Abinitio
Exposure to tools like Service Now, JIRA
Hands on experience in Ansible, Python, Shell Scripting
Familiarity with NDM, SFTP
Basics of Networking
Job Expectations:
Design, code , test and deliver software to automate manual operation work
Partner with different application teams throughout the life cycle to understand their application infrastructure monitoring and apply site reliability principles to baseline and set up SLOs for critical components
Identify
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s