Jobs and Careers
Q2

Senior Site Reliability Engineer

Q2
Aspen Lake 3, United States, United Statesfull_timeVerifiedPosted 8 Jan 2025

About the role

As passionate about our people as we are about our mission.

What We’re All About:

Q2 is proud of delivering our mobile banking platform and technology solutions, globally, to more than 22 million end users across our 1,300 financial institutions and fintech clients.  At Q2, our mission is simple: Build strong, diverse communities by strengthening their financial institutions. We accomplish that by investing in the communities where both our customers and employees serve and live.

What Makes Q2 Special?

Being as passionate about our people as we are about our mission. We celebrate our employees in many ways, including our “Circle of Awesomeness” award ceremony and day of employee celebration among others! We invest in the growth and development of our team members through ongoing learning opportunities, mentorship programs, internal mobility, and meaningful leadership relationships. We also know that nothing builds trust and collaboration like having fun. We hold an annual Dodgeball for Charity event at our Q2 Stadium in Austin, inviting other local companies to play, and community organizations we support to raise money and awareness together.

The Hosting Platform Team comprised of Platform and Site Reliability Engineering are experts at distilling customer needs and designing integrated systems.  Our highly collaborative and resourceful team is entrusted with implementing new capabilities and customizing end-to-end solutions.  From developing deep Q2 application knowledge, managing our container orchestration platform, supporting private and public cloud environments, learning how to leverage automation to drive efficiencies, and troubleshooting critical infrastructure, your opportunities to make a significant impact at Q2 are limitless.

Responsibilities

If you were working for us as a Senior Platform Engineer focused on designing, implementing and scaling mission critical systems, here are some of the things you would have done last week:

  • Design and develop infrastructure solutions for enterprise business applications to be hosted in datacenter and public cloud (AWS and Azure)

  • Partner with Software Engineers and Database Administrators to deliver a 10x increase in full stack performance

  • Collaborate with Security Engineering to embed best in class threat management capabilities

  • Join forces with Site Reliability Engineering to optimize production, performance and implement automation to minimize risk and reduce toil.

  • Maintain, test and optimize disaster recovery services to ensure we have a contingency plan to failover our highly available platform

  • Serve as the Subject Matter Expert, distill complex technical issues into consumable communication and lead cross-functional teams.

  • Optimize storage costs and data consumption services

  • Promptly restore system services by diagnosing problems, making critical decisions and remediating root causes.

  • Share your knowledge with a peer, document your process, and facilitate lessons learned reviews

  • Participate in an on-call rotation to assist with customer-impacting outages

  • Typically requires a minimum of 12 years of related experience with a Bachelor’s degree; or 6 years and a 
    Master’s degree; or a PhD with 3 years experience; or equivalent experience. Some barriers to entry exist at this level, requiring department review.

  • Experience designing and integrating complex systems, performing high-level optimizations, and leading large-scale cross-functional initiatives

  • Advanced knowledge of Linux and Windows Systems Administration

  • Fluent in VMWare and cloud environments such as AWS and Azure

  • Strong understanding of load balancing and traffic management

  • Understanding of security protocols and threat management implementation

  • Knowledge of CI/CD Pipelines implementation for applications and infrastructure

  • Proficiency with Automation / Orchestration tools such as Ansible and HashiCorp Consul, Nomad, Vault, Packer, and Terraform or equivalent technologies

  • Real-world experience integrating Data Eventing / Messaging Bus capability via technology such as Kakfa or Rabbit MQ is a plus

  • Capable of embedding Monitoring / Logging technologies such as Elastic Stack, Prometheus / Grafana or equivalent technologies

  • Knowledge of best practices of running applications in containerized environments including health checks and rolling update strategies

  • Experience with scripting / development languages such as Go, Python, Bash, or PowerShell

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Q2

View company profile →