Technical Leader, Site Reliability Engineer
CiscoAbout the role
Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received.
Meet the Team
Cisco’s Collaboration Business Unit empowers people and organizations worldwide to connect, communicate, and innovate seamlessly.
You will collaborate with a global team of software engineers and SREs responsible for delivering extraordinary collaboration experiences at scale. Our team supports backend services deployed worldwide and works closely with development, product, and operations partners to ensure reliability and performance.
Webex is powering the shift to the hybrid workforce, helping people stay connected in a rapidly evolving digital world. We cultivate a startup-like culture that values innovation, ownership, and collaboration, while offering the scale and impact of a global technology leader.
Who we are
Cisco Collaboration is the global engine of hybrid work, delivering the secure, AI-driven communication platform that connects millions of users and empowers the world’s largest enterprises to innovate together. The Persistence Team is a foundational component of this ecosystem. We provide Database-as-a-Service (DBaaS) at a large scale, managing distributed data footprints globally across private datacenters and AWS. We are responsible for the deployment, reliable operations, and resilience of the technology that powers the entire Cisco Collaboration Suite.
Who You’ll Work With
You will join a globally distributed team of experts with diverse backgrounds who take pride in maintaining high availability with zero-downtime migrations. Our footprint is vast; you will collaborate with technical leads and architects across the entire Webex portfolio. You’ll also partner with our security teams to ensure our US Federal environments remain compliant. You’ll participate in a follow-the-sun on-call rotation during your regions daytime hours.
About you
You are a distributed big data enthusiast who believes that if you have to do it twice, you should automate it. You thrive on solving complex "stateful" problems in a "stateless" world and are passionate about building resilient infrastructure that can survive regional outages.
Some of the things you will work on
Architect & Deploy: Design and implement large-scale solutions for Cassandra, Kafka, and OpenSearch across hybrid environments (AWS & Private Cloud).
Automate Everything: Lead cloud engineering projects using Terraform, Ansible, and Kubernetes to improve service reliability and deployment velocity.
Scale & Optimize: Lead capacity planning and performance tuning for high-throughput production systems.
Ensure Integrity: Manage automated workflows for backups, security patching, and incident management with a focus on quality assurance in live environments.
Stakeholder Leadership: Act as a consultant to application teams, guiding them on data access patterns and cloud migration strategies.
Minimum Qualifications
8+ years hands on experience (deploy, monitor, scale, and upgrades) with distributed database technology such as CassandraDB, Opensearch, Kafka, PostgreSQL
5+ Years managing AWS infrastructure at scale.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s