Senior Manager, Site Reliability
PindropAbout the role
Senior Manager, Site Reliability
US-Remote
Who we are
Are you passionate about innovating at the intersection of technology and personal security? At Pindrop, we recognize that the human voice is a unique personal identifier, increasingly susceptible to sophisticated fraud, including the threat of deepfakes. We're leading the way in developing cutting-edge authentication, fraud prevention, and deepfake detection. Our mission is to provide seamless and secure digital experiences, safeguarding the most personal aspect of our identity: our voice. Here, you'll be part of a team driven by values of Innovation, Customer Advocacy, Excellence, and Impact. We're not just creating a safer digital landscape by fortifying trust and integrity with those we serve, we’re also building a dynamic, supportive workplace where your contributions make a real difference.
Headquartered in Atlanta, GA, Pindrop is backed by world-class investors such as Andreessen-Horowitz, IVP, and CapitalG.
Our tech stack
-
Redhat/CentOs/Ubuntu Linux
-
Chef/Ansible/Terraform/Packer/Helm/Flux
-
Python, Go and Bash
-
Jenkins/Github Actions, Github, Nginx, VMWare/Proxmox, Splunk, ELK, Artifactory, Prometheus/Grafana, Kubernetes (EKS/GKE/Helm/Flex),
-
AWS/GCP infrastructure services such as Compute, Postgres, DynamoDB, Blob Storage/S3, Lambda
What you’ll do
At Pindrop, we commit to creating Evangelical Customers for Life; because of this, customer experience is at the forefront of everything we do. To help us build functional systems that improve the customer experience we are now looking for an experienced Sr. Manager of SRE. You will be responsible to help improve Reliability, Scalability and Security of our SaaS platform. You will be responsible for overall infrastructure for our Cloud platform and Datacenter setup, deploying security updates, Reporting Uptime and other relevant SLAs and managing and evolving our incident management process.
- Manage, lead, and uplevel the SRE team as well as the wider Engineering team.
- Collaborate with other Engineering managers to evolve our Infrastructure to the next level of reliability and scalability.
- Lead your teams to deliver with a high level of executive and achievement.
- Define and manage project plans and deliverables for the SRE group.
- Well-versed in SRE tools, concepts and methodologies.
- Help the team Investigate and resolve technical issues.
- Work across various Engineering teams (DevOps, Platform, Product) to identify areas of improvement in our SaaS platform.
- Set strategic goals for operational efficiency and increased productivity.
Who you are
- You possess a problem-solving attitude and are adaptable in the face of change
- You are positive and enthusiastic with a high level of cross-functional partnership across other technical teams and Pindrop as a whole
- You bring a collaborative approach and excite your teams by creating a high-performing team culture
- You are an excellent communicator with high quality presentation skills as well as verbal and written communication skills
- You are resilient in the face of challenges, change, and ambiguity. Being the “unsung hero” is just another day on the job as you keep our infrastructure in excellent working order
- You are optimistic and believe that you can make a problem into a solution
- You are resourceful, excited to uncover innovative solutions and teach yourself something new when needed
- You take accountability, do the things you say you’ll do, under-promise and over-deliver
- You are nimble and adaptable when priorities change and continue to see the “forest through the trees”
Your skill-set
- Bachelor's degree (or equivalent) in computer science or related field, or equivalent experience required.
- 6+ years experience directly and formally managing and leading SRE team and function
- Prior professional experience as a SRE Engineer or similar role
- Architectural experience building and managing large scale SaaS systems
- Experience with multiple cloud technologies required (one of: AWS or GCP)
- Knowledge of Terraform, Ansible, Python or Golang required
- Working knowledge of Kubernetes and Infrastructure as a Code concepts.
- Well versed with incident management process, uptime SLA, and, error budget etc.
- Possess technical confidence and familiarity with SRE tools, techniques, and more importantly, mindset
- Proven ability to develop innovative solutions that lead to increased productivity
Nice to
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s