Jobs and Careers
GA

Dev Ops & Site Reliability Engineer (SRE)

Gather
Palo Alto, United Statesfull_timeVerifiedPosted 16 Apr 2024

About the role

Our Mission at Gather

You generate enormous amounts of personal data when you use the internet. This data is extremely powerful and could make your life easier, better, more magical. So why aren't you using it? 

At Gather, we’ve developed a product that effortlessly enables you to consolidate your digital world – from your Twitter likes to your Kindle highlights – with a single click.

Thanks to our unique data access approach, we're pioneering the definitive personal AI assistant. It seamlessly merges GPT's problem-solving prowess with deep context about your life. Whether it's acting as a memory aid, providing insights about your life, or anticipating your future needs, Gather's AI intuitively understands you from the moment you two meet.

Our team consists of individuals who embody a big vision, show a lot of hustle, and share lots of laughter.  The office exudes palpable energy, and we are eager to welcome the next team member!

Join us at Gather, and play a key role in building a future where personal data and AI intersect to empower the individual.

We are based in-person in Palo Alto and offer relocation assistance as needed to new employees.

Role Overview:

We are currently seeking an exceptional DevOps/SRE to become a valuable member of our team. In this role, you will play a pivotal part in overseeing our application stack and cloud infrastructure, ensuring seamless orchestration and management of our services. Your responsibilities will encompass the design, development, and maintenance of our internal automation tools, which are crucial for efficiently managing our service lifecycle. Additionally, you'll be tasked with diagnosing and resolving runtime issues spanning the various tiers of our hosting stack.

Core Responsibilities:

  • Manage containerized applications using technologies like Docker and Kubernetes

  • Implement monitoring, logging, and proactive issue identification

  • Architect and manage cloud-based infrastructure on GCP

  • Design resilient infrastructure for high availability and disaster recovery

  • Automate infrastructure setup and configuration using tools like Terraform, Ansible, or Puppet

  • Handle incident response and contribute to problem-solving efforts when necessary

  • Foster collaboration between development and operations teams

Nice To Have

  • Establish and optimize CI/CD pipelines for automated software delivery

Your qualifications:

  • Bachelor's degree in Computer Science, Engineering, or a related field

  • 3+ years experience in building and maintaining cloud-native production infrastructure

  • Strong passion for meticulously documenting and automating intricate data systems.

  • Proficiency in Infrastructure as Code (IaC) practices.

  • A solid grasp of cutting-edge monitoring solutions and techniques.

  • Expertise in at least one modern programming language such as Python, Go, or similar.

  • Team player who is driven by ensuring the highest level of product quality

  • Desirable: Previous exposure to troubleshooting and enhancing production infrastructure.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Gather

View company profile →