Jobs and Careers
KL

Senior Site Reliability Engineer- Platform Services

Klaviyo
United Statesfull_timeVerifiedPosted 5 Mar 2024
💰 $235,200/yr

About the role

At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo (we call ourselves Klaviyos) brings to our workplace each and every day. We believe everyone deserves a fair shot at success and appreciate the experiences each person brings beyond the traditional job requirements. If you’re a close but not exact match with the description, we hope you’ll still consider applying. Want to learn more about life at Klaviyo? Visit careers.klaviyo.com to see how we empower creators to own their own destiny.

Senior Site Reliability Engineer - Platform Services 

Engineers come to Klaviyo with experience in a variety of languages and from a number of disciplines. All engineers are expected to become extremely proficient in the technologies we use (not exhaustive):

  • Python, Golang, Bash
  • GRPC, Protobuf, Django, Gunicorn
  • Apache Pulsar, RDS Aurora MySQL, NGINX, Terraform, Redis, Cloudflare
  • Prometheus, Grafana, Splunk, Cloudwatch
  • Amazon Web Services (EC2, RDS, etc.), Kubernetes on EKS, Istio

The SRE Platform Services owns runtime services used by the product development teams. These services include common database and database proxies, edge gateway routing and proxies, and asynchronous task management systems. Our two biggest areas of planned work in 2024 are: 1) Klaviyo Messaging System, our asynchronous processing framework built in Python and Apache Pulsar, and 2) a modern API gateway to route external traffic to team owned services and provide common functionality like auth and internationalization.

As a Senior Site Reliability Engineer you will take a large role in these foundational Klaviyo services and make a big impact on the productivity of our product engineering teams and customers.

How You'll Make a Difference

  • Ship foundational services to enable Klaviyo engineering to move faster with confidence
  • Design and develop systems and processes that enable highly available & scalable systems
  • Design, build and deliver software to dramatically improve the availability, scalability, latency, and efficiency of Klaviyo’s services
  • Achieve break-throughs in systems throughput by identifying and eliminating bottlenecks
  • Leverage technology such as Python, AWS, Apache Pulsar, Django, Kubernetes, Bash, Terraform, MySQL, and Redis to advance Klaviyo’s platform
  • Champion best practices by actively collaborating with other teams in a culture that values whiteboarding and technical design review
  • Contribute to the company as a subject matter expert in multiple areas, constantly pushing yourself to be a better engineer and to level up all of your peers within your team and within Klaviyo.
  • Mentor and pair with other Klaviyo engineers to build better software by focusing on performance, self-healing system, configuration as code; defensive programming, application security, etc.
  • Participate in periodic on call duties with a focus on solving issues when they are discovered, preventing recurrences and minimizing alert fatigue 
  • Prototype and advocate for architectural improvements to achieve breakthrough results in Klaviyo systems’ operational scalability and reliability
  • Work hand-in-hand with product-facing engineers to ship impactful code
  • Perform quantitative investigation to understand and scale Klaviyo systems and manage the cross-functional effort to resolve scalability issues
  • Produce and advocate for preventative, upstream solutions with internal stakeholders and external vendors and dependencies
  • Confidently make informed, data-driven choices in a fast paced environment with competing priorities

Who You Are 

  • Knowledge of Linux operating systems and computer networking
  • Experience writing code in a programming language such as Python, Ruby, Go, etc.
  • Experience administering cloud-based infrastructure (e.g. AWS)
  • Ability to troubleshoot production issues related to computer infrastructure, configuration, monitoring, deployments, and continuous integration and delivery
  • Ability and willingness to learn
  • Ability to communicate clearly and mentor and coach others on a team
  • Ability to participate in an on-call rotation

The pay range for this role is listed below. Sales roles are also eligible for variable compensation and hourly non-exempt roles are eligible for overtime in accordance with applicable law. This role is eligible for benefits, including: medical, dental and vision coverage, health savings accounts, flexible spending accounts, 401(k), flexible paid time off and company-paid holidays and a culture of learning that includes a learning allowance and access to a professional coac

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Klaviyo

View company profile →