Jobs and Careers
BO

Senior Site Reliability Engineer, Platform Foundation

Box
Redwood City, United Statesfull_timeVerifiedPosted 16 Nov 2023
💰 $218,000/yr($174,500/yr$218,000/yr)

About the role

WHAT IS BOX?

Box is the market leader for Cloud Content Management. Our mission is to power how the world works together. Box is partnering with enterprise organizations to accelerate their digital transformation by creating a single platform for secure content management, collaboration and workflow. We have an amazing opportunity to further establish ourselves as leaders in the space, and we need strong advocates to help us achieve that goal. By joining Box, you will have the unique opportunity to help capture a majority of this developing market and define what content management looks like for the digital enterprise. Today, Box powers 100,000+ businesses, including many top Fortune 500 companies who trust our secure collaboration platform to manage the entire content lifecycle.

WHY BOX NEEDS YOU 

The Platform Foundation team's mission is to transform the way Boxers build services by providing a common platform to configure, deploy, and run Box services seamlessly, securely, reliably, and efficiently in optimal regions of the world, so that they can deliver awesome products.  Our team owns Box's Platform as a Service (PaaS) based on Kubernetes. With Box being an early adopter, supporter, and contributor of Kubernetes, you will work for the company that introduced kubectl, kube-state-metrics, kube-applier, and kube-exec-controller to the open-source community. Every team and service at Box relies on our platform and we are looking for a passionate technical leader for this team. This is where you come in! 

We are seeking an experienced Site Reliability Engineer to develop our Kubernetes platform and maintain its stability as well as drive the functionality, capability and efficiency of our infrastructure to the next level. You will be the technical lead inside the cloud native team to design, develop, automate, and continuously improve cluster life cycle management, platform services and pipelines, such as monitoring, alerting, logging, tracing, CI/CD, etc. We are looking for a passionate engineer to join and collaborate with Cloud Native partner teams, Box developer community and open-source communities to advance Kubernetes and Cloud Native technologies.

LOCATION

This role is located at our Redwood City office. Employees based from a Box office hub are expected to visit the office 2+ days per week on average.

WHAT YOU'LL DO 

  • Design, deploy, and maintain Kubernetes clusters across GKE and EKS environments
  • Develop and maintain automation tools for deploying and managing Kubernetes clusters
  • Monitor and troubleshoot GKE and EKS components to ensure high availability and performance in a large-scale enterprise environment
  • Implement security best practices for Kubernetes infrastructure and services
  • Be part of our K8s oncall rotation and participate in incident response and work to reduce the MTTR over time.
  • Proactively and continuously improve the reliability, scalability, and performance of our Kubernetes infrastructure
  • Manage our CI/CD pipelines and other DevOps tools and processes
  • Work closely with developers to ensure that their applications are properly deployed and configured in Kubernetes

WHO YOU ARE

Required skills:

  • 5+ years of experience with deploying, managing and operating multiple GKE/EKS clusters & large-scale container workloads in prod environments.
  • 5+ years experience with managing and deploying IaC via Terraform
  • 3+ years of experience with monitoring & observability tools such as Prometheus, Grafana, and Elasticsearch
  • 3+ years of hands-on experience in automation in declarative fashion using Terraform, Ansible and in imperative languages like shell scripting and python scripting
  • Experience with Kubernetes CNI deployment and troubleshooting, including (but not limited to) the following CNIs: Calico, Cilium
  • Good understanding and working knowledge on CI/CD processes with products like Jira, Git, Artifactory, Jenkins and etc across multiple systems (application & infrastructure)
  • Good understanding and working knowledge about various microservices like spring boot or nodejs

Preferred skills:

  • Kubernetes operator and controller development.
  • FinOps knowledge
  • Experience with performing

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Box

View company profile →