Jobs and Careers
BI

Senior Infrastructure Engineer - InfraOps

BitGo
Palo Alto, United Statesfull_timeVerifiedPosted 10 Dec 2025
💰 $220,000/yr($180,000/yr$220,000/yr)

About the role

BitGo is the leading infrastructure provider of digital asset solutions, delivering custody, wallets, staking, trading, financing, and settlement services from regulated cold storage. Since our founding in 2013, we have focused on enabling our clients to securely navigate the digital asset space. With a global presence and multiple Trust companies, BitGo serves thousands of institutions, including many of the industry's top brands, exchanges, and platforms, and millions of retail investors worldwide. As the operational backbone of the digital economy, BitGo handles a significant portion of Bitcoin network transactions and is the largest independent digital asset custodian, and staking provider, in the world. For more information, visit www.bitgo.com.

BitGo is seeking a highly experienced DevOps/SRE Engineer to lead and architect our highly available digital asset infrastructure on Kubernetes. This pivotal role demands a proven leader capable of driving strategic initiatives, guaranteeing robust performance, and ensuring unparalleled reliability across our global operations. The successful candidate will proactively define and implement advanced monitoring and security frameworks, significantly enhancing network integrity, optimizing operational efficiency, and delivering a stable, cost-efficient, and highly scalable platform that empowers our developers and solidifies user trust. This position blends deep expertise in both web2 and web3 technologies, directly contributing to the security and scalability of over $100 billion in digital assets and shaping the future of our infrastructure.

This role is on-site in Palo Alto (CA, US) or San Francisco (CA, US) and requires participation in a 24/7 on-call rotation, including weekend coverage.

Responsibilities:

  • Architect, design, and champion the adoption of cutting-edge Infrastructure as Code (IaC) tooling and automation solutions across the organization, setting best practices and driving innovation.
  • Lead cross-functional collaborations with engineering and business teams to proactively identify and address complex infrastructure requirements, ensuring the delivery of highly scalable, resilient, and performant solutions that align with strategic business objectives.
  • Evaluate, integrate, and strategically deploy advanced open-source and commercial tools to significantly enhance our security posture, infrastructure capabilities, and consistently meet evolving business demands.

Drive cost optimization initiatives across cloud infrastructure including capacity planning, resource right-sizing, and reserved instance strategies.

  • Define, own, and execute the technical roadmaps for critical system components, ensuring seamless alignment with organizational strategic objectives and long-term vision.
  • Drive operational excellence, reliability, and performance of critical client and internal systems through proactive project leadership, sophisticated incident response, and mentorship within on-call rotations.

Required:

  • Extensive and demonstrable experience securing, scaling, and operating multiple complex environments on Kubernetes, coupled with deep expertise in associated tooling (ArgoCD, GitOps, Grafana) and advanced Terraform implementations.
  • Strong Linux systems administration skills, including performance tuning, troubleshooting, and security hardening.
  • Deep understanding of networking fundamentals: VPCs, security groups, load balancers, CNI plugins, and network policy management.
  • Proficiency in at least one high-level programming language, preferably Go, with strong bash scripting capabilities.
  • Deep operational expertise with relational and NoSQL databases (including advanced connection maintenance, intricate slow query analysis,index management) as well as large-scale object storage solutions.
  • Proven track record with Github Actions and architecting robust CI/CD pipelines.
  • Experience with major public cloud providers and advanced container orchestration/fleet management strategies, including sophisticated cost optimisation and capacity planning. We operate primarily on AWS but believe skills in one cloud provider translate to others.
  • Minimum of five years of progressive experience building, securing, and maintaining complex, high-stakes production systems.
  • Exceptional analytical and communication skills, coupled with a highly inquisitive and proactive disposition.

Preferred:

  • Secur

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

BitGo

View company profile →