Senior Network Developer, Infrastructure & Observability
CoreWeaveAbout the role
CoreWeave is the AI Hyperscaler™, delivering a cloud platform of cutting edge services powering the next wave of AI. Our technology provides enterprises and leading AI labs with the most performant, efficient and resilient solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers covering every region of the US and across Europe. CoreWeave was ranked as one of the TIME100 most influential companies of 2024.
As the leader in the industry, we thrive in an environment where adaptability and resilience are key. Our culture offers career-defining opportunities for those who excel amid change and challenge. If you’re someone who thrives in a dynamic environment, enjoys solving complex problems, and is eager to make a significant impact, CoreWeave is the place for you. Join us, and be part of a team solving some of the most exciting challenges in the industry.
CoreWeave powers the creation and delivery of the intelligence that drives innovation.
What You’ll Do
We’re seeking a Senior Network Developer to join our Network Development team. Our mission is to design, automate, and enhance the network infrastructure and network observability systems of CoreWeave’s entire GPU cloud network. As part of this role, you may focus on network infrastructure automation, building software-driven solutions to streamline deployments, operating our network observability stack, developing real-time monitoring and telemetry systems to ensure seamless network performance.You’ll play a key role in automating, optimizing, and scaling our network, working closely with Network Engineering, Site Reliability, and Platform teams to build reliable, self-healing, and intelligent network solutions. Your mission? To make our network highly automated, observable, and self-sufficient—ensuring smooth operations while eliminating human intervention wherever possible.
- Design, develop, and maintain software-driven solutions for network automation and observability, ensuring seamless operations at scale.
- Write code in Python, Go, and Bash to automate network provisioning, monitoring, and performance optimization.
- Build and manage network telemetry pipelines, using tools like Prometheus, Grafana, Alertmanager, gNMI, and SNMP to provide deep insights into network health.
- Implement Zero Touch Provisioning (ZTP) and infrastructure-as-code methodologies to streamline network configuration and deployment.
- Develop and integrate network monitoring tools, aggregating logs, metrics, and events from multiple platforms (e.g., Arista EOS, NVIDIA Cumulus Linux, Nokia SR OS, SR Linux).
- Collaborate with security and platform teams to containerize network applications and integrate observability across infrastructure.
- Participate in on-call rotations, troubleshooting network-related issues, and supporting operations teams with deep technical insights.
- Engage in code reviews, design discussions, RFCs, and architectural decisions to ensure high-quality software development practices.
- Guide junior engineers and foster a culture of collaboration, learning, and continuous improvement.
Investing in our people is one of our top priorities, and we value candidates who can bring their diversified experiences to our teams. Here are some qualities we’ve found compatible with our team. We'd love to talk about whether this aligns with your experience and interests and what you’re excited to work on next.
Who You Are
Minimum Qualifications
- 7+ years of experience as a Network Engineer, Software Developer, SRE, or Systems Administrator, ideally in large-scale enterprise or cloud settings.
- Expertise in Python, Go, and Bash, with experience in automation, scripting, and infrastructure-as-code.
- Experience with Prometheus, Grafana, Alertmanager, gNMI, SNMP, Ansible, Jinja, and NetBox.
- Solid Linux networking knowledge with hands-on experience in routing, switching, and troubleshooting.
- Experience working with networking platforms such as:
- NVIDIA Cumulus Linux
- Nokia SR OS and SR Linux
- Arista EOS
- Comfortable working in a Kubernetes-based environment, deploying containerized workloads for network services.
- Passion for automation-first approaches, eliminating manual work through software-driven solutions.
- A collaborative mindset, able to mentor team members, share knowledge, and contribute to team growth.
Our compensation reflects the cost of labor across several US geographic markets. The base pay for this position ranges from $175,000-$220,000. Pay is based on a number of factors including market location and may vary depending on job-related knowledge, skills, and experience.
What We Offer
The ra
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s