Jobs and Careers
ME
Network Production Engineer
MetaMenlo Park, United Statesfull_timeVerifiedPosted 27 Jan 2025
💰 $251,000/yr($177,000/yr – $251,000/yr)
About the role
Description
The Network Infrastructure team is responsible for designing, building and operating one of the largest networks in the world. Networking is at the core of all Meta products and experiences, and we are looking for seasoned Production Engineers who are interested in solving complex technical challenges in the Backbone or Datacenter Network domains.
Production Network Engineers at Meta are hybrid software and network engineers who keep reliability and scalability in mind as they work on different parts of the lifecycle (designing, building, and operating our worldwide network). This role offers an opportunity to solve the scaling challenges of supporting billions of people using our family of apps; to cutting-edge challenges in AI workloads that power new Meta products.
We make the global Datacenter front end network fleet reliable and available for all Infrastructure services to use. Contribute to the company's mission by operating and bringing to production all the new network products that enable networking for AI training and Inference.
A Network Production Engineer in this role would support leading Meta’s server fleet connections to the network, specifically at the NIC layer of the fleet. They would operate at a unique intersection of low level systems engineering and handle the challenge of operating a massively distributed fleet that is uniquely available at Meta. Network Production Engineers are exposed to bleeding edge technology being developed internally at Meta to optimize our server network communication stack.Network Production Engineer Responsibilities
The Network Infrastructure team is responsible for designing, building and operating one of the largest networks in the world. Networking is at the core of all Meta products and experiences, and we are looking for seasoned Production Engineers who are interested in solving complex technical challenges in the Backbone or Datacenter Network domains.
Production Network Engineers at Meta are hybrid software and network engineers who keep reliability and scalability in mind as they work on different parts of the lifecycle (designing, building, and operating our worldwide network). This role offers an opportunity to solve the scaling challenges of supporting billions of people using our family of apps; to cutting-edge challenges in AI workloads that power new Meta products.
We make the global Datacenter front end network fleet reliable and available for all Infrastructure services to use. Contribute to the company's mission by operating and bringing to production all the new network products that enable networking for AI training and Inference.
A Network Production Engineer in this role would support leading Meta’s server fleet connections to the network, specifically at the NIC layer of the fleet. They would operate at a unique intersection of low level systems engineering and handle the challenge of operating a massively distributed fleet that is uniquely available at Meta. Network Production Engineers are exposed to bleeding edge technology being developed internally at Meta to optimize our server network communication stack.Network Production Engineer Responsibilities
- Conceive, develop, and deploy systems and tools to keep the network running reliably and efficiently.
- Demonstrated experience on complex technical issues across networks, ranging from automated tooling to hardware failures and network issues.
- Develop documentation, write and review code, and debug the hardest problems, live, on some of the largest and most complex networks and systems in the world.
- Participate in a weekly on-call rotation and be an escalation contact for service incidents.
- Lead projects to address hard technical challenges, directly contributing to roadmaps and partner alongside the best engineers in the industry to develop reliable and scalable network and software solutions for our global data center fleet.
- Proactively find gaps that impact multiple teams, come up with the execution plan, and drive the project directly and through influence of other teams.
- Contribute to the overall team growth and development through peer mentorship.
- Collaborate effectively with team members and partners in other global regions (i.e. EMEA), including flexibility to work during global-friendly hours as needed
- Global travel 10-15% of the time.
- Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience.
- 7+ years of relevant experience developing scalable and reliable systems and/or networks
- Experience coding in higher-level languages (e.g., Python, C++, Go, etc.)
- Experience in developing and understanding network device configuration for at least one vendor (Juniper, Cisco, Arista, Brocade, etc.)
- Experience in configuration and maintenance of network devices and NMS systems, or applications such as web servers, load balancers, relational databases, storage systems and messaging systems
- Experience learning software, frameworks and APIs
- BS or MS in Computer Science, Computer Engineering, or Network Engineering
- 7+ in TCP/IP and IPv6
- 10+ years experience designing and operating data center networks
- 7+ years experience in one or more of BGP, MPLS, ISIS or similar routing protocols - knowledge in typical configurations and performance tuning
- 7+ years experience coding in higher-level languages (e.g., Python, C++, Go, etc.)
- Experience understanding network hardware and topology failures
- Understanding of AI training workloads and demands they exert on networks
- Experience leading and setting technical direction for a team of 6+ engineers
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s