Jobs and Careers
ME
Network Production Engineer - Backbone
MetaMenlo Park, United Statesfull_timeVerifiedPosted 15 May 2024
💰 $251,000/yr($177,000/yr – $251,000/yr)
About the role
You will be joining the team that is responsible for the end-to-end health (performance and reliability) of Meta's backbone networks. You will build tools and use automation to efficiency scale how we mitigate real-time impact to the network, identify and investigate long-term trends into performance and risks in our backbone, and drive innovative solutions to monitor and improve Meta's current and future backbone network products.
Our backbones continue to rapidly expand globally, driven most recently through the network demands that our AGI journey brings. We support both our "Classic Backbone", that transports traffic destined to people using Meta's products, and our "Express Backbone", that handles machine to machine traffic between our Data Centers.
Engineers that typically thrive in this role are hybrid software and network engineers who are curious about how systems work, how they fail, and how we can increase their reliability. You have the opportunity to dig into interesting challenges in the networking and software domains, at a scale that offers new challenges on a daily basis.Network Production Engineer - Backbone Responsibilities
Our backbones continue to rapidly expand globally, driven most recently through the network demands that our AGI journey brings. We support both our "Classic Backbone", that transports traffic destined to people using Meta's products, and our "Express Backbone", that handles machine to machine traffic between our Data Centers.
Engineers that typically thrive in this role are hybrid software and network engineers who are curious about how systems work, how they fail, and how we can increase their reliability. You have the opportunity to dig into interesting challenges in the networking and software domains, at a scale that offers new challenges on a daily basis.Network Production Engineer - Backbone Responsibilities
- Write and review code, develop documentation and capacity plans, and debug the hardest problems, live, on some of the largest and most complex networks and systems in the world
- Participate in a weekly on-call rotation and be an escalation contact for service incidents
- Perform deep dives on complex technical issues across networks, ranging from automated tooling to hardware failures and network issues
- Manage and maintain multi-vendor, multi-protocol backbone and edge networks
- Analyze data to diagnose and identify root causes to network issues
- Define, develop, and optimize automated network monitoring systems to mitigate and remediate network events
- Proactively find gaps that impact multiple teams, come up with the execution plan, and drive the project directly and through influence of other teams
- Contribute to team growth and development through peer mentorship, set technical direction for the team
- Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience.
- 2+ years experience coding in higher-level languages (e.g., Python, C++, Go, etc.)
- 8+ years experience in one or more of BGP, MPLS, ISIS or similar routing protocols - knowledge in typical configurations and performance tuning
- 8+ years experience understanding and mitigating network hardware and topology failures
- Experience in configuration and maintenance of network devices and NMS systems, or applications such as web servers, load balancers, relational databases, storage systems and messaging systems
- Experience learning software, frameworks and APIs
- 5+ years experience developing and understanding network device configuration for at least one vendor (Juniper, Cisco, Arista, Brocade, etc.)
- Expert knowledge in routing and switching - hardware design and knowledge of forwarding and data planes
- BS or MS in Computer Science, Computer Engineering, or Network Engineering
- Expert knowledge of TCP/IP and IPv6
- Expert knowledge of traffic engineering and performance tuning in backbone networks
- 8+ years experience coding in higher-level languages (e.g., Python, C++, Go, etc.)
- Experience operating and designing SDN-based backbone networks
- Experience working in a multi-vendor network environment
- Experience with developing distributed systems and operating them at scale
- Experience with automation frameworks and tools such as Ansible, Puppet, or Chef
- Experience leading and setting technical direction for a team of 10+ engineers
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s