Jobs and Careers
ME
Network Engineer, Operations & Support
MetaUnited Statesfull_timeVerifiedPosted 9 Jul 2026
About the role
Meta is seeking a Network Engineer to support the reliability and performance of its production network infrastructure. In this role, you will own operational responsibilities across network systems that underpin Meta's global services, driving incident response, change execution, and continuous improvement across production environments. You will work closely with network engineering, deployment, and cross-functional operations teams to ensure network services meet reliability and performance standards at scale.
Responsibilities
Own incident response and troubleshooting for production network issues, including fault isolation, root cause analysis, and coordinating restoration across network layers
* Execute and validate network changes, including device configurations, routing policy updates, and capacity adjustments in production environments
* Monitor production network health using observability tooling and telemetry, identifying anomalies and driving resolution before customer impact occurs
* Participate in on-call rotations and lead operational response to network events, communicating status, risks, and mitigation plans to stakeholders
* Contribute to root cause analysis and post-incident reviews, identifying systemic gaps and driving corrective actions to reduce repeat incidents
* Partner with network engineering and deployment teams to support network builds, expansions, and migrations, validating operational readiness prior to cutover
* Identify opportunities to improve operational workflows, runbooks, and automation to increase reliability and reduce time-to-resolution
* Maintain accurate documentation of network topology, configurations, and operational procedures to support handoffs and audit requirements
* Advise cross-functional partners on network operational constraints and tradeoffs when evaluating changes or new deployments
* Leverage AI-integrated workflows to accelerate diagnostics, streamline documentation, and improve operational efficiency across the team
Qualifications
Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
* 6+ years of experience in production network operations, including hands-on troubleshooting of routing, switching, and transport protocols in large-scale environments
* Experience with network protocols including BGP, OSPF, IS-IS, MPLS, and Ethernet at scale
* Experience operating in on-call, incident-driven environments with accountability for restoration timelines and stakeholder communication
* Experience executing and validating network configuration changes in production environments with change management discipline
* Experience collaborating with engineering and cross-functional teams to align on operational priorities and communicate technical tradeoffs Experience supporting large-scale data center or backbone network environments, including optical transport and interconnect operations
* Demonstrated ability to integrate AI tools to optimize operational workflows and drive measurable improvements in efficiency or quality
* Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
* Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
* Familiarity with network observability platforms, telemetry pipelines, and alerting frameworks used to monitor production infrastructure health
* Experience with network automation and scripting (e.g., Python, Ansible) to reduce manual operational toil and improve repeatability
* Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
Responsibilities
Own incident response and troubleshooting for production network issues, including fault isolation, root cause analysis, and coordinating restoration across network layers
* Execute and validate network changes, including device configurations, routing policy updates, and capacity adjustments in production environments
* Monitor production network health using observability tooling and telemetry, identifying anomalies and driving resolution before customer impact occurs
* Participate in on-call rotations and lead operational response to network events, communicating status, risks, and mitigation plans to stakeholders
* Contribute to root cause analysis and post-incident reviews, identifying systemic gaps and driving corrective actions to reduce repeat incidents
* Partner with network engineering and deployment teams to support network builds, expansions, and migrations, validating operational readiness prior to cutover
* Identify opportunities to improve operational workflows, runbooks, and automation to increase reliability and reduce time-to-resolution
* Maintain accurate documentation of network topology, configurations, and operational procedures to support handoffs and audit requirements
* Advise cross-functional partners on network operational constraints and tradeoffs when evaluating changes or new deployments
* Leverage AI-integrated workflows to accelerate diagnostics, streamline documentation, and improve operational efficiency across the team
Qualifications
Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
* 6+ years of experience in production network operations, including hands-on troubleshooting of routing, switching, and transport protocols in large-scale environments
* Experience with network protocols including BGP, OSPF, IS-IS, MPLS, and Ethernet at scale
* Experience operating in on-call, incident-driven environments with accountability for restoration timelines and stakeholder communication
* Experience executing and validating network configuration changes in production environments with change management discipline
* Experience collaborating with engineering and cross-functional teams to align on operational priorities and communicate technical tradeoffs Experience supporting large-scale data center or backbone network environments, including optical transport and interconnect operations
* Demonstrated ability to integrate AI tools to optimize operational workflows and drive measurable improvements in efficiency or quality
* Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
* Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
* Familiarity with network observability platforms, telemetry pipelines, and alerting frameworks used to monitor production infrastructure health
* Experience with network automation and scripting (e.g., Python, Ansible) to reduce manual operational toil and improve repeatability
* Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s