Network Production Engineer, Delivery Engineering
Meta Menlo Park, California, United States · $184K–$257K/yr
Software Development · 10,001+ employees
About the role
You will design, build, and operate global datacenter networks while developing automation tools to ensure maximum reliability and scalability. The role involves leading cross-functional projects to improve operational efficiency and participating in on-call rotations to resolve production challenges.
What they look for
Requirements
Candidates must have a bachelor's degree in a technical field and at least 8 years of experience in developing scalable systems or networks. Proficiency in higher-level programming languages and experience with network device configuration are required.
Benefits
Full description
The Network Infrastructure team is responsible for designing, building, and operating one of the largest networks in the world. Networking is at the core of all Meta products and experiences, and we are looking for Network Production Engineers who are interested in solving complex technical challenges in the Backbone, Datacenter, and AI Network domains.
The scale of the network and its continuous expansion presents an opportunity to work on and solve interesting engineering challenges in the datacenter network domain. We create new and innovative ways of designing and operating our global datacenter networks and do it at scale with efficiency.
Production Network Engineers at Meta are hybrid software and network engineers who design, build, and operate our worldwide network. This team owns the complete lifecycle of the network, which includes areas of planning, design, product definition, QA, deployment, and monitoring. Simple, elegant, and scalable network design, automation, and data analytics are the key to meeting our demand. In this role, you will be part of a team that is responsible for conceiving design solutions, developing and deploying network software, systems, and tools that keep the network operating at maximum reliability, scalability, and efficiency.
This role offers an opportunity to solve scaling challenges supporting billions of people using our family of apps, as well as challenges in AI workloads that power new Meta products.
Responsibilities
- Establish and implement global best practices and design new, scalable network solutions
- Conceptualize, build, and maintain automation and tools to support new product introductions, network deployment, release engineering, and operations
- Design and develop solutions that scale across a variety of hardware platforms of network equipment
- Lead enhancements of automation for continuous integration, validations, testing infrastructure, release, and configuration management across our global backbone, data center, and edge networks
- Work closely with our hardware, software, and sourcing teams to develop new networking solutions and influence the future of networking and its associated infrastructure
- Conduct thorough investigations into complex technical issues across networks, ranging from automated tooling to hardware failures and network issues
- Develop operational process improvements and implement them in scalable, automated workflows to enhance operational efficiency
- Help increase operational efficiency between peers and cross-functional teams by identifying roadblocks, designing and delivering automation solutions, and driving change
- Proactively find gaps that impact multiple teams, come up with the execution plan, drive the project, and influence other teams to reach their goals
- Participate in an oncall rotation to learn from real-world production challenges and take the lessons to improve current and future generation products
- Contribute to the team's growth and development through peer mentorship
Minimum Qualifications
- Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
- Bachelor's degree in Computer Science, Computer Engineering, a relevant technical field, or equivalent practical experience
- 8+ years of relevant experience in developing scalable and reliable systems and/or networks
- Experience coding in higher-level languages (e.g., Python, C++, Go, Rust, etc.)
- Experience working with software frameworks and APIs
- Engineering degree, or a related technical discipline, or equivalent experience
- Experience in the configuration and maintenance of network devices and NMS systems, or applications such as web servers, load balancers, relational databases, storage systems, and messaging systems
- Experience in developing and understanding network device configurations for at least one vendor (Juniper, Cisco, Arista, Brocade, etc.)
Preferred Qualifications
- Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
- Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
- Experience with network physical infrastructure, cables and connectors, and their implementation
- Understanding of different Optics and internals of a switch ASIC
- Familiarity with Linux-based systems
- Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
- Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
- Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)
- Experience in adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
$184,000/year to $257,000/year + bonus + equity + benefits