Senior Infrastructure Engineer - Infrastructure Security and Core Services
NVIDIA Santa Clara, California, United States · $208K–$334K/yr
Computer Hardware Manufacturing · 10,001+ employees
About the role
Lead the architecture, design, and deployment of global-scale manufacturing sites and their connectivity to AI factories. Implement and refine compute, storage, security, and telemetry practices to ensure high availability and performance for mission-critical manufacturing workloads.
What they look for
Requirements
Requires a Master's or PhD in a technical field and over 12 years of experience in building and managing large-scale hybrid networks. Candidates must possess expert knowledge in networking, compute, and storage technologies, along with strong automation scripting skills.
Benefits
Full description
NVIDIA has been redefining computer graphics, PC gaming, and accelerated computing for 30 years. It’s an outstanding legacy of innovation that’s motivated by extraordinary technology—and outstanding people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, encouraging environment where everyone is inspired to do their best work. Come join our team and see how you can make a lasting impact on the world!
The Contract Manufacturing infrastructure team is seeking experienced candidates who have a passion for infrastructure security, compute services and storage needs that are designed for manufacturing workloads. This candidate would need to redesign the NVIDIA manufacturing sites core services, enable new sites and troubleshoot incidents quickly knowing the revenue impact of a factory line being down. This is a hands-on engineer position focused on the development and deployment of ultra-high-speed, resilient, and scalable core services for GPU-accelerated OT and IT environments, manufacturing sites. Outstanding problem-solving abilities and a comprehensive understanding of the computer, storage, routing, switching, automation and deep understanding of fundamental network theory is also critical to your success at NVIDIA.
What you will be doing
- Lead the architecture, design, and deployment of global‑scale manufacturing sites and their connectivity to AI factories and offices. Architect and build CPU‑based compute, storage, and GPU/HPC clusters.
- Design high‑performance OT and IT networks both for NVIDIA and Partner connectivities to support both general compute workloads and GPU‑dense AI/ML training and inference environments.
- Partner with systems, Operations teams, supply chain partners, OS, GPU, storage, and HPC product teams to deliver scalable, highly available network architectures and connectivity solutions that can evolve with rapid growth in both compute and GPU capacity.
- Implement and refine compute, storage and security, telemetry, and performance‑engineering practices across the infrastructure to detect issues early and continually improve end‑to‑end application experience.
- Manage life cycle management, come up with design infrastructures, revenue generating manufacturing sites.
- Define and enforce security, compliance, and reliability standards for all infrastructure components supporting mission critical manufacturing, R&D workloads, NPI designs in a secure way.
- Collaborate with Operations, Manufacturing Partners and engineering teams to develop “NVIDIA on NVIDIA” reference architectures and best‑practice solutions for large‑scale compute and AI data designs, using NVIDIA products like Mellanox and GPUs.
- Be able to troubleshoot quickly DNS, DHCP, and other full connectivity stack infrastructures.
- Experience with Test Engineering topologies and experience supporting infrastructure on the factory floor.
What we need to see
- MS or PhD in Electrical Engineering, Computer Science, Computer Engineering, Artificial Intelligence, Data Science, Mathematics, Statistics, or equivalent experience.
- 12+ years of experience in building, managing and supporting large scale hybrid networks, developing automation pipelines with Python, Ruby, Go or other languages used in infrastructure automation.
- Expert in networking technologies especially Mellanox, Compute technologies - Dell, Openshift, cloud compute services like Dell etc and Storage Services - Pure, Netapp
- Experience designing compute and storage architectures for data centers, offices, Manufacturing environments and Labs
- Experience building data lakes, caching layers and cloud based infrastructure services with the ability to take requirements into new designs as businesses transform and product requirements change
- Experience with designing Test Automation infrastructures in manufacturing and strong scripting skills
- Designing both CPU and GPU workloads
NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The NVIDIA factories that produce GPUs, our inventions, serve as the visual cortex of modern computers and are at the heart of our products and services. We need to build secure networks for these critical revenue generating manufacturing sites. Our work opens up new universes to explore, enables outstanding creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. NVIDIA is looking for exceptional people like you to help us accelerate the next wave of artificial intelligence.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 208,000 USD - 333,500 USD.
You will also be eligible for equity and benefits.
Applications for this job will be accepted at least until September 13, 2026.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.