Senior Datacenter Technical Program Manager, At-Scale AI Clusters
Nvidia · Santa Clara, CA
About this role
Nvidia is hiring a senior-level Technical Program Manager in the software engineering function based in Santa Clara, CA. The posting calls out experience with Deep Learning, Splunk, Prometheus, Grafana.
- Role
- Technical Program Manager
- Function
- software engineering
- Level
- senior
- Track
- Individual contributor
- Employment
- Full-time
- Location
- Santa Clara, CA
- Posted
- May 15, 2026
More roles at Nvidia
Job description
from Nvidia careersNVIDIA is looking for a highly-motivated Technical Program Manager (TPM) to join our Applied Systems Engineering Team to drive datacenter integration for the next generation of NVIDIA AI supercomputing systems. This TPM will play a crucial role throughout the lifecycle of the latest AI systems at scale, from datacenter design and requirements definition, through systems integration of AI clusters into the datacenter environment, and support for these systems as they enter production.
This role will drive collaboration between engineering leaders across multiple hardware and software teams, helping us work together to build AI supercomputers for NVIDIA engineers and develop reference architectures to advise customers and partners.
What you’ll be doing:
Collaborate with outstanding engineers and architects to build and deploy large scale GPU computing systems based on NVIDIA's reference supercomputing architectures
Lead the integration of new AI clusters with datacenter facilities with demanding requirements on power, cooling, and instrumentation
Coordinate design and fit-out of new datacenter builds, working with both internal engineering teams and external contractors
Own and produce detailed documentation for the end-to-end process for datacenter fit-out and integration
Communicate internally with engineering leadership to prioritize and address key issues essential to the success of our largest customers
What we need to see: