Posted at: 17 April
Datacenter Technical Program Manager
Company
NVIDIA Corporation is a Santa Clara-based technology company specializing in designing GPUs and AI solutions for gaming, professional visualization, and cloud services, operating in both B2B and B2C markets globally.
Remote Hiring Policy:
NVIDIA supports flexible remote work arrangements and hires from various regions globally, including the Americas, Europe, Asia, and the Middle East, with roles that may require collaboration across time zones.
Job Type
Full-time
Allowed Applicant Locations
United States
Salary
$168,000 to $258,750 per year
Job Description
NVIDIA is looking for a highly-motivated Technical Program Manager (TPM) to join our Applied Systems Engineering Team to drive datacenter integration for the next generation of NVIDIA AI supercomputing systems. This TPM will play a crucial role throughout the lifecycle of the latest AI systems at scale, from datacenter design and requirements definition, through systems integration of AI clusters into the datacenter environment, and support for these systems as they enter production.This role will drive collaboration between engineering leaders across multiple hardware and software teams, helping us work together to build AI supercomputers for NVIDIA engineers and develop reference architectures to advise customers and partners.What you’ll be doing:Collaborate with outstanding engineers and architects to build and deploy large scale GPU computing systems based on NVIDIA's reference supercomputing architecturesLead the integration of new AI clusters with datacenter facilities with demanding requirements on power, cooling, and instrumentationCoordinate design and fit-out of new datacenter builds, working with both internal engineering teams and external contractorsOwn and produce detailed documentation for the end-to-end process for datacenter fit-out and integrationCommunicate internally with engineering leadership to prioritize and address key issues essential to the success of our largest customersWhat we need to see:BS in Applied Science or Engineering (or equivalent experience)8+ years of overall experienceExperience with high-performance computing systems and GPU clusters deployed in on-premises datacentersA passion for understanding challenging technical problems and driving the process of finding a solutionStrong teamwork and interpersonal skills, to facilitate building a collaborative workflow for coordination between many teamsWays to stand out from the crowd:Understanding of datacenter design, including familiarity with power and cooling technologiesExpertise in system monitoring and instrumentation of large clusters, using technologies such as Prometheus, Grafana, Splunk, Modbus, and BACNetExperience working with the engineering or academic research community supporting high-performance computing or deep learningYour base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until April 20, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.