NVIDIA Logo

NVIDIA

Senior Datacenter Technical Program Manager, At-Scale AI Clusters

Reposted 25 Days Ago
In-Office or Remote
Hiring Remotely in CA, USA
168K-259K Annually
Senior level
In-Office or Remote
Hiring Remotely in CA, USA
168K-259K Annually
Senior level
The role involves integrating AI supercomputing systems into datacenters, collaborating with engineering teams, and producing documentation for processes related to datacenter integration and fit-out.
The summary above was generated by AI

NVIDIA is looking for a highly-motivated Technical Program Manager (TPM) to join our Applied Systems Engineering Team to drive datacenter integration for the next generation of NVIDIA AI supercomputing systems. This TPM will play a crucial role throughout the lifecycle of the latest AI systems at scale, from datacenter design and requirements definition, through systems integration of AI clusters into the datacenter environment, and support for these systems as they enter production.

This role will drive collaboration between engineering leaders across multiple hardware and software teams, helping us work together to build AI supercomputers for NVIDIA engineers and develop reference architectures to advise customers and partners.

What you’ll be doing:

  • Collaborate with outstanding engineers and architects to build and deploy large scale GPU computing systems based on NVIDIA's reference supercomputing architectures

  • Lead the integration of new AI clusters with datacenter facilities with demanding requirements on power, cooling, and instrumentation

  • Coordinate design and fit-out of new datacenter builds, working with both internal engineering teams and external contractors

  • Own and produce detailed documentation for the end-to-end process for datacenter fit-out and integration

  • Communicate internally with engineering leadership to prioritize and address key issues essential to the success of our largest customers

What we need to see:

  • BS in Applied Science or Engineering (or equivalent experience)

  • 8+ years of overall experience

  • Experience with high-performance computing systems and GPU clusters deployed in on-premises datacenters

  • A passion for understanding challenging technical problems and driving the process of finding a solution

  • Strong teamwork and interpersonal skills, to facilitate building a collaborative workflow for coordination between many teams

Ways to stand out from the crowd:

  • Understanding of datacenter design, including familiarity with power and cooling technologies

  • Expertise in system monitoring and instrumentation of large clusters, using technologies such as Prometheus, Grafana, Splunk, Modbus, and BACNet

  • Experience working with the engineering or academic research community supporting high-performance computing or deep learning

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until June 12, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar Jobs

2 Hours Ago
Remote or Hybrid
United States
81K-109K Annually
Junior
81K-109K Annually
Junior
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Drive aftersales performance across a defined region by partnering with dealership leadership to execute Customer Care & Aftersales initiatives, grow revenue, improve customer retention and NPS, analyze performance data, develop business plans, resolve aftersales issues (warranty, goodwill, technical support), and deliver targeted operational improvements through frequent dealer visits.
Top Skills: Automotive Parts And Service SystemsData Analytics ToolsDealer Operating ReportsFixed Ops Analysis ToolsExcelSales Reporting Tool (Srt)
2 Hours Ago
Remote or Hybrid
United States
76K-114K Annually
Mid level
76K-114K Annually
Mid level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
The District Manager, OnStar will drive OnStar activation targets, revenue growth, train dealerships, and conduct performance analysis to improve sales and customer onboarding.
Top Skills: B2B SalesMarketing
2 Hours Ago
Remote or Hybrid
95K-110K Annually
Senior level
95K-110K Annually
Senior level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
The HR Manager will support West Coast Studio Operations by managing employee relations, fostering a positive work environment, and driving HR initiatives such as onboarding, performance management, and talent development.

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account