Skip to main content
NVIDIA

Senior Datacenter Product Development Engineer - CA, Santa Clara, United States

NVIDIA US, CA, Santa Clara, United States Full-time 13 days ago
Product Engineering & Quality

$200,000 - $322,000 USD yearly

Originally posted 23 September 2026 by the employer.

This role focuses on the Manufacturing, Test, and Failure Analysis/recovery process for new datacenter servers.

About the role

This role involves owning, developing, and delivering the Manufacturing, Test, and Failure Analysis/recovery process for new datacenter servers and assemblies. This includes incorporating NVIDIA's latest GPU, CPU, and high-speed interconnect technologies.

What you'll do

  • Engage early with hardware, firmware, software engineering, manufacturing partners, and operations units.
  • Develop and implement comprehensive product development plans.
  • Participate in Design for Manufacturing / Test (DFX) process and reviews.
  • Work with factory personnel and engineering teams during prototype/validation builds, driving solutions and reporting status.
  • Review build dashboards, analyze throughput, yields, and failure categories, and publish daily status reports.
  • Develop documentation and SOPs for product assembly, test, debug, and failure analysis.
  • Debug server platforms, work with CM Failure Analysis, Test Engineers, and Operators, and perform rework/retesting.

What you'll need

  • BS/MS or equivalent experience in Electrical Engineering and 12+ years of relevant work experience.
  • Solid understanding of complex compute/AI server system architecture and rack scale systems (servers, backend network switches, TOR switches, power shelves, cabling).
  • Proven skills and experience in design and manufacture of complex hardware platforms.
  • Excellent complex hardware platforms debugging and rational problem solving skills; ability to review logfiles for initial triage.
  • Familiarity with power and thermal system design for datacenter rack server systems.
  • Familiarity with diagnostics/test software as used in a factory setting.
  • CPU, GPU, HBM, PCIe expertise, and comfortable writing Python code and test scripts.
  • Familiarity with linux command line, running commands, and reviewing output.

Skills: datacenter servers, manufacturing, test, failure analysis, DFX, server platforms debugging

Apply now
US, CA, Santa Clara, United States
On-site
Full-time
$200,000 - $322,000 USD yearly
13 days ago

Share this job