$148,000 - $287,500 USD yearly
Originally posted 1 October 2026 by the employer.
Join the team building some of the largest and fastest AI Compute systems in the world.
About the role
This role focuses on deploying, managing, and validating AI Compute/HPC infrastructure.
What you'll do
- Deploy, manage, and validate AI Compute/HPC infrastructure in Linux-based environments. - Act as a domain expert for customers during planning and implementation. - Create handover documentation and perform knowledge transfers for customer support. - Provide feedback to internal teams, including bug reporting, documenting workarounds, and suggesting improvements.
What you'll need
- 8+ years providing in-depth support and deployment services for hardware and software products. - Knowledge and experience with Linux system administration, process management, package management, task scheduling, kernel management, boot procedures/troubleshooting, performance reporting/optimization/logging, network-routing/advanced networking. - Cluster management and provisioning technologies for bare-metal servers. - Four-year degree in Computer Science, Electrical or Computer Engineering, or equivalent experience. - Scripting proficiency (Bash, Python, Ansible). - Experience with schedulers such as SLURM, LSF, UGE. - Ability to travel to customer sites within the United States up to 20% of the time. - Experience with benchmarking tools such as HPL, NCCL tests, MLPerf, and Kubernetes experience.
Nice to have
- InfiniBand experience. - Experience with GPU focused hardware/software. - Experience with MPI (Message Passing Interface). - Storage technologies such as Lustre or GPFS. - Familiarity with OEM GPU platforms.
Skills: AI Compute, HPC infrastructure, Linux system administration, Cluster management, GPU focused hardware/software
This role has been open 0 days — well below the 51-day median for Infrastructure/Platform Software roles.
Infrastructure/Platform Software · Infrastructure Platform
|
Open roles in category
866
|
Median days open
51 d
|
Median salary
$236k
|
See the full market breakdown ▾Category comparison, skills in demand, and who else is hiring
| Metric | NVIDIA | All employers we track in this specialty (866 roles · 84 employers) |
|---|---|---|
| Open roles in this specialty | 245 | 866 |
| Open roles in the wider Software, Firmware & Systems family | 961 | 5930 · 145 employers |
| Median days open | 43 d | 51 d (−8 d vs this employer) |
| Median salary (USD postings) | — | $236k |
Skills observed across this category: AI Compute, HPC infrastructure, Linux system administration, Cluster management, GPU focused hardware/software
Who's hiring in this category
- NVIDIA (this employer) · 241 open roles · median 43 d
- Qualcomm · 57 open roles · median 65 d
- AMD · 46 open roles · median 23 d
- Cerebras · 39 open roles · median 67 d
- Graphcore · 37 open roles · median 94 d
- Micron Technology · 30 open roles · median 51 d
How we counted: 866 open Infrastructure/Platform Software (Infrastructure Platform) roles from 84 employers tracked in the SemiconductorJobs index, counted 2 Oct 2026. Specialty figures count only roles carrying this exact specialty label, so an employer's related work in neighbouring specialties is not included there — it is counted in the wider Software, Firmware & Systems family row. Figures refresh nightly.