Skip to main content
NVIDIA

NIM Solution Architect

NVIDIA 2 Locations Full-time 15 days ago
Applications & FAE

Originally posted 21 September 2026 by the employer — open 0 days.

Focus on NVIDIA Inference Microservices (NIM).

About the role

This role focuses on NVIDIA Inference Microservices (NIM), inference/RL rollout performance, and AI workflow enablement. The position involves model optimization, inference infrastructure, and customer solution delivery for LLM, VLM, and other generative AI workloads.

What you'll do

  • Drive implementation, deployment, and optimization of NVIDIA Inference Microservices (NIM) solutions for enterprise and industry AI workloads.
  • Package and serve open-source, NVIDIA, and customer-proprietary models through NIM with standardized, containerized APIs.
  • Optimize high-volume inference and rollout workloads for LLMs and VLMs.
  • Evaluate and tune NIM models.
  • Deliver technical projects, demos, and client support tasks.
  • Provide technical support and guidance to customers, facilitating adoption and implementation of NVIDIA technologies and products.

What you'll need

  • Master’s degree or higher in Computer Science, Machine Learning, Electrical Engineering, Mathematics, or a related technical field, or equivalent experience.
  • 2+ years of hands-on experience in machine learning engineering, applied research, LLM/VLM inference, or RL rollout.
  • Production-quality Python and PyTorch skills, including distributed GPU training, solution, profiling, debugging, and memory optimization.
  • Working knowledge of transformer architectures, performance optimization, rollout sampling strategies, structured generation, and model-quality evaluation.

Skills: NIM Solution Architect, customer solution delivery, LLM, VLM, inference infrastructure, technical support and guidance

Market context

This role has been open 0 days — well below the 41-day median for Applications FAE roles.

Applications FAE · AI ML Hardware

Open roles in category
96
Median days open
41 d
Median salary
$261k
See the full market breakdown ▾Category comparison, skills in demand, and who else is hiring
How NVIDIA compares in Applications FAE hiring
Metric NVIDIA All employers we track in this specialty (96 roles · 18 employers)
Open roles in this specialty 61 96
Median days open 31 d 41 d (−10 d vs this employer)
Median salary (USD postings) — $261k

Skills observed across this category: NIM Solution Architect, customer solution delivery, LLM, VLM, inference infrastructure, technical support and guidance

Who's hiring in this category

  • NVIDIA (this employer) · 58 open roles · median 37 d
  • AMD · 10 open roles · median 44 d
  • NXP Semiconductors · 6 open roles · median 12 d
  • Tenstorrent · 6 open roles · median 116 d
  • Qualcomm · 3 open roles · median 76 d
  • Arm Holdings · 1 open role · median 87 d

How we counted: 96 open Applications FAE (AI ML Hardware) roles from 18 employers tracked in the SemiconductorJobs index, counted 21 Sept 2026. Specialty figures count only roles carrying this exact specialty label, so an employer's related work in neighbouring specialties is not included there. Figures refresh nightly.

Apply now
2 Locations
On-site
Full-time
15 days ago

Share this job