Originally posted 21 September 2026 by the employer — open 0 days.
Focus on NVIDIA Inference Microservices (NIM).
About the role
This role focuses on NVIDIA Inference Microservices (NIM), inference/RL rollout performance, and AI workflow enablement. The position involves model optimization, inference infrastructure, and customer solution delivery for LLM, VLM, and other generative AI workloads.
What you'll do
- Drive implementation, deployment, and optimization of NVIDIA Inference Microservices (NIM) solutions for enterprise and industry AI workloads.
- Package and serve open-source, NVIDIA, and customer-proprietary models through NIM with standardized, containerized APIs.
- Optimize high-volume inference and rollout workloads for LLMs and VLMs.
- Evaluate and tune NIM models.
- Deliver technical projects, demos, and client support tasks.
- Provide technical support and guidance to customers, facilitating adoption and implementation of NVIDIA technologies and products.
What you'll need
- Master’s degree or higher in Computer Science, Machine Learning, Electrical Engineering, Mathematics, or a related technical field, or equivalent experience.
- 2+ years of hands-on experience in machine learning engineering, applied research, LLM/VLM inference, or RL rollout.
- Production-quality Python and PyTorch skills, including distributed GPU training, solution, profiling, debugging, and memory optimization.
- Working knowledge of transformer architectures, performance optimization, rollout sampling strategies, structured generation, and model-quality evaluation.
Skills: NIM Solution Architect, customer solution delivery, LLM, VLM, inference infrastructure, technical support and guidance
This role has been open 0 days — well below the 41-day median for Applications FAE roles.
Applications FAE · AI ML Hardware
|
Open roles in category
96
|
Median days open
41 d
|
Median salary
$261k
|
See the full market breakdown ▾Category comparison, skills in demand, and who else is hiring
| Metric | NVIDIA | All employers we track in this specialty (96 roles · 18 employers) |
|---|---|---|
| Open roles in this specialty | 61 | 96 |
| Median days open | 31 d | 41 d (−10 d vs this employer) |
| Median salary (USD postings) | — | $261k |
Skills observed across this category: NIM Solution Architect, customer solution delivery, LLM, VLM, inference infrastructure, technical support and guidance
Who's hiring in this category
- NVIDIA (this employer) · 58 open roles · median 37 d
- AMD · 10 open roles · median 44 d
- NXP Semiconductors · 6 open roles · median 12 d
- Tenstorrent · 6 open roles · median 116 d
- Qualcomm · 3 open roles · median 76 d
- Arm Holdings · 1 open role · median 87 d
How we counted: 96 open Applications FAE (AI ML Hardware) roles from 18 employers tracked in the SemiconductorJobs index, counted 21 Sept 2026. Specialty figures count only roles carrying this exact specialty label, so an employer's related work in neighbouring specialties is not included there. Figures refresh nightly.