Director of AI Workload Intelligence - San Jose, CA, United States
$250,000 - $300,000 USD yearly
Originally posted 1 September 2026 by the employer — open 20 days.
About the role
This Director-level position leads global initiatives to analyze LLM and multimodal workloads for HBF applicability, bridging AI intelligence with memory strategy.
What you'll do
- Lead global collaboration and drive AI workload-based HBF technology development from research to PoC and productization.
- Analyze AI models, algorithms, and workload trends for HBF applicability.
- Classify HBF-relevant memory objects such as weights, KV cache, activations, embeddings, adapters, and MoE experts.
- Define model-level criteria for data placement, caching, prefetching, and offload decisions.
- Analyze HBF integration points in vLLM, SGLang, TensorRT-LLM, PyTorch, and ONNX Runtime.
- Develop memory-use taxonomy and HBF applicability criteria.
What you'll need
- Proven experience leading global AI system architecture initiatives and cross-functional teams.
- Hands-on experience in AI inference systems, ML systems, LLM serving, AI model analysis, AI framework/runtime analysis, or heterogeneous accelerator software.
- Strong understanding of Transformers, LLMs, MoE, recommendation models, embedding/retrieval workloads, and multimodal inference.
- Experience with PyTorch, Hugging Face, vLLM, SGLang, TensorRT-LLM, ONNX Runtime, Triton Inference Server, or equivalent AI inference stacks.
- Experience in latency, throughput, token throughput, memory footprint, bandwidth, and data movement analysis.
Skills: AI workload, HBF technology, LLMs, AI inference SW Stack, memory objects
This role has been open 20 days — well below the 67-day median for AI/ML Hardware Engineering roles.
AI/ML Hardware Engineering · AI ML Hardware
|
Open roles in category
522
|
Median days open
67 d
|
Median salary
$235k
|
See the full market breakdown ▾Category comparison, skills in demand, and who else is hiring
| Metric | SK Hynix America | All employers we track in this specialty (522 roles · 76 employers) |
|---|---|---|
| Open roles in this specialty | 3 | 522 |
| Open roles in the wider Software, Firmware & Systems family | 5 | 5873 · 143 employers |
| Median days open | 177 d | 67 d (+110 d vs this employer) |
| Median salary (USD postings) | — | $235k |
Skills observed across this category: AI workload, HBF technology, LLMs, AI inference SW Stack, memory objects
Who's hiring in this category
- Qualcomm · 107 open roles · median 97 d
- NVIDIA · 86 open roles · median 68 d
- AMD · 46 open roles · median 60 d
- Micron Technology · 32 open roles · median 40 d
- Mobileye · 25 open roles · median 54 d
- Analog Devices · 14 open roles · median 25 d
How we counted: 522 open AI/ML Hardware Engineering (AI ML Hardware) roles from 76 employers tracked in the SemiconductorJobs index, counted 22 Sept 2026. Specialty figures count only roles carrying this exact specialty label, so an employer's related work in neighbouring specialties is not included there — it is counted in the wider Software, Firmware & Systems family row. Figures refresh nightly.