Staff Engineer, ML Profiling Tools - San Jose, California, United States
$163,000 - $253,000 USD yearly
Originally posted 2 October 2026 by the employer.
Focus on ML Profiling Tools for AI/ML workloads.
About the role
This role involves designing and developing scalable platforms to handle computational and memory requirements of AI/ML workloads while minimizing energy consumption and maximizing performance. This position will focus on ML Profiling Tools.
What you'll do
- Develop and maintain web-based visual analytics tools on AI/ML workloads.
- Build interactive data visualizations (flame graphs, call trees, timeline charts) for multi-dimensional infrastructure telemetry datasets.
- Partner with compiler engineers and hardware architects to translate performance profiling metrics into intuitive developer workflows.
- Analyze and optimize developer experience for machine learning performance profiling, designing tooling to diagnose and eliminate inference bottlenecks.
- Communicate with stakeholders to ensure systems are delivered on time and within budget.
What you'll need
- Experience designing and implementing rich, interactive data visualizations and user interfaces using modern web technologies.
- Strong background or interest in creating developer tools, IDE extensions, or complex diagnostic dashboards.
- Proven ability to translate massive, unstructured, or multidimensional telemetry/profiling data into intuitive, human-readable visual representations.
- Strong track record of designing workflows for technical users (software engineers, data scientists, or researchers), focusing on minimizing cognitive load and streamlining root-cause diagnosis.
- Hands-on experience with, or strong curiosity about, systems profiling and tracing tools (e.g., Perfetto, Chrome Tracing, NVIDIA Nsight Systems/Compute, PyTorch Profiler, TensorBoard, eBPF).
- Baseline understanding of computing architectures (CPUs, GPUs, TPUs, interconnects/networking) and typical performance bottlenecks (memory bandwidth, compute utilization, synchronization stalls).
Nice to have
- Working knowledge of modern machine learning frameworks (e.g., PyTorch, JAX, TensorFlow) and execution paradigms (LLM inference, pipeline/tensor parallelism, kernel dispatch).
Skills: ML Profiling Tools, AI/ML workloads, performance profiling, data visualization, computing architectures
This role has been open 0 days — well below the 51-day median for Infrastructure/Platform Software roles.
Infrastructure/Platform Software · Infrastructure Platform
|
Open roles in category
874
|
Median days open
51 d
|
Median salary
$236k
|
See the full market breakdown ▾Category comparison, skills in demand, and who else is hiring
| Metric | Samsung Semiconductor | All employers we track in this specialty (874 roles · 84 employers) |
|---|---|---|
| Open roles in this specialty | 2 | 874 |
| Open roles in the wider Software, Firmware & Systems family | 13 | 5917 · 145 employers |
| Median days open | 12 d | 51 d (−39 d vs this employer) |
| Median salary (USD postings) | — | $236k |
Skills observed across this category: ML Profiling Tools, AI/ML workloads, performance profiling, data visualization, computing architectures
Who's hiring in this category
- NVIDIA · 245 open roles · median 44 d
- Qualcomm · 57 open roles · median 66 d
- AMD · 47 open roles · median 24 d
- Cerebras · 40 open roles · median 68 d
- Graphcore · 37 open roles · median 95 d
- Micron Technology · 30 open roles · median 52 d
How we counted: 874 open Infrastructure/Platform Software (Infrastructure Platform) roles from 84 employers tracked in the SemiconductorJobs index, counted 3 Oct 2026. Specialty figures count only roles carrying this exact specialty label, so an employer's related work in neighbouring specialties is not included there — it is counted in the wider Software, Firmware & Systems family row. Figures refresh nightly.