NVIDIA Corporation · Santa Clara, CA

Salary
$184,000–$356,500from the description
Posted
Jul 9, 2026
Location
Santa Clara, CA
Last confirmed open
Jul 20, 2026

What this job asks for AI summary

A senior developer relations role focused on growing an ecosystem of partners building AI software for local and on-device environments — covering models, inference runtimes, agent platforms, and developer tools running on PCs and workstations. The position involves managing strategic partner relationships end-to-end, driving technical integrations with GPU-accelerated AI software, and feeding partner insights back into internal product and engineering teams. Best suited to someone with deep AI/ML software experience and a background in technical partnerships or developer advocacy.

Senior level · 8+ years · Santa Clara, CA · Full-time

Must have (6)
AI/MLgenerative AIinference frameworksvLLM, Sglang, Llama.cpp or OllamaGPU accelerationmodel quantization
Nice to have (4)
diffusion modelsagentic AImulti-GPU inferencelow-latency serving

“or” means any one of them counts — you don't need all of them.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

What gives you an edge
vLLM3%

Rare in this occupation — lead with these, and say what you built with them.

What the occupation pays Median $188,486 (middle half $139,113–$224,920). This posting is about at that midpoint.

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (8)

This is a Developer Relations / Technical Partnerships role — not a pure software engineering or pure business role. 15-1299 (Computer Occupations, All Other) is the best fit; 15-1252 (Software Developers) is the runner-up given the strong hands-on AI software engineering requirement.

The posting covers two salary bands: Level 4 ($184,000–$287,500) and Level 5 ($224,000–$356,500). The min and max reported here span both bands combined. Equity and benefits are also offered.

No specific work location is stated in the posting; NVIDIA's headquarters is Santa Clara, CA, used as the default metro. The role may be based at another NVIDIA office.

The degree requirement accepts 'equivalent experience' in lieu of a Bachelor's or Master's, so no hard degree gate is set.

vLLM, SGLang, Llama.cpp, and Ollama are listed together as examples of inference frameworks under the required qualifications section; vLLM is used as the primary skill name with the others as alternatives.

Skills such as model quantization, memory management, and hardware-aware optimization are listed together in the required section as aspects of GPU/local AI performance knowledge; quantization is captured as the representative named technique.

LLMs, VLMs, diffusion models, speech models, retrieval systems, and agentic patterns appear under 'Ways to stand out from the crowd' (the preferred/nice-to-have section).

Multi-GPU inference, low-latency serving, and related performance-critical areas also appear only under the preferred 'Ways to stand out' section.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗