Santa Clara, CA

Salary
$168,000–$322,000from the description
Posted
Aug 9, 2026
Location
Santa Clara, CA
Last confirmed open
Sep 24, 2026

What this job asks for AI summary

This Senior SRE role at NVIDIA focuses on designing, deploying, and optimizing on-premises HPC storage infrastructure augmented by cloud computing. The engineer will build automation tooling for large-scale storage environments, work with distributed and parallel file systems, and collaborate with engineering teams to align infrastructure with data-intensive workload requirements. It suits an experienced infrastructure engineer with deep HPC storage and cloud operations expertise.

Senior level · 8+ years · Santa Clara, CA · Full-time

Must have (12)
NetApp or Pure StorageMinIOLustre or GpfsPythonBashGolangAWS, Azure or GCPPrometheusElasticsearchKibanaSplunkZabbix
Nice to have (4)
InfiniBand or RoceSlurm, Pbs or LsfDockerKubernetes

“or” means any one of them counts — you don't need all of them.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

Roughly 5 people in the Santa Clara, CA area plausibly meet what this posting asks for (network and computer systems administrators). range 1–7

Applicant volume Light — Few people clear these requirements, so an application that does clear them gets looked at. Worth applying to even if you miss a nice-to-have.

What gives you an edge
Zabbix4%Golang14%

Rare in this occupation — lead with these, and say what you built with them.

What won't set you apart
Bash70%Python51%AWS50%Prometheus45%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $136,293 (middle half $104,938–$171,398). This posting is about at that midpoint.

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Aug 11, 2026. It is a model, not a headcount.

Why we read it this way (6)

The compensation range spans two internal levels: Level 4 ($168,000–$270,250) and Level 5 ($200,000–$322,000); the posting does not specify which level this opening targets.

SOC classification is a genuine judgment call: the role involves significant automation and tooling development (pointing toward 15-1252 Software Developers), but its primary framing is operating and optimizing HPC storage infrastructure (pointing toward 15-1244). 15-1244 was chosen as the primary code given the infrastructure operations emphasis, with 15-1252 as the runner-up.

The monitoring stack (Prometheus, Grafana, Elasticsearch, Kibana, Splunk, Zabbix) is listed under the required qualifications section with 'such as' framing, indicating the specific tools are interchangeable examples of a required monitoring capability; all are marked required accordingly.

Python, Bash, and Golang are listed together as a single scripting/programming requirement; they are emitted as separate skills since they are distinct technologies used alongside each other.

RDMA (InfiniBand/RoCE), HPC cluster management tools (Slurm/PBS/LSF), and containerization (Docker, Kubernetes) appear under 'Ways To Stand Out Of The Crowd' — the preferred/nice-to-have section.

No specific location is stated in the posting beyond NVIDIA's headquarters context; Santa Clara, CA (CBSA 41940) is used as the best-fit metro. The posting does not explicitly state remote eligibility.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗