NVIDIA Corporation · Hillsboro, ORremote

Salary
$184,000–$356,500from the description
Posted
Jun 28, 2026
Location
Hillsboro, OR
Last confirmed open
Jul 21, 2026

What this job asks for AI summary

A senior data engineering role focused on building and optimizing distributed data pipelines and microservices that process large volumes of autonomous vehicle data to support AI training and data mining. Day-to-day work spans ETL pipeline development, workflow orchestration, RAG workflow design, and deploying AI models, using technologies such as Kubernetes, Spark, vector databases, and big data engines. Best suited to engineers with deep experience in scalable distributed systems and big data processing.

Senior level · 8+ years · Remote · Full-time

Must have (8)
Python or GoKubernetesHelmHiveParquetSQLMilvusETL pipelines
Nice to have (4)
SparkLLMsRAGNVIDIA RAPIDS

“or” means any one of them counts — you don't need all of them.

Posted 2 times — it's one opening, so apply once.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

Roughly 880 people nationally plausibly meet what this posting asks for (database architects). range 180–1,300

Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.

What won't set you apart
SQL90%Python55%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $139,500 (middle half $109,370–$169,290). This posting is about at that midpoint.

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (8)

The posting spans two internal levels (Level 4 and Level 5) with distinct salary bands: $184,000–$287,500 (L4) and $224,000–$356,500 (L5). The pay band fields reflect the full combined range across both levels.

SOC classification is a genuine judgment call: the role centers on designing and building data pipelines, ETL systems, and distributed data architectures (pointing to 15-1243 Database Architects), but also involves shipping microservices and distributed applications in Python/Go (pointing to 15-1252 Software Developers). 15-1243 was chosen because data pipeline design and architecture is the primary framing.

Milvus is listed as an example of a vector database ('e.g., Milvus') under the required qualifications section. It is treated as a hard gate on vector database proficiency, with Milvus as the named representative; no other specific vector DB is named.

ETL pipelines and big data engines appear in the required section as a capability gate rather than a named tool — captured as 'ETL pipelines' since no specific engine (beyond Hive/Parquet/Spark) is named there.

Spark, LLMs, RAG, and NVIDIA RAPIDS all appear under the 'Ways to Stand Out from the Crowd' / 'Eagerness to learn' sections and are treated as preferred. Note that RAG workflows are also mentioned in the main responsibilities, but the JD does not gate candidates on prior RAG experience in the requirements block.

Degree requirement is None: the JD explicitly accepts 'equivalent experience' in lieu of a BS, and the MS path is an alternative, not a minimum floor.

The caller has declared this a fully-remote role; no metro is inferred.

Caller marked this a fully-remote role — scored against the national candidate pool.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗