Senior ML Engineer, Computer Vision - Applied AI at Uber
Seattle, WA
$202,000–$224,000from the description
Jul 17, 2026
Seattle, WA
Jul 21, 2026
What this job asks for AI summary
This role centers on building and deploying computer vision and vision-language models for document intelligence tasks — such as identity verification, receipt transcription, and menu digitization — at large scale. The work spans the full model lifecycle: training and experimentation through to production deployment, monitoring, and optimization for throughput, latency, and efficiency, including edge and mobile targets. It suits an ML engineer with deep hands-on experience in computer vision or multimodal systems who is comfortable bridging research-oriented model development with production engineering.
Senior level · 5+ years · Remote · Full-time
“or” means any one of them counts — you don't need all of them.
We read this from the posting text with AI. Skim the description below before ruling yourself out.
How this req sits in the market our data
Roughly 3,550 people nationally plausibly meet what this posting asks for (data scientists). range 1,500–5,300
Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.
Rare in this occupation — lead with these, and say what you built with them.
Most people in this occupation already list these. Still required — just not what gets you shortlisted.
What the occupation pays Median $122,874 (middle half $87,544–$162,374). This posting is about at that midpoint.
Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.
Why we read it this way (8)
The SOC classification is a close call: this role sits at the intersection of ML research (model training, architecture design — 15-2051 Data Scientists) and production software engineering (deployment, scaling, optimization — 15-1252 Software Developers). 15-2051 is chosen as primary because the core emphasis is on developing, training, and evaluating ML/vision models, with production engineering as a supporting concern.
The compensation figures cited ($202,000–$224,000/year) are location-specific to San Francisco CA, Seattle WA, and Sunnyvale CA. The caller has designated this role as fully remote/national, so these figures may not reflect the full remote pay range.
The posting states employees must spend at least 50% of their time in-office unless approved for full remote work; the caller has overridden this to remote=true.
PyTorch, JAX, and TensorFlow are listed together as interchangeable framework options ('such as PyTorch, JAX, or TensorFlow Lite') in the required qualifications; PyTorch is used as the primary with the others as alternatives.
TensorFlow Lite and ONNX appear together under Preferred qualifications as edge/mobile optimization tools and are treated as interchangeable alternatives.
'Computer vision or multimodal systems' is the stated focus area for the 5+ years experience requirement; this is captured via the overall years minimum and the computer vision skill rather than as a separate the years demanded on a named tool.
Skills such as object detection, segmentation, OCR, document layout understanding, VLMs, distributed training, and edge deployment (TensorFlow Lite/ONNX/quantization) all appear exclusively under Preferred Qualifications.
Caller marked this a fully-remote role — scored against the national candidate pool.
Read the full posting
The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.