Los Angeles, CAremote

Salary
$114,486–$163,552from the description
Posted
Aug 16, 2026
Location
Los Angeles, CA
Last confirmed open
Sep 24, 2026

What this job asks for AI summary

This role sits on Spotify's Voice Production team and is responsible for defining and measuring the quality of Spotify's AI-generated voices across products such as AI DJ, Generative AI Ads, and Personalized Podcasts. The analyst independently designs evaluation rubrics, runs qualitative and quantitative evaluation campaigns, and benchmarks Spotify's voice models against competitors and third-party providers. It suits someone with a background in product quality, data curation, or model evaluation who has a strong ear for audio and voice characteristics.

Senior level · New York-Newark-Jersey City, NY-NJ-PA · Full-time

Quick apply — this platform usually takes a CV and a few fields.

Must have (2)
MOS-style evaluationsqualitative evaluation
Nice to have (2)
generative AIspeech technology

We read this from the posting text with AI. Skim the description below before ruling yourself out.

Why we read it this way (6)

This role is unusually difficult to classify by SOC code. Its primary work is designing and running human-judgment evaluations of AI voice models — a blend of product quality analysis, data curation, and operations research. 15-2031 (Operations Research Analysts) is the closest fit for the rubric-design and quantitative-synthesis work, but 15-1299 (Computer Occupations, All Other) is a reasonable alternative given the AI/ML evaluation context.

The posting names two work locations — New York and Los Angeles — but no single primary site is specified. New York is listed first and used here; the Los Angeles CBSA (31080) is equally valid.

The skills list is thin by necessity: the posting describes methodological competencies (evaluation design, MOS-style testing, qualitative synthesis, voice/audio quality judgment) rather than named software tools or platforms. No specific technology stack is required or mentioned beyond broad references to generative AI and speech technology.

MOS (Mean Opinion Score) is a standard speech-quality evaluation methodology referenced explicitly in the posting and is treated as a required competency.

No minimum years of overall experience are stated in the posting.

Ignored 1 non-technology phrase(s) as skills (responsibilities/concepts, not named tools): evaluation rubric design.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗