Truthset, Inc. · Oakland, CA

Salary
$150,000–$180,000
Posted
Jul 12, 2026
Location
Oakland, CA
Last confirmed open
Jul 20, 2026

What this job asks for AI summary

A data engineering role focused on building and maintaining large-scale data pipelines that move terabytes of structured data to enterprise clients and internal teams. Day-to-day work involves writing Scala-based ETL code, automating data ingestion into cloud warehouses, deploying batch workflows via orchestration tools, and building internal monitoring dashboards. Suits engineers with hands-on experience in distributed cloud environments and data pipeline design.

Mid level · 3+ years · Remote · Bachelor's required · Full-time

Must have (6)
Python, Scala or JavaSparkAWS EMRSnowflake, Databricks or RedshiftSQLGitHub
Nice to have (3)
TerraformAirflowdbt

“or” means any one of them counts — you don't need all of them.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

Roughly 2,900 people nationally plausibly meet what this posting asks for (database architects). range 1,200–4,350

Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.

What won't set you apart
SQL90%Python55%Snowflake42%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $142,568 (middle half $111,775–$173,013).

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (6)

The role sits at the boundary between data/pipeline architecture (15-1243) and software development (15-1252) — it involves both designing data pipelines and writing custom Scala/Python ETL code. 15-1243 was chosen as primary given the emphasis on data pipeline design, warehousing, and data modeling.

The Core Qualifications list Python/Scala/Java as interchangeable options ('proficiency in one or more… such as'), so they are captured as a single must-have skill with alternatives rather than separate hard gates.

Snowflake, Databricks, and Redshift are listed as interchangeable cloud data warehouse options in the requirements section; Snowflake is named first and used as the primary with the others as alternatives.

Scala, Terraform, Airflow, and dbt all appear under 'Ideal Qualifications' (framed as 'familiarity with'), so they are marked preferred. Note that Scala also appears in the required programming-language list as one of several acceptable options, but the Ideal Qualifications section separately calls out industry Scala experience as a bonus, reinforcing its preferred status as a standalone skill.

No compensation figures were provided — only a description of benefits and equity potential.

This posting reads as a fully-remote role, so it was scored against the national candidate pool rather than a single metro.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗