Pittsburgh, PA

Salary
$100,000–$120,000
Posted
Jul 20, 2026
Location
Pittsburgh, PA
Last confirmed open
Jul 22, 2026

What this job asks for AI summary

An entry-level data engineering position focused on building and maintaining data pipelines, ETL/ELT processes, and Lakehouse solutions on the Databricks platform using PySpark and SQL. The role involves working under senior engineers on tasks ranging from performance tuning and data quality to CI/CD and documentation. It suits a recent master's graduate with foundational knowledge of Apache Spark, Python, and cloud-based data platforms.

Junior level · Pittsburgh, PA · Master's required

Must have (5)
DatabricksApache SparkPythonSQLGit
Nice to have (5)
CI/CDAzure, AWS or GCPDelta LakeAzure Data Factory or Microsoft FabricPower BI or Tableau

“or” means any one of them counts — you don't need all of them.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

Roughly 6 people in the Pittsburgh, PA area plausibly meet what this posting asks for (database architects). range 4–9

Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.

What won't set you apart
SQL90%Python55%Databricks42%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $156,160 (middle half $124,540–$192,911).

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (6)

The role is classified as a data pipeline/lakehouse engineering role — the primary day-to-day work is building ETL/ELT pipelines and Lakehouse solutions (closer to 15-1243 Database Architects / data engineers) rather than general software development (15-1252), though the boundary is genuinely ambiguous given the PySpark coding emphasis.

A Master's degree in a relevant technical field is explicitly listed as a hard requirement with no 'or equivalent experience' escape clause — the degree requirement is set to Masters accordingly.

Databricks and Apache Spark appear in the Required Qualifications section framed as 'familiarity with' and 'knowledge of' — softer language than typical hard gates, but they are the core platform of the role and PySpark is called out explicitly as a key coding responsibility, so all three are treated as required.

Familiarity with Medallion Architecture is listed under Preferred Qualifications only.

No compensation figures are provided in the posting.

Relocation assistance to Pittsburgh, PA is explicitly offered.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗