Baltimore, MDremote

Salary
Posted
Jul 26, 2026
Location
Baltimore, MD
Last confirmed open
Jul 28, 2026

What this job asks for AI summary

A data engineering role focused on designing and maintaining scalable ETL/ELT pipelines and data validation workflows for commercial and government healthcare clients. The engineer will work across cloud-native AWS services, Python-based processing scripts, and orchestration tools to support analytics, regulatory reporting, and system modernization. The role also involves contributing to CI/CD practices, data quality checks, and documentation of data lineage.

Mid level · 4+ years · Bachelor's required

Quick apply — this platform usually takes a CV and a few fields.

Must have (5)
SQLPythonPandas, Spark or NumPyGitHub Actions or JenkinsAWS Glue, Aws Emr, Hadoop, Aws Glue Workflows or Aws Cdk
Nice to have (3)
AuroraSnowflakedbt or Dataform

“or” means any one of them counts — you don't need all of them.

Posted 2 times — it's one opening, so apply once.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

What won't set you apart
SQL90%Python55%Pandas40%GitHub Actions40%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $142,568 (middle half $111,775–$173,013).

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (5)

The role sits at the boundary between data pipeline/ETL engineering (15-1243 Database Architects) and general software development (15-1252 Software Developers); the emphasis on pipeline design, data schemas, orchestration, and ELT architecture tips it toward 15-1243.

The 4+ years requirement is stated at the role level ('4+ years of experience in a data engineering, data pipeline, or ETL/ELT development role'), so it is captured as the overall experience minimum rather than tied to any individual technology.

Pandas, PySpark, and NumPy are listed together as examples of Python libraries; Pandas is used as the primary skill name with the others as alternatives since the posting treats them as interchangeable options.

AWS Lambda appears both as a required orchestration option (alternatives to AWS Step Functions) and as a preferred standalone AWS service; it is captured once under the required orchestration skill and once as a preferred AWS service to reflect both contexts.

Healthcare domain experience (healthcare data systems, CMS data environments) and AWS certifications are noted as preferred but name no specific technology tool, so they are omitted from the skills list per extraction rules.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗