SOSi · Huntsville, AL

Salary
Posted
Jun 30, 2026
Location
Huntsville, AL
Last confirmed open
Jul 21, 2026

What this job asks for AI summary

A senior Data Engineer role embedded with a government customer in Huntsville, Alabama, focused on building and maintaining large-scale data pipelines, ETL processes, and ML infrastructure on supercomputing resources handling petabyte-scale data. The work involves automating manual workflows into MLOps-based systems, designing data architectures, and implementing CI/CD and DevSecOps practices, in close collaboration with Data Scientists. A Top Secret/SCI clearance and at least seven years of relevant experience are required.

Senior level · 7+ years · Huntsville, AL · TS/SCI clearance · Full-time

Must have (6)
PythonGitYAMLDockerSQLCI/CD
Nice to have (11)
PyTorch or TensorFlowPostgreSQLMongoDBNeo4jKubernetesKafkaAirflowSparkGitLabArgoHarness

“or” means any one of them counts — you don't need all of them.

Posted 3 times — it's one opening, so apply once.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

What won't set you apart
SQL90%Python55%Docker53%CI/CD45%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $104,100 (middle half $101,657–$142,578).

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (7)

SOC classification is a genuine toss-up: the role is titled 'Data Engineer' and centers on building pipelines, ETL, and data infrastructure (15-1243 Database Architects), but the heavy emphasis on writing software, CI/CD, MLOps, and extending ML infrastructure also makes 15-1252 Software Developers a strong fit. 15-1243 was chosen as the primary because pipeline/warehouse construction and data architecture are the stated core duties.

The stack section ('The team will work with technologies including…') lists PyTorch/TensorFlow, PostgreSQL, MongoDB, Neo4j, Weaviate, Kubernetes, Kafka, Airflow, Spark, GitLab, Argo, and Harness as context rather than gated requirements; these are marked preferred accordingly.

Weaviate (a vector database) appears in the stack narrative but was omitted as a standalone skill because it is a very niche tool listed only in passing stack context alongside better-known databases already captured.

The degree requirement (BS or MS in a quantitative field) appears only under 'Preferred Qualifications', so no minimum degree is hard-required.

An active TS/SCI clearance is an explicit hard gate ('Top Secret Security Clearance with SCI eligibility' listed under Qualifications).

No compensation figures are provided in the posting.

Requires a TS/SCI clearance — the cleared population is a small fraction of this occupation, so the real candidate pool is materially smaller than the estimate below, which does not model clearance.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗