Senior Python Developer with Spark
CGI Technologies and Solutions, Inc. · Reston, VA
$97,300–$156,700from the description
Jul 14, 2026
Reston, VA
Jul 21, 2026
What this job asks for AI summary
A senior data engineering role centered on building and tuning large-scale data pipelines using Python and Apache Spark within a cloud-native AWS environment, with a focus on financial services data. Day-to-day work involves Spark performance optimization, advanced SQL for hierarchical datasets, and integrating AWS services such as EMR, Redshift, and Glue. Suits an experienced data engineer comfortable with distributed systems, CI/CD tooling, and cross-functional collaboration.
Senior level · 8+ years · Washington-Arlington-Alexandria, DC-VA-MD-WV · Bachelor's required · Full-time
We read this from the posting text with AI. Skim the description below before ruling yourself out.
How this req sits in the market our data
Roughly 25 people in the Washington-Arlington-Alexandria, DC-VA-MD-WV area plausibly meet what this posting asks for (database architects). range 5–40
Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.
Most people in this occupation already list these. Still required — just not what gets you shortlisted.
What the occupation pays Median $166,452 (middle half $130,866–$200,954). This posting is about at that midpoint.
Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.
Why we read it this way (7)
The role sits at the boundary between data/pipeline engineering (15-1243 Database Architects) and general software development (15-1252 Software Developers). The heavy emphasis on Spark performance tuning, ETL pipeline design, data modeling, and big-data ecosystems tips it toward 15-1243, but the breadth of AWS service development and API work makes 15-1252 a credible alternative.
The posting lists specific AWS services (EMR, Lambda, Step Functions, EventBridge, Redshift, S3, Glue) as a single grouped requirement; each is emitted as a separate skill because they are distinct, separately-marketed services rather than interchangeable alternatives.
Hadoop and Hive appear in the required qualifications under 'big data ecosystems such as Hadoop, Hive, and EMR.' The 'such as' qualifier means the specific tools are interchangeable, but the capability itself is a hard gate; they are emitted as separate required skills since they are distinct technologies rather than alternatives to each other.
CI/CD, GitLab, and Terraform are listed together in the required qualifications ('CI/CD pipelines and tools such as GitLab and Terraform'). CI/CD is the capability gate; GitLab and Terraform are named tools within that requirement and are emitted separately as required.
The degree requirement states 'Bachelor's degree in Computer Science, Information Systems, or a related field' with no 'or equivalent experience' language, so it is treated as a hard requirement.
The role is hybrid (not fully remote) at a client site in Reston, VA.
'Experience in financial services or regulated environments,' 'Exposure to data visualization tools,' and 'Familiarity with event-driven architectures on AWS' appear under the 'Desired Skillset' heading and are treated as preferred. 'Data visualization' and 'event-driven architecture' are the only ones concrete enough to emit as skills.
Read the full posting
The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.