Irving, TX

Salary
$125,760–$188,640from the description
Posted
Aug 12, 2026
Location
Irving, TX
Last confirmed open
Sep 24, 2026

What this job asks for AI summary

A senior data engineering role at Citi focused on designing and operating large-scale data pipelines, Lakehouse architectures, and federated query solutions across cloud platforms. The position involves hands-on work with Databricks, Snowflake, Starburst/Trino, and PySpark, while also supporting AI/ML initiatives such as RAG and Agentic AI systems. It carries technical leadership responsibilities including mentoring junior engineers and serving as a stakeholder-facing point of contact.

Senior level · 6+ years · Dallas-Fort Worth-Arlington, TX · Full-time

Must have (9)
PythonPySparkDatabricksDelta LakeAb InitioSnowflakeStarburst or TrinoApache IcebergAWS, GCP, Azure, Azure Data Factory or Google Cloud Composer
Nice to have (7)
Redshift, Azure Synapse Analytics or BigQueryPandasNumPyDaskDockerKubernetesUnity Catalog

“or” means any one of them counts — you don't need all of them.

Posted 4 times — it's one opening, so apply once.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

Roughly 6 people in the Dallas-Fort Worth-Arlington, TX area plausibly meet what this posting asks for (database architects). range 3–10

Applicant volume Moderate — A normal amount of company. The rare requirements below are what will separate a shortlisted application from the rest.

What gives you an edge
Ab Initio3%

Rare in this occupation — lead with these, and say what you built with them.

What won't set you apart
Python55%Databricks42%Snowflake42%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $158,173 (middle half $132,644–$169,109). This posting is about at that midpoint.

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Aug 15, 2026. It is a model, not a headcount.

Why we read it this way (7)

The SOC classification is a close call between 15-1243 (Database Architects, covering data pipeline and warehouse design) and 15-1252 (Software Developers). The role's primary emphasis on designing Lakehouse/data warehouse architecture, ETL/ELT pipelines, and federated query solutions tips it toward 15-1243, but the production-grade coding expectations (PySpark, Python, Ab Initio) give 15-1252 a credible claim.

Ab Initio (GDE, Co>Operating System, Conduct>It) is listed under the core required technologies section with 'strong, practical experience' language — treated as a hard gate despite being an older ETL tool not commonly seen alongside modern cloud stacks.

The cloud provider requirement (AWS, GCP, or Azure) is framed as 'at least one major cloud provider' — AWS is listed as the primary name with GCP and Azure as explicit alternatives.

Cloud-native services (AWS Glue/Lambda/S3/Redshift, Azure Data Factory/Synapse, Google Cloud Composer/Dataflow/BigQuery) appear as illustrative examples of the cloud-native services requirement rather than individually gated skills; captured as preferred.

Docker and Kubernetes appear only under 'Recommended Qualifications' — treated as preferred.

The 6–10 year experience range is stated under 'Recommended Qualifications'; however, given the depth of required expertise across many technologies and the leadership expectations, 6 years is used as the overall years minimum.

A Bachelor's degree or equivalent experience is accepted, so the degree requirement is set to None.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗