Salary
Posted
Jun 28, 2026
Location
Last confirmed open
Jul 20, 2026

What this job asks for AI summary

A data engineering role focused on building and maintaining healthcare data integration pipelines using PySpark and HL7 standards. Day-to-day work involves designing ETL/ELT workflows, transforming and validating HL7 messages, and ensuring interoperability across healthcare systems. Suited to engineers with a solid background in distributed data processing and healthcare data exchange, with familiarity with GCP or FHIR being a bonus.

Mid level · 4+ years · National · Bachelor's required · Full-time

Must have (3)
PySpark · 2+ yrsHL7SQL
Nice to have (4)
SMILE CDRGCP, AWS or AzureInformaticaFHIR

“or” means any one of them counts — you don't need all of them.

We read this from the posting text with AI. Skim the description below before ruling yourself out.

How this req sits in the market our data

What gives you an edge
HL72%

Rare in this occupation — lead with these, and say what you built with them.

What won't set you apart
SQL90%

Most people in this occupation already list these. Still required — just not what gets you shortlisted.

What the occupation pays Median $139,500 (middle half $109,370–$169,290).

Estimated from BLS employment for this occupation and area, per-skill prevalence across our listing corpus, and published wage benchmarks — as of Jul 28, 2026. It is a model, not a headcount.

Why we read it this way (6)

This role is based in Chennai, India (IN-TN-Chennai), so US-based compensation benchmarks do not apply.

The degree requirement states 'Bachelor's or Master's degree' — Bachelor's is treated as the minimum hard floor.

FHIR appears both in the job responsibilities (as an HL7 version variant) and explicitly under Preferred Skills; it is marked preferred based on its placement in the 'Good to Have' section.

The SOC classification is a close call: the role centers on building and optimizing data pipelines and ETL/ELT workflows (pointing to 15-1243 Database/Data Architects), but the heavy PySpark and distributed-processing coding emphasis could also support 15-1252 Software Developers.

Total years minimum is set to 4, the role-level requirement stated ('4+ years in Data Engineering'); the 2-year PySpark figure is captured on that skill's years demanded.

This role's work location reads as outside the US — the candidate pool, comp, and contention benchmarks here are US-only (BLS employment, Adzuna/USAJOBS demand, and certified H-1B wages), so treat them as a rough US reference, not a local market.

Read the full posting

The employer publishes the full description on their own site — read it there ↗. Or sign in to read it here — it's free, and it also lets you track this application.

Apply

Apply on employer site ↗