Data Engineering Fabric (Anantapur)

Data Engineering Fabric (Anantapur)

20 Sep
|
EXL
|
Anantapur

20 Sep

EXL

Anantapur

Job Description

Role Purpose

n

Build and operate the data pipelines that feed the Entity Hub. This role lands all six in-scope sources into Fabric, implements standardization and transformation logic, and maintains the data quality checks and monitoring that the entity resolution engine depends on. Reliable, observable ingestion is the foundation the entire programme rests on.

n

Key Responsibilities

n

- n
- Ingestion development — build and maintain pipelines to land the six in-scope sources (Secretary of State, D&B;, ARROW, E1, hCue, DocCentral) into the Fabric Bronze/raw layer.n
- Mirroring & CDC — implement Fabric Mirroring for supported structured sources and establish change-data-capture patterns; implement watermark/incremental load logic where mirroring is unavailable.n
- Raw layer management — maintain one Delta table per source on an append-only basis, retaining evidence records and full source provenance.n
- Standardization & transformation — implement name normalization, address parsing and attribute standardization logic in Spark notebooks; support identifier-spine construction.n
- Data quality — implement data quality checks, validation rules, threshold alerts and exception handling; support reconciliation against source.n
- Pipeline operations — schedule, monitor and troubleshoot pipeline runs; investigate failures and performance issues; maintain run documentation.n




- Performance tuning — optimise Spark jobs, Delta file sizes, partitioning and pipeline efficiency to manage Fabric capacity consumption.n
- Documentation — produce and maintain source-to-target mappings, transformation logic documentation and lineage records.n

n

Must-Have Qualifications

n

- n
- 4+ years hands-on data engineering with robust PySpark and SQLn
- Production experience building ingestion pipelines from multiple heterogeneous sourcesn
- Working knowledge of Delta Lake and medallion/lakehouse architecturen
- Experience implementing incremental loads and CDC-style processingn
- Experience implementing data quality checks and troubleshooting pipeline failuresn

n

Nice-to-Have

n

- n
- Microsoft Fabric hands-on experience (Mirroring, Copy Jobs, Environments)n
- Exposure to entity/master data standardization (name and address parsing)n
- Familiarity with libraries such as Great Expectations for data qualityn
- Experience optimising for Fabric capacity/CU consumptionn

n

Key Deliverables Owned

n

- n
- Operational ingestion pipelines for all agreed sourcesn
- Bronze/raw layer with one Delta table per source and CDC retainedn
- Standardization and parsing transformation logicn
- Data quality checks, monitoring and exception handlingn
- Source-to-target mapping and run documentationn

n

📌 Data Engineering Fabric (Anantapur)
🏢 EXL
📍 Anantapur

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: data engineering fabric (anantapur) / anantapur

Subscribe to this job alert:

Get the latest job offers by email for: data engineering fabric (anantapur) / anantapur