13 Sep
|
Finarb
|
Secunderabad
13 Sep
Finarb
Secunderabad
Job DescriptionData QA Engineer
NLocation: Kolkata, India (Onsite/Hybrid)
NExperience: 2–4 years
NEmployment Type: Full-time
NABOUT THE ROLE
nWeare hiring a Data QA Engineer for our data engineering team. This role is responsible for
Nvalidating ETL pipelines and data platforms built on Azure Data Factory, Databricks, and
NMicrosoft Fabric, along with the downstream tables, reports, and models they feed. The core
Nresponsibility is verifying data correctness, which is distinct from confirming that a pipeline
Nexecuted without errors —the two are frequently conflated, and this role exists to keep them
Nseparate.
nKEY RESPONSIBILITIES
N
n
- Design and execute test plans covering source-to-target validation and transformation logic
N
- for ETL pipelines
n
- Write SQL and PySpark scripts to verify accuracy, completeness, and consistency in Delta
N
- Lake tables
n
- Perform regression testing on every pipeline change;
a successful pipeline run does not
N
- guarantee correct output, and validation must be independent of execution status
N
- Build and maintain reusable data quality checks (e.G., Outstanding Expectations, dbt tests, or
N
- custom PySpark frameworks) in place of one-off manual queries
N
- Reconcile data between source systems and target Lakehouse/Warehouse layers, and
N
- investigate root cause when discrepancies are found
N
- Validate schema conformance, null/duplicate handling, referential integrity, and business
N
- rule adherence across Bronze/Silver/Gold layers
N
- Validate Microsoft Fabric artifacts — Lakehouses, Warehouses, and semantic models —
N
- including DirectLake mode behavior and cross-domain data consistency
N
- Apply consistent QA methodology across tools;
the underlying platform(Databricks, Fabric,
N
- or otherwise) should not change how rigorously data is validated
N
- Document test cases anddefects with enough detail for engineers to act on them without
N
- requiring additional clarification
N
- Work directly with the data engineering team on requirement clarification and defect
N
- resolution
n
nREQUIRED SKILLS
N
n
- Strong SQL, with the ability to write validation queries independently
N
- Working proficiency in Python/PySpark for scripting data checks
N
- Solid understanding of ETL concepts: staging, transformations, incremental loads, SCD
N
- handling
n
- Hands-on experience with Azure Data Factory and Databricks
N
- Working knowledge of Microsoft Fabric (Lakehouse, Warehouse, semantic models)
N
- Familiarity with Delta Lake and medallion architecture
N
- Ability to read transformation logic and determine expected output
N
- General data QA methodology that transfers across tools and platforms, not skills tied to a
N
- single vendor stack
N
- Experience with defect tracking and structured test documentation (Jira, TestRail, or
N
- equivalent)PREFERRED QUALIFICATIONS
N
- Experience with a data quality framework (Great Expectations, Deequ, dbt tests)
N
- Familiarity with Unity Catalog and general data governance/lineage tooling
N
- Experience validating pipelines in a regulated domain (pharma, healthcare, finance) where
N
- lineage and auditability are requirements
N
- Exposure to CI/CD for test automation
N
- Azure, Databricks, or Fabric certification
N
nQUALIFICATIONS
nBachelor's degree in Computer Science, IT, or a related field
N2–4 years of experience in data QA, data validation, or ETL testing;
manual/UI testing
nexperience without pipeline exposure does not meet this requirement
NCANDIDATE FIT
nThis role requires the ability to determine why a data discrepancy occurred, not simply flag
Nthat one exists. Candidates whose QA background is primarily manual/UI testing and who are
Nseeking to transition into data-focused work should not apply for this position;
the required
ndata depth is expected from day one.
📌 Data Quality Assurance Lead (Secunderabad)
🏢 Finarb
📍 Secunderabad