02 Aug
|
RapidBrains
|
Pune
Job Title : Data QA Engineer
Experience: 5+ Years
Location : Pune
Notice Period: Immediate Joiners Preferred
We are looking for an experienced Data QA Engineer to drive data validation and regression testing across enterprise data pipelines spanning legacy and contemporary cloud platforms. This role focuses on ensuring data accuracy, integrity, completeness, and consistency through source-to-target validation, ETL testing, automation frameworks, and data quality assurance.
The ideal candidate should possess strong expertise in Python, SQL, ETL Testing, Data Validation, and automation-first testing approaches, along with exposure to capital markets or financial data environments.
Key Responsibilities
- Perform end-to-end source-to-target data validation across legacy and cloud-based data platforms.
- Validate data accuracy, completeness, consistency, referential integrity, schema, metadata, lineage, and business rules.
- Build reusable automation frameworks using Python, PySpark, and SQL.
- Perform ETL/ELT testing, regression testing, reconciliation, and transformation validation.
- Develop automated validation scripts, reusable test cases, and regression suites.
- Write and optimize complex SQL queries and stored procedures.
- Validate batch jobs, cloud pipelines,
post-trade data flows, and downstream reporting.
- Integrate automated data validation into CI/CD pipelines using Azure DevOps, Jenkins, or GitLab.
- Identify data anomalies, transformation failures, duplicate records, and integrity issues.
- Work closely with engineering, business, and QA teams to improve data quality and release confidence.
Required Skills
- 5+ years of experience in Data QA, ETL Testing, or Data Engineering QA
- Strong programming skills in Python
- Excellent SQL skills (complex queries, joins, stored procedures)
- Hands-on experience with ETL/ELT Testing
- Data Validation & Source-to-Target Testing
- Regression Testing
- PySpark
- Data Quality Validation
- Schema Validation
- Reconciliation Testing
- Metadata & Lineage Validation
- CI/CD tools (Azure DevOps, Jenkins, GitLab)
- JIRA / XRay
- Agile methodology
Nice to Have
- Capital Markets or Post-Trade domain experience
- Azure, AWS, Azure Data Factory (ADF)
- Databricks
- Hadoop / Spark
- Kafka
- Airflow / AutoSys / Control-M
- MongoDB, Cassandra, DynamoDB
- Kubernetes & Containers
- DevSecOps exposure
- AI-assisted development tools such as GitHub Copilot or Claude
📌 Data QA Engineer (Pune)
🏢 RapidBrains
📍 Pune