Job Description
Role Summary
The Data Engineer designs, builds, and operates the ingestion, transformation, and Lakehouse solutions that power Masco's enterprise POS and adjacent commercial data. This role delivers high-quality, reliable, and secure data pipelines against the standards and specifications set by the Enterprise Data Architect and the direction of the Data Engineering Leader. The engineer contributes across the full pipeline lifecycle — ingestion, data quality, transformation, and enrichment — including the pipeline-side execution of the enterprise attribution crosswalk. Depending on assignment, the Data Engineer may focus more heavily on ingestion or on quality and pipeline development, but the role covers all aspects.
What You'll Own
Ingestion & Pipeline Development
Design, build, and maintain automated ingestion pipelines and multi-source integration for retailer, HQ, and BU data on the Databricks and Azure data stack.
Develop RESTful APIs and API-based integrations for retailer and third-party data acquisition.
Deliver critical-path pipelines and validation for priority data feeds.
Integrate with orchestration tools and cloud or hybrid storage systems to enable end-to-end data workflows.
Contribute to Lakehouse solutions using ACID-compliant storage layers, schema enforcement, and versioning for reliable data management.
Data Quality, Validation & Monitoring
Run data-quality checks at the pipeline level — missing values, outliers, consistency, and cross-source reconciliation.
Implement monitoring, validation, and automated notifications within the data lifecycle to protect performance, quality, and availability.
Support pipeline monitoring, incident response, and continuous improvement in partnership with the Data Engineering Leader.
Maintain change control and testing processes for modifications to pipelines and data models.
Attribution & Master Data Support
Build the ingestion side of the attribution crosswalk and master data foundations against
📌 Data Engineer (India)
🏢 Masco
📍 India