08 Sep
|
Questhiring
|
Gurugram
08 Sep
Questhiring
Gurugram
Job Description
n
Role: Data Engineer II – GCP Data Platform
n
Objective of the Role
n
We are looking for a hands-on Data Engineer II with strong Google Cloud Platform experience to build scalable data pipelines, transformations and reusable data-platform capabilities.
n
You will work across Data Engineering and Data Platform Engineering , independently delivering production-grade data products while contributing reusable components and engineering patterns that can be adopted across multiple teams and use cases.
n
The role requires strong hands-on development skills and the ability to take a solution from design through implementation, testing, deployment and production support.
n
You Will
n
n
- Design, develop and maintain production-grade data pipelines on Google Cloud Platform .
n
- Build scalable batch and incremental data-processing solutions .
n
- Develop complex transformations using BigQuery, SQL and Dataform / dbt .
n
- Build and maintain data-processing workflows using Cloud Composer / Apache Airflow .
n
- Develop reusable transformation components, orchestration patterns, libraries and templates.
n
- Implement metadata-driven and configuration-driven processing where appropriate.
n
- Build distributed data-processing solutions using Dataflow / Apache Beam and/or Dataproc / Spark .
n
- Design data models for analytical and downstream data-product requirements.
n
- Implement incremental and idempotent processing patterns.
n
- Define and maintain schemas and data contracts.
n
- Implement automated data-quality checks, validation and reconciliation.
n
- Capture and integrate metadata and lineage into data-processing workflows.
n
- Optimise BigQuery queries and pipelines for performance and cloud cost.
n
- Implement automated testing and integrate data workloads with CI/CD pipelines.
n
- Build monitoring and operational controls for production pipelines.
n
- Troubleshoot production issues and perform root-cause analysis.
n
- Contribute to reusable data-platform capabilities and engineering standards.
n
- Collaborate with Data Engineers, Platform Engineers, DevOps, Architects and business stakeholders.
n
n
n
You Must Have
n
n
- 4+ years of hands-on Data Engineering experience building production data solutions.
n
- Strong practical experience with Google Cloud Platform (GCP) .
n
- Strong hands-on experience with BigQuery .
n
- Advanced SQL skills including complex joins, CTEs, window functions and query tuning.
n
- Strong Python development skills for data processing, automation and testing.
n
- Experience developing production ETL / ELT pipelines .
n
- Hands-on experience with Cloud Composer / Apache Airflow .
n
- Experience with Dataform and/or dbt .
n
- Experience designing batch and incremental processing pipelines.
n
- Experience with Dataflow / Apache Beam or Dataproc / Spark / PySpark .
n
- Robust understanding of data modelling, including normalization, denormalization and dimensional modelling.
n
- Experience working with large-scale datasets.
n
- Experience with incremental and idempotent pipeline patterns.
n
- Understanding of data contracts and schema evolution.
n
- Experience implementing data-quality validation and reconciliation.
n
- Understanding of metadata and data lineage.
n
- Experience with Git, automated testing and CI/CD .
n
- Experience troubleshooting and supporting production data pipelines.
n
- Ability to independently design solutions rather than only implement predefined specifications.
n
n
Technical Skills
n
n
Cloud & Storage
n
n
- Google Cloud Platform
n
- BigQuery
n
- Google Cloud Storage
n
n
n
Transformation & Orchestration
n
n
- Advanced SQL
n
- Dataform / dbt
n
- Cloud Composer / Apache Airflow
n
n
n
Data Processing
n
n
- Dataflow / Apache Beam
n
- Dataproc / Spark / PySpark
n
n
n
Programming & Engineering
n
n
- Python
n
- Git
n
- Automated testing
n
- CI/CD
n
- Monitoring and troubleshooting
n
n
n
Data Engineering Capabilities
n
n
- ETL / ELT
n
- Batch processing
n
- Incremental and idempotent pipelines
n
- Data modelling
n
- Data contracts
n
- Data quality and reconciliation
n
- Metadata and lineage
n
- Query performance optimisation
n
- Cloud-cost optimisation
n
- Reusable data-platform components
n
n
n
Good to Have
n
n
- Experience with Apache Iceberg or modern lakehouse table formats.
n
- Experience with streaming or event-driven processing.
n
- Practical understanding of Data Product / Data Mesh principles .
n
- Experience with metadata-driven or configuration-driven processing.
n
- Experience with Change Data Capture.
n
- Familiarity with Terraform / Infrastructure as Code .
n
- Experience building reusable components used by multiple engineering teams.
n
n
n
Strong Interpersonal Skills
n
n
- Strong ownership mindset and ability to take engineering work through production.
n
- Solid analytical, debugging and problem-solving skills.
n
- Ability to communicate technical decisions clearly.
n
- Comfortable participating in design and code reviews.
n
- Ability to collaborate with engineering and business stakeholders.
n
- Ability to work independently while seeking guidance for complex architectural decisions.
n
- Comfortable working within distributed and multicultural teams.
n
n
📌 Data Engineer (Gurugram)
🏢 Questhiring
📍 Gurugram