- Data Engineer with GCP + Pyspark, Scala Should know how to develop, deploy , monitor and optimize Etl workflows for batch with Scala, Pyspark, Airflow, BQ , Dataproc and GCP environments
- Experience in Medallion architecture with proper Data quality checks and summarization activities
- Experience in performance optimizations, Observability and runbook creations for continuous optimizations.
- Positive knowledge of Airflow orchestrations including dependency handing, idempotent execution etc.
- Exposure to data governance would be a plus.