Role & responsibilities Strong expertise in Apache Spark (batch + streaming) and Hive. Proficiency in Python, Scala, or Java. Knowledge of orchestration tools (Airflow / Control-M) and SQL transformation frameworks (DBT preferred). Experience working with Kafka, Solace, and object stores (S3, MinIO). Exposure to Docker/Kubernetes for deployment. Hands on experience of data Lakehouse formats (Iceberg, Delta Lake, Hudi).
Responsibilities
- Solid expertise in Apache Spark (batch + streaming) and Hive.
- Proficiency in Python, Scala, or Java.
- Knowledge of orchestration tools (Airflow / Control-M) and SQL transformation frameworks (DBT preferred).
- Experience working with Kafka, Solace, and object stores (S3, MinIO).
- Exposure to Docker/Kubernetes for deployment.
- Hands on experience of data Lakehouse formats (Iceberg, Delta Lake, Hudi).
📌 Data Engineer - PySpark (Chennai)
🏢 Tekskills
📍 Chennai
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.