Design, develop, and maintain scalable data pipelines and ETL/ELT processes.
Develop data processing solutions using Python, PySpark, and Scala.
Write and optimize complex SQL queries.
Work with AWS/Azure/GCP cloud services for data engineering solutions.
Perform data transformation, cleansing, and validation.
Optimize Spark jobs and data pipelines for performance and scalability.
Troubleshoot data pipeline issues and support production environments.
Collaborate with data scientists, analysts, and engineering teams.