Responsibilities:
- Build and optimize ETL pipelines using PySpark and Spark SQL
- Process large datasets from S3 / RDBMS
- Handle data transformation, joins, and aggregations
- Tune Spark jobs for performance and cost
Skills:
- Strong PySpark, Spark SQL, and SQL
- AWS EMR and S3 experience
- Knowledge of Parquet, partitioning, joins
Positive to Have:
- Airflow, Snowflake, Trino
- Data lake and performance tuning experience
Outcome:
- Deliver scalable, high-performance data pipeline
📌 Pyspark Developer (Chennai)
🏢 Relevantz Technology Services
📍 Chennai
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.