- Develop and maintain scalable data processing applications using PySpark.
- Design, build, and optimize ETL/ELT pipelines for large datasets.
- Process structured and unstructured data using Apache Spark.
- Collaborate with data engineers, data scientists, and business stakeholders to understand requirements.
- Write productive Spark jobs and optimize performance for distributed environments.
- Integrate data from multiple sources such as databases, APIs, and cloud storage.
- Perform data quality checks, validation, and troubleshooting.
- Monitor and maintain production data pipelines.
- Implement best practices for coding, testing, deployment, and documentation.
- Work with cloud platforms such as AWS, Azure, or GCP.
📌 Pyspark DeveloperF discussion - Chennai
🏢 Tata Consultancy Services
📍 Chennai
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.