Job Title:
PySpark DeveloperLocation:
Chennai / Bangalore / Hyderabad / PuneNotice Period:
Immediate to 30 Days
Job Description :We are seeking a skilled
PySpark Developer
with strong experience in Python, PySpark, SQL, and Data Warehousing concepts. The ideal candidate will be responsible for designing, developing, and optimizing large-scale data processing pipelines and ETL solutions using Spark-based technologies.
Key Responsibilities :Design, develop, and maintain scalable ETL/ELT pipelines using PySpark.Build and optimize Spark jobs for performance, reliability, and scalability.Process and transform large datasets using Spark SQL, DataFrames, and RDDs.Develop batch and real-time data processing solutions.Integrate data pipelines with Hive, HDFS, Snowflake, Redshift, and other data platforms.Collaborate with data engineers, analysts,
and business stakeholders.Monitor data pipelines, troubleshoot issues, and ensure SLA compliance.Follow coding best practices, version control, and CI/CD processes.Work with Hadoop ecosystem tools and cloud platforms such as AWS, Azure, or GCP.
Required Skills:Strong hands-on experience in
Python and PySparkExpertise in
Spark SQL, DataFrames, and RDDsGood knowledge of
Hadoop (Hive, HDFS, YARN)Solid
SQL
and query optimization skillsExperience with
Data Warehousing conceptsKnowledge of
Parquet, Avro, JSON
data formatsExperience with
Git
version controlFamiliarity with
Airflow, Oozie, or similar scheduling toolsExposure to
AWS, Azure, or GCP
is an added advantage
📌 PySpark Developer | Python | SQL (Mumbai)
🏢 TCS
📍 Mumbai