We are looking for a skilled PySpark Developer to join our data engineering team
The ideal candidate should have strong experience in big data processing, distributed computing, and building scalable data pipelines using PySpark
Key Responsibilities
Develop and optimize data pipelines using PySpark
Work with large datasets in distributed environments
Design, build, and maintain ETL processes
Collaborate with data analysts and data scientists for data requirements
Ensure data quality, integrity, and performance tuning
Work on Hadoop ecosystem tools and cloud platforms if required
Required Skills
Robust experience in PySpark and Python
Hands-on with Spark SQL, DataFrames, RDDs
Positive knowledge of Hadoop ecosystem (HDFS, Hive, etc)
Experience in ETL development and data processing
Understanding of big data architecture and distributed systems
Positive to Have
Experience with AWS / Azure / GCP
Knowledge of Kafka or streaming tools
Familiarity with data warehousing concepts
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Pyspark Developer Bengaluru
🏢 Mobile Programming
📍 Bengaluru
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.