- Design, develop, test, deploy and maintain large-scale data pipelines using PySpark on GCP.
- Collaborate with cross-functional teams to gather requirements and deliver high-quality data solutions.
- Ensure scalability, reliability, and performance of the data pipeline by monitoring its metrics and implementing improvements.
- Troubleshoot issues related to data processing, storage, and retrieval from various sources.
Job Requirements :
- 7-9 years of experience in Data Engineering with expertise in Python programming language.
- Robust understanding of Big Data technologies such as Hadoop ecosystem (HDFS) and Spark (PySpark).
- Proficiency in working with Google Cloud Platform (GCP) services including Storage Buckets, Dataproc Clusters, Pub/Sub Topics etc.
📌 Data Engineer (Chennai)
🏢 IonIdea
📍 Chennai
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.