- Design, develop, test, deploy, and maintain large-scale data pipelines using Hadoop, Spark, Hive, and other big data technologies.
- Collaborate with cross-functional teams to gather requirements and design solutions meeting business needs.
- Develop complex SQL queries to extract insights from large datasets and provide actionable recommendations to stakeholders.
- Troubleshoot data processing, storage, and retrieval issues.
- Work with big data technologies such as Spark, Hadoop, and Hive to analyze and process large datasets.
- Ensure data quality and integrity by implementing data validation and testing procedures.
Job Requirements :
- Strong expertise in Java programming language and PySpark framework.
- Proficiency in Hadoop ecosystem (HDFS) including Hive for querying large datasets.
- Experience with big data technologies such as Spark, Hadoop, and Hive.
- Solid understanding of data engineering principles and practices.
- Ability to work collaboratively with cross-functional teams.
- Strong problem-solving skills and attention to detail.
📌 Data Engineer (Pune)
🏢 IRIS SOFTWARE
📍 Pune
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.