Key Responsibilities
- Design, develop, and maintain scalable data processing systems using Spark and Hadoop ecosystem tools
- Develop and optimize batch and real-time data pipelines using Spark / PySpark
- Work extensively with Hive, Impala, and Sqoop for data ingestion and querying
- Write effective and complex SQL queries for data extraction, transformation, and analysis
- Perform data modeling and ensure optimal performance tuning of large-scale data systems
- Collaborate with cross-functional teams including data engineers, analysts, and business stakeholders
- Monitor and troubleshoot data pipelines to ensure data quality and reliability
- Implement best practices for big data processing, security, and governance.
Required Skills & Qualifications
- 4-6 years of experience in Big Data development
- Strong hands-on experience with Apache Spark and Hadoop ecosystem
- Proficiency in Python and/or Scala
- Strong expertise in SQL
- Experience with Hive, Impala, Sqoop
- Familiarity with Shell Scripting
- Good understanding of data modeling concepts
- Experience in performance tuning and optimization of big data applications.
📌 Big Data Developer (Pune)
🏢 Datametica
📍 Pune
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.