07 Aug
|
Datametica
|
Pune
We are looking for a skilled Big Data Developer with strong expertise in Spark and Hadoop ecosystem technologies. The ideal candidate will have hands-on experience building scalable data pipelines, optimizing data processing workflows, and working with large datasets in distributed environments.
Key Responsibilities
- Design, develop, and maintain scalable data processing systems using Spark and Hadoop ecosystem tools
- Develop and optimize batch and real-time data pipelines using Spark / PySpark
- Work extensively with Hive, Impala, and Sqoop for data ingestion and querying
- Write efficient and complex SQL queries for data extraction, transformation, and analysis
- Perform data modeling and ensure optimal performance tuning of large-scale data systems
- Collaborate with cross-functional teams including data engineers, analysts,
and business stakeholders
- Monitor and troubleshoot data pipelines to ensure data quality and reliability
- Implement best practices for big data processing, security, and governance.
Required Skills & Qualifications
- 4-6 years of experience in Big Data development
- Strong hands-on experience with Apache Spark and Hadoop ecosystem
- Proficiency in Python and/or Scala
- Strong expertise in SQL
- Experience with Hive, Impala, Sqoop
- Familiarity with Shell Scripting
- Valuable understanding of data modeling concepts
- Experience in performance tuning and optimization of big data applications.
📌 Big Data Developer (Pune)
🏢 Datametica
📍 Pune