Data Engineer | Hadoop, Spark, PySpark, Python
Experience: 4+ Years
Location: Chennai / Bangalore
Work Mode: Hybrid
Job Description
We are hiring an experienced Data Engineer with solid expertise in Hadoop, Apache Spark, PySpark, Python, Hive, Advanced SQL, and Unix/Linux. The ideal candidate will design, develop, and optimize scalable data pipelines and big data solutions while working with large-scale distributed data processing systems.
Key Responsibilities
- Design and develop scalable ETL/ELT data pipelines using PySpark and Apache Spark.
- Build and optimize data processing workflows on the Hadoop ecosystem.
- Develop data ingestion and transformation frameworks using Python and Spark.
- Write optimized SQL queries for data extraction, transformation, and reporting.
- Develop and maintain Hive tables, partitioning, and performance tuning.
- Monitor, troubleshoot,
and optimize production data pipelines.
- Perform data validation, quality checks, and root cause analysis.
- Collaborate with cross-functional teams in an Agile environment.
Mandatory Skills
- Apache Hadoop
- Apache Spark & Spark Core
- PySpark
- Python
- Hive
- Advanced SQL
- Unix/Linux & Shell Scripting
- ETL/ELT Development
- Big Data Technologies
- Data Pipeline Development
Preferred Qualifications
- BE/B.Tech, MCA, M.Tech, or equivalent.
- Strong analytical and problem-solving skills.
- Experience working in Agile/Scrum environments.
Interested candidate share me your updated cv on
[email protected]
Whatsapp-(phone hidden)
📌 Data Engineer (Tamil Nadu)
🏢 EVOKE HR
📍 Tamil Nadu