- Design and develop Big Data solutions using Hadoop technologies.
- Develop and maintain data ingestion and ETL pipelines.
- Process large-scale structured and unstructured datasets.
- Build Hive scripts, MapReduce jobs, and Spark applications.
- Import and export data using Sqoop and other ingestion tools.
- Implement real-time and batch data processing solutions.
- Optimize Hadoop cluster performance and data processing workflows.
- Work with HDFS, Hive, HBase, and Kafka environments.
- Perform data validation, cleansing, and quality checks.
- Troubleshoot production issues and optimize job execution.
- Collaborate with data engineers, analysts, and business stakeholders.
- Create technical documentation and support deployment activities.