- Design, develop, test, deploy and maintain large-scale data pipelines using AWS Glue to extract insights from various sources.
- Collaborate with cross-functional teams to identify business requirements and design solutions that meet those needs.
- Develop ETL processes using PySpark on Amazon EC2 instances running Airflow workflows in an EMR cluster.
- Troubleshoot issues related to data quality, performance, and scalability of the pipeline.
Job Requirements :
- 6-10 years of experience in Data Engineering with expertise in AWS Glue.
- Robust understanding of SQL concepts and ability to write complex queries for querying large datasets.
- Experience working with big-data technologies such as Hadoop (Hive), Spark (PySpark) and NoSQL databases like Redshift.
📌 AWS Data Engineer (Hyderabad)
🏢 Naukri Assist
📍 Hyderabad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.