01 Aug
|
Important Business
|
Pune
01 Aug
Important Business
Pune
Job Title: PySpark Data Engineer
Experience: 6+ Years
Location: Hyderabad/ Pune
Employment Type: Full -Time
Job Summary:
We are looking for a skilled and experienced PySpark Data Engineer to join our growing data engineering team. The ideal candidate will have 6+ years of experience in designing and implementing data pipelines using PySpark, AWS Glue, and Apache Airflow, with solid proficiency in SQL. You will be responsible for building scalable data processing solutions, optimizing data workflows, and collaborating with cross -functional teams to deliver high -quality data assets.
Requirements
Key Responsibilities:
- Design, develop, and maintain large -scale ETL pipelines using PySpark and AWS Glue.
- Orchestrate and schedule data workflows using Apache Airflow.
- Optimize data processing jobs for performance and cost -efficiency.
- Work with large datasets from various sources, ensuring data quality and consistency.
- Collaborate with Data Scientists, Analysts, and other Engineers to understand data requirements and deliver solutions.
- Write efficient, reusable, and well -documented code following best practices.
- Monitor data pipeline health and performance; resolve data -related issues proactively.
- Participate in code reviews, architecture discussions, and performance tuning.