Location: Pune / Chennai / Bangalore/Delhi NCR
Experience: 3.5+ years
Job Summary:
We are looking for a Data Engineer with strong expertise in AWS, PySpark, and Databricks. The ideal candidate will design, build, and optimize scalable data pipelines and work on large-scale distributed data processing systems.
Key Responsibilities:
Design and develop scalable data pipelines using PySpark
Build and maintain data solutions on Databricks platform
Work with AWS services for data storage, processing, and orchestration
Optimize data workflows for performance, scalability, and reliability
Handle large datasets and ensure productive data processing and transformation
Collaborate with cross-functional teams (data analysts, scientists, stakeholders)
Ensure data quality, integrity, and governance standards
Troubleshoot and improve existing data systems
Mandatory Skills:
Robust experience in PySpark
Hands-on expertise with Databricks
In-depth knowledge of AWS (S3, EMR, Glue, Lambda, etc.)
Solid experience in Data Engineering concepts and ETL pipelines
Positive understanding of distributed data processing
Strong SQL skills
Qualifications:
Bachelors/Master’s degree in Computer Science, Engineering, or related field
3.5+ years of experience in Data Engineering
Soft Skills:
Strong problem-solving and analytical ability
Effective communication and teamwork
Ability to work in a fast-paced environment
📌 Aws & Pyspark Data Engineer Immediate Joiners Only Pune
🏢 EY
📍 Pune