Location: Pune / Chennai / Bangalore/Delhi NCR
Experience: 3.5+ years
Job Summary:
We are looking for a Data Engineer with solid expertise in AWS, PySpark, and Databricks. The ideal candidate will design, build, and optimize scalable data pipelines and work on large-scale distributed data processing systems.
Key Responsibilities:
- Design and develop scalable data pipelines using PySpark
- Build and maintain data solutions on Databricks platform
- Work with AWS services for data storage, processing, and orchestration
- Optimize data workflows for performance, scalability, and reliability
- Handle large datasets and ensure efficient data processing and transformation
- Collaborate with cross-functional teams (data analysts, scientists, stakeholders)
- Ensure data quality, integrity, and governance standards
- Troubleshoot and improve existing data systems
Mandatory Skills:
- Strong experience in PySpark
- Hands-on expertise with Databricks
- In-depth knowledge of AWS (S3, EMR, Glue, Lambda, etc.)
- Solid experience in Data Engineering concepts and ETL pipelines
- Good understanding of distributed data processing
- Strong SQL skills
Qualifications:
- Bachelors/Master’s degree in Computer Science, Engineering, or related field
- 3.5+ years of experience in Data Engineering
Soft Skills:
- Strong problem-solving and analytical ability
- Effective communication and teamwork
- Ability to work in a fast-paced environment
📌 AWS & Pyspark- Data Engineer-Immediate joiners only (Bengaluru)
🏢 EY
📍 Bengaluru