An experienced
Data Engineer
with solid expertise in
Spark, Scala, and Cloudera Data Platform (CDP)
or a strong
Scala
background. The ideal candidate will be responsible for designing, developing, and optimizing scalable data pipelines and distributed data processing solutions.
Key Responsibilities
- Design, develop, and maintain scalable data pipelines using
Scala
and
Apache Spark
.
- Build and optimize batch and real-time data processing solutions.
- Work with structured and unstructured datasets to deliver high-quality data solutions.
- Optimize Spark jobs for performance, scalability, and reliability.
- Collaborate with cross-functional teams to understand business requirements and implement data engineering solutions.
- Ensure data quality, governance, and best coding practices.
- Troubleshoot production issues and provide performance tuning.
- Participate in code reviews and mentor junior team members.
Required Skills
- 5+ years of experience in Data Engineering.
- Strong hands-on experience with
Scala
.
- Extensive experience with
Apache Spark
.
- Strong SQL programming skills.
- Experience with distributed data processing and ETL development.
- Good understanding of Hadoop ecosystem concepts.
- Experience with Git and CI/CD practices.
- Strong debugging and performance tuning skills.
- Excellent analytical and problem-solving abilities.
Good to Have
- Experience with cloud platforms (AWS/Azure/GCP).
- Knowledge of Kafka or other streaming technologies.
- Experience with Airflow or similar orchestration tools.
- Exposure to Delta Lake, Databricks, or Iceberg.
📌 Data engineer (Pimpri-Chinchwad)
🏢 Rosemallow Technologies
📍 Pimpri-Chinchwad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.