- Design, build, and maintain scalable data pipelines using PySpark and Python
- Develop and optimize complex SQL queries for large datasets
- Implement and manage ETL/ELT processes ensuring data quality and reliability
- Collaborate with business and product teams to translate data requirements into solutions
- Build and maintain data warehouse solutions
- Handle large-scale data processing using Hadoop/Big Data technologies
- Perform performance tuning and optimization of data workflows
Required Skills:
- Robust hands-on experience with PySpark and Python
- Advanced proficiency in SQL
- Solid experience in ETL processes and data warehousing
- Familiarity with Hadoop ecosystem and Big Data technologies
- Experience working with large datasets in distributed environments
- Good communication and business understanding
Good to Have:
- Experience with Apache Airflow
- Exposure to cloud platforms (AWS, GCP, Azure)
- Knowledge of data lakes and up-to-date data architectures
- Experience with streaming tools (Kafka, Spark Streaming)
At Indium, diversity, equity, and inclusion (DEI) are core values. We promote DEI through a dedicated council, expert sessions, and tailored training, fostering an inclusive workplace. Our initiatives, like the WE@IN women empowerment program and DEI calendar, create a culture of respect and belonging. Recognized with the Human Capital Award, we're committed to an environment where everyone thrives. Join us in building a diverse, innovative workplace.
Indium do not solicit or accept any form of fees or payment at any stage of the hiring process.
📌 Data Engineer (India)
🏢 Indium software
📍 India
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.