06 Sep
|
Sparix Global
|
India
06 Sep
Sparix Global
India
Job Summary (List Format):
- Lead the development of data pipelines to support data science and analytics use cases, primarily using Hadoop ecosystem tools (Hadoop, HDFS, Hive, Impala).
- Mandatory expertise in Scala for developing, testing, and implementing production-ready code; knowledge of Java and Spark also required.
- Design and implement data ingestion pipelines using tools like SQOOP and Spark.
- Act as a technical expert and squad lead: Oversee a team of 3-5 data engineers with minimal supervision, guiding them through data transformation projects on Hadoop.
- Participate in solution design and requirements analysis for data engineering projects.
- Integrate and collaborate with IT and business project teams to deliver data products.
- Utilize CI/CD tools such as Git, Jenkins, and Nexus for deployment and version control.
- Apply Agile methodologies using Jira and Confluence for project management and documentation.
- Work with relational databases and demonstrate advanced SQL skills.
- Good-to-have: Experience with Unix and HDFS commands, Pyspark, SAS, and exposure to data analytics and visualization tools like Tableau.
- Solid communication skills required for team collaboration and stakeholder interaction.
📌 Spark/Scala Developer (India)
🏢 Sparix Global
📍 India