15 Sep
|
Genpact
|
Bengaluru
Role & responsibilities In this role, the Databricks Developer is responsible for solving the real-world cutting-edge problem to meet both functional and non-functional requirements.
Responsibilities
- Maintains close awareness of new and emerging technologies and their potential application for service offerings and products.
- Work with architect and lead engineers for solutions to meet functional and non functional requirements.
- Demonstrated knowledge of relevant industry trends and standards.
- Demonstrate strong analytical and technical problem-solving skills.
- Must have experience in Data Engineering domain.
Preferred candidate profile
- Must have experience in Data Engineering domain .
- Must have implemented at least 4 project end-to-end in Databricks.
- Must have at least experience on databricks which consists of various components as below
- Must have skills: Azure data factory, Azure data bricks, Python and Pyspark
- Expert with database technologies and ETL tools.
- Hands-on experience on designing and developing scripts for custom ETL processes and automation in Azure data factory, Azure databricks, Python, Pyspark etc.
- Good knowledge of AZURE, AWS, GCP Cloud platform services stack
- Hands-on experience on designing and developing scripts for custom ETL processes and automation in Azure data factory, Azure databricks, Delta lake, Databricks workflows orchestration, Python,
Pyspark etc.
- Good Knowledge on Unity Catalog implementation.
- Good Knowledge on integration with other tools like DBT, other transformation tools.
- Good knowledge on Unity Catalog integration with Snowlflake
- Must be well versed with Databricks Lakehouse concept and its implementation in enterprise environments.
- Must have good understanding to create complex data pipeline
- Must have good knowledge of Data structure & algorithms.
- Must be strong in SQL and sprak-sql.
- Must have strong performance optimization skills to improve efficiency and reduce cost.
- Must have worked on both Batch and streaming data pipeline.
- Must have extensive knowledge of Spark and Hive data processing framework.
- Must have worked on any cloud (Azure, AWS, GCP) and most common services like ADLS/S3, ADF/Lambda, CosmosDB/DynamoDB, ASB/SQS, Cloud databases.
- Must be strong in writing unit test case and integration test
- Must have strong communication skills and have worked on the team of size 5 plus
- Must have great attitude towards learning recent skills and upskilling the existing skills.
Preferred
Qualifications
- Good to have Unity catalog and basic governance knowledge.
- Good to have Databricks SQL Endpoint understanding.
- Good To have CI/CD experience to build the pipeline for Databricks jobs.
📌 Senior Principal Consultant - Databricks Developer (Bengaluru)
🏢 Genpact
📍 Bengaluru