Required Skill Set: Data Lake architecture, Azure Services ADLS, ADF, Azure Databricks, Synapse
ADF, ADB, SQL, Pyspark, python
Good Knowledge of Data Brick lake house and Azure Data Lake concept
- Knowledge of Data Bricks delta concept
Delta live tables (DLT)
- Strong hands-on experience in ELT pipeline development using Azure Data factory and Databricks Autoloader, Notebook scripting and Azure Synapse Activity Copy, Data Flow Task
- Robust knowledge of metadata-driven data pipeline, metadata management, dynamic logic
- In-depth knowledge of data storage solutions, including Azure Data Lake Storage (ADLS), and Azure Serverless SQL Pool.
- Experience with data transformation using Spark, and SQL technologies. • Solid understanding of design patterns,
and best practices of the cloud stack.
- Experience with code management and version control using Git or similar tools.
- Strong problem-solving and debugging skills in ETL workflows and data pipelines.
- Strong understanding of Azure Data bricks and Azure Synapse internals features and capabilities.
- Knowledge of Azure DevOps and continuous integration and deployment (CI/CD) process.
- Knowledge of data quality and data profiling techniques, with experience in data validation and data cleansing.