Required Skill Set: Data Lake architecture, Azure Services ADLS, ADF, Azure Databricks, Synapse
ADF, ADB, SQL, Pyspark, python
Good Knowledge of Data Brick lake house and Azure Data Lake concept
Knowledge of Data Bricks delta concept
Delta live tables (DLT)
Solid hands-on experience in ELT pipeline development using Azure Data factory and Databricks Autoloader, Notebook scripting and Azure Synapse Activity Copy, Data Flow Task
Robust knowledge of metadata-driven data pipeline, metadata management, dynamic logic
In-depth knowledge of data storage solutions, including Azure Data Lake Storage (ADLS), and Azure Serverless SQL Pool.
Experience with data transformation using Spark, and SQL technologies. • Solid understanding of design patterns,
and best practices of the cloud stack.
Experience with code management and version control using Git or similar tools.
Strong problem-solving and debugging skills in ETL workflows and data pipelines.
Robust understanding of Azure Data bricks and Azure Synapse internals features and capabilities.
Knowledge of Azure DevOps and continuous integration and deployment (CI/CD) process.
Knowledge of data quality and data profiling techniques, with experience in data validation and data cleansing.