01 Oct
|
Sama AI
|
Chennai
Role Overview
Serves as the core technical authority for designing, modernizing, and optimizing the Databricks Lakehouse architecture, code refactoring, and data execution pipelines.
Key Roles & Responsibilities
- Refactoring & Modernization: Refactor legacy Synapse T-SQL stored procedures, scripts, and views into native PySpark, Databricks SQL, or Delta Live Tables (DLT).
- Storage Optimization: Transition legacy Synapse distribution schemas (HASH, ROUND_ROBIN, REPLICATE) to Delta Lake Liquid Clustering (CLUSTER BY) for optimal query performance and reduced maintenance.
- Pipeline Design: Design productive batch and streaming ingestion patterns leveraging Auto Loader and Delta Lake Lakehouse mechanisms.
- Performance Tuning: Execute comprehensive tuning using OPTIMIZE, Z-ORDER, compute cluster sizing, and Serverless warehouse setup.
EDUCATION & QUALIFICATIONS
Bachelor's degree in Computer Science/IT/MSC/MBA. Any Related Field.
Key Skills Required
- Architecture & Governance: End-to-end Target State Architecture blueprints utilizing Unity Catalog and Delta Live Tables; articulating design decisions, build vs. buy evaluations, and balancing performance against licensing costs.
- Databricks Ecosystem: Databricks Runtime, Delta Lake, Delta Live Tables (DLT).
- Performance & Storage: Liquid Clustering, Partitioning Strategies, Auto Loader, DB SQL Serverless.
- Programming Languages: Python, PySpark, Advanced SQL (Scala is optional).
CI/CD & Tools: Databricks Asset Bundles (DABs), dbx, Git integration, MLflow
Pay: ₹1,500,000.00 - ₹3,000,000.00 per year
Application Question(s)
- Immediate to max 30 days preferred
Experience:
- data bricks: 10 years (Preferred)
Work Location: In person
📌 Data bricks Architect (Chennai)
🏢 Sama AI
📍 Chennai