Role Overview - Databricks:
• We are seeking a skilled Databricks Engineer to design, develop, and optimize large scale data processing solutions using the Databricks Lakehouse Platform. The ideal candidate will have solid experience in big data technologies, cloud platforms, and data engineering best practices to enable analytics, machine learning, and data driven decision making.
Key Responsibilities
• Design, develop, and maintain data pipelines using Databricks, Apache Spark, and Delta Lake • Build ETL/ELT workflows to ingest, transform, and process structured and unstructured data • Optimize Spark jobs for performance, scalability, and cost efficiency • Implement data models and manage data lakes/ lakehouses using Delta tables • Collaborate with data scientists and analysts to support machine learning and advanced analytics • Integrate Databricks with cloud services (Azure, AWS, or GCP) • Ensure data quality, reliability, and governance across pipelines • Implement CI/CD and version control for Databricks notebooks and jobs • Monitor jobs, troubleshoot failures, and resolve performance issues • Follow security and compliance best practices Required Skills &
Qualifications
• Bachelor's degree in Computer Science, Engineering, or related field • 5+ years of experience in Data Engineering or Big Data development • Strong hands on experience with Databricks • Proficiency in Apache Spark and distributed data processing • Strong programming skills in Python and/or Scala • Knowledge of cloud platforms (AWS, Azure, or GCP).
• Experience working with Delta Lake, Parquet, ORC • Solid knowledge of SQL for data transformation and analysis • Experience with cloud platforms (Azure Databricks / AWS Databricks / GCP Dataproc)
• Understanding of data warehousing concepts and data modeling • Familiarity with Git and CI/CD practices Preferred Qualifications:
• Experience with Apache Airflow, Azure Data Factory, or similar orchestration tools • Knowledge of machine learning workflows on Databricks • Experience with Unity Catalog, data governance, and access controls • Databricks or cloud certifications (Azure/AWS/GCP) • Exposure to real-time data processing (Kafka, Spark Streaming) ECMS REQ ID 547517 Project Location 1 MAH | PUNE Project Location 2 (No Value) Project Location 3 (No Value) Project Location 4 (No Value) Delivery SPOC Mahesh_kakade Relevant Experience 7-8 years Mandatory skills Apache Spark (Core)
Databricks Platform
Programming Languages
Data Engineering Fundament als
Delta Lake
Cloud Fundamentals (at least one)
Version Control
Desired skills Structured Streaming
Performance Optimization
Databricks Advanced Features
CI/CD & Autom ation
Data Modeling & Analytics
DevOps / MLOps Exposure
Security & Compliance
Domain (Industry) MFG BGC (Before onboarding / After onboarding) PostOB Total Experience (Ex. 5-7 Years) 7-8 years BGC Details Education Verification - Highest Education Earned
- Employment Verification - Last 5 years of th e employment check to be verified (Recent to Past)
- Criminal Record check - Criminal record check should be performed through law firms / e-COURT for last 7 years where ever the candidate has resided
- Global & India Database Check - Both India specific and global database searches like CBI, FBI, Interpol most wanted, OFAC, SDN, media & internet searches etc. to be verify
BGC Vendor Any NASSCOM registered vendor Mode of Interview Face to Face WFO / WFH / Hybrid WFO Please enter shift timings General shift Shift Timings General
📌 Databricks (India)
🏢 Clifyx
📍 India