Data Engineering Lead (Pune)

Data Engineering Lead (Pune)

03 Sep
|
Calfus
|
Pune

03 Sep

Calfus

Pune

About Us:

At Calfus, we are known for delivering cutting-edge AI agents and products that transform businesses in ways previously unimaginable. We empower companies to harness the full potential of AI, unlocking opportunities they never imagined possible before the AI era. Our software engineering teams are highly valued by customers, whether start-ups or established enterprises, because we consistently deliver solutions that drive revenue growth. Our ERP solution teams have successfully implemented cloud solutions and developed tools that seamlessly integrate with ERP systems, reducing manual work so teams can focus on high-impact tasks.

None of this would be possible without talent like you! Our global teams thrive on collaboration, and we’re actively looking for skilled professionals to strengthen our in-house expertise and help us deliver exceptional AI, software engineering, and solutions using enterprise applications.

As one of the fastest-growing companies in our industry, we take pride in fostering a culture of innovation where new ideas are always welcomed—without hesitation. We are driven and expect the same dedication from our team members. Our speed, agility, and dedication set us apart, and we perform best when surrounded by high-energy, driven individuals.

To continue our rapid growth and deliver an even greater impact, we invite you to apply for our open positions and become part of our journey!

About the Role

We're looking for a Data Engineering Lead to architect and own enterprise-scale data platforms end to end — Databricks, cloud-native Lakehouse architecture, and modern orchestration — empowering the organization to make reliable, governed, data-driven decisions. You'll be the senior technical voice on the team, with BI and reporting as a downstream output of the platform you build, not the platform itself.

What You'll Do

- Data Platform Architecture: Design and own end-to-end data pipeline architecture — ETL/ELT, orchestration,



and data modeling — across cloud-native platforms (Azure Databricks, Azure Data Factory, and/or Snowflake).
- Lakehouse Engineering: Architect Medallion (Bronze/Silver/Gold) or equivalent layered data platform designs using Delta Lake; own Unity Catalog governance, access control, and data lineage.
- Data Integration: Oversee ETL/ELT pipeline development and orchestration (Apache Airflow, ADF), including legacy migration support (e.g., SSIS, Informatica, SAS) where applicable.
- Data Modelling: Own dimensional modeling and SCD (Type 1/2) implementation for accurate historical and current-state reporting.
- Database Management: Utilize SQL to write complex queries, stored procedures, and manage large-scale data transformations.
- Reporting Enablement: Ensure the platform reliably feeds downstream BI tools (Power BI or similar) with clean, curated, analytics-ready data.
- Collaboration: Work closely with stakeholders to gather requirements and translate them into technical specifications and architecture designs.
- Performance Optimization: Analyze and optimize Spark/pipeline workloads for performance, scalability, and reliability.
- Data Governance: Implement RBAC, data classification, and lineage tracking to ensure compliance and audit readiness.
- Team Leadership: Mentor and guide data engineers, fostering a culture of continuous learning and architectural rigor.

On your first day, we'll expect you to have:

- Bachelor's degree in Computer Science, Information Systems, Data Science, or a related field.
- 8+ years of experience in data engineering,



with demonstrated Lead-level ownership.
- Deep, hands-on expertise in Azure Databricks (PySpark, Delta Lake) and/or Snowflake, plus ADF or equivalent orchestration.
- Strong data modeling experience — dimensional modeling, SCD, Medallion or equivalent layered architecture.
- Strong proficiency in SQL, including advanced query writing and database management.
- Robust programming foundation in Python:

1. Data manipulation and analysis using Pandas, NumPy, and PySpark.
2. Data serialization and formats — JSON, CSV, Parquet.
3. Database interaction to query cloud-based data warehouses.
4. Pipeline orchestration using Airflow — scripting and automation.
5. Cloud services such as S3 and AWS Lambda for infrastructure management.

- Familiarity with cloud-native databases such as Snowflake, Postgres, Redshift.
- Code quality and management using version control and collaborative workflows.
- Full lifecycle experience — you've taken a design from blueprint through production support.

We'd be super excited if you have:

- Relevant cloud/platform certifications (Databricks, Snowflake, Azure).
- BI/reporting tool exposure (Power BI, Tableau, or similar).
- Experience with visualization tools such as QuickSight, Plotly, or Dash.
- Ability to interact with REST APIs and perform web scraping tasks.
- Azure SDK familiarity.

Benefits:

At Calfus, we value our employees and offer a robust benefits package. This includes medical, group, and parental insurance, coupled with gratuity and provident fund options. Further, we support employee wellness and provide birthday leave as a valued benefit.

Calfus Inc. is an Equal Opportunity Employer.

We believe diversity drives innovation. We’re committed to creating an inclusive workplace where everyone—regardless of background, identity, or experience—has the opportunity to thrive. We welcome all applicants!

📌 Data Engineering Lead (Pune)
🏢 Calfus
📍 Pune

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: data engineering lead (pune) / pune