17 Sep
|
Summit Lake
|
Noida
About the Company
Summit Lake Digital is an Databricks + Anthropic consulting partner helping US-based clients turn complex enterprise data into governed, analytics- and AI-ready lakehouses.
We’re building our India engineering team and looking for a Senior Data Engineer who enjoys being hands-on, solving real client problems, and taking ownership from design through production. Joining us early means your contribution goes beyond delivering pipelines—you’ll help shape our engineering practices, reusable solutions, and how we deliver for clients.
About the Role
- You’ll build the pipelines, ingestion frameworks, and transformations at the heart of our client engagements. Working closely with the Lead Data Engineer / Architect, you’ll turn architectural direction into reliable, tested, production-ready data products.
- This is a hands-on engineering role for someone who enjoys both building and improving: asking thoughtful questions, working through evolving requirements, and following a solution through to delivery.
Why join us?
- Ownership with visible impact: Take responsibility for substantial parts of client solutions and see your work put into production.
- A voice in how we build: Contribute ideas, improve standards, and help establish engineering practices as the company grows.
- Technical depth: Work across Databricks, cloud data platforms, governance, and the data foundations that support AI.
- Collaboration and growth: Work closely with the Lead Data Engineer / Architect, with room to grow toward technical leadership.
- Client exposure:
Collaborate with client stakeholders to understand business challenges and turn them into reliable data solutions.
Your role
You’ll build the pipelines, ingestion frameworks, and transformations at the heart of our client engagements. Working closely with the Lead Data Engineer / Architect, you’ll turn architectural direction into reliable, tested, production-ready data products.
This is a hands-on engineering role for someone who enjoys both building and improving: asking thoughtful questions, working through evolving requirements, and following a solution through to delivery.
What you’ll do
- Build and maintain bronze, silver, and gold pipelines on Databricks using PySpark, SQL, and Delta Lake.
- Integrate enterprise sources such as Salesforce, QuickBooks, Accounting Seed, UKG, travel and expense systems, and files using Lakeflow Connect, Auto Loader, Azure Data Factory, AWS Glue, or partner connectors.
- Implement access controls and PII handling through Unity Catalog.
- Develop tested, documented transformations, with a focus on data quality, maintainability, and performance.
- Troubleshoot production issues and tune Spark workloads for reliability and efficiency.
- Use Git and Databricks Asset Bundles to support CI/CD and consistent deployments.
- Partner with analytics engineers to make curated gold models available through Power BI and Genie.
- Collaborate with client subject matter experts to clarify requirements and validate outcomes.
- Contribute reusable components and practical engineering standards that help the team scale.
What you’ll bring
- 4+ years of data engineering experience, including 2+ years of hands-on Databricks experience.
- Strong PySpark, advanced SQL, and Delta Lake skills, with practical knowledge of medallion architecture.
- Experience with Databricks Workflows / Lakeflow and Git-based CI/CD, plus exposure to Unity Catalog.
- Experience with Azure or AWS—either is welcome.
- A strong ownership mindset: you follow through, raise issues early, and care about the quality of what you deliver.
- Clear English communication and confidence collaborating with technical and business stakeholders.
- Comfort in an early-stage environment where requirements evolve and your ideas can help shape the approach.
- Availability to work from India with planned overlap during US business hours.
Outstanding to have
- Databricks Certified Data Engineer Associate or Professional certification.
- Terraform or other infrastructure-as-code experience.
- Deeper expertise in Spark performance tuning.
- Familiarity with Salesforce, finance, or HR data.
- Exposure to Mosaic AI or data workflows supporting generative AI.
Ready to help build what comes next? If you want hands-on technical work, meaningful ownership, and the opportunity to contribute to a company from an early stage, we’d love to hear from you.
📌 Data Engineer (Noida)
🏢 Summit Lake
📍 Noida