13 Sep
|
Pine Labs
|
Uttar Pradesh
13 Sep
Pine Labs
Uttar Pradesh
Role Purpose
We are looking for a highly skilled and motivated Data engineering leader with 5-10 years of experience to join our growing team. You will create data pipelines to move data from source to data lake, Maintain the data lake, real time transform the data to uncover actionable insights and contribute to strategic decision-making across various departments. The ideal candidate will have strong analytical and programming skills, a solid understanding of data pipeline and programming logic, and the ability to collaborate effectively with cross-functional teams.
You will be helping all application teams by providing data interface, also will be ensuring that data analytics team is getting required data for their initiatives The responsibilities we entrust you
Own the data engineering roadmap – from streaming ingestion to analytics-ready layers. Lead and evolve our real-time data pipelines using Apache Kafka/MSK, Debezium CDC, and Apache Pinot. Design and maintain custom-built Java/Python-based frameworks for data ingestion, validation, transformation, and replication.
Design and manage data lakes using Apache Iceberg, Glue, and Athena over S3. Define and enforce data modelling standards across Pinot, Redshift and warehouse layers. Lead implementation and scaling of open-source tools and orchestration platforms (Airflow).
Drive cloud-native deployment strategies using ECS/EKS, Terraform/CDK, Docker. Collaborate with ML, product, and business teams to support advanced analytics and AI use cases. Mentor data engineers and evangelize modern architecture and engineering best practices.
Ensure observability, lineage, data quality, and governance across the data stack.
Wha mattters in this role:
Must-Have:
Proven expertise in Kafka/MSK, Debezium, and real-time event-driven architectures Hands-on experience with Pinot, Redshift, Rocks DB or no SQL DB Strong background in custom tooling using Java and Python Experience with Apache Airflow, Superset, Iceberg, Athena, and Glue Strong AWS ecosystem knowledge (IAM, S3, Lambda, Glue, ECS/EKS, Cloud Watch, etc.) Deep understanding of data lake architecture, streaming vs batch processing, CDC concepts Familiarity with modern data formats (Parquet, Avro) and storage abstractions Nice-to-Have:
Exposure to dbt, Trino/Presto, Click House, or Druid Familiarity with data security practices, encryption at rest/in-transit, GDPR/PCI compliance Experience with Dev Ops practices, Git Hub Actions, CI/CD, and Terraform/Cloud Formation Education
Bachelor's or Master's degree in Computer Science, Engineering, or a related field. Communication & Ownership:
Strong communication, interpersonal, and conflict resolutio n skills Able to work both independently and as part of team Location: Sector 62, Noida
Things you should be comfortable with:
Working from office: 5 days a week Pushing the boundaries: Have a big idea? See something that you feel we should do but haven't done? We will hustle hard to make it happen. We encourage out of the box thinking, and if you bring that with you, we will make sure you get a bag that fits all the energy you bring along. What we value inour people:
You take the shot: You decide rapid and deliver right. You are the CEO of what you do: You show ownership and make things happen. You sing your work like an artist: You seek to learn and take pride in thework you do.
📌 Lead Engineer - Data Engineering (Uttar Pradesh)
🏢 Pine Labs
📍 Uttar Pradesh