05 Sep
|
Hireologist
|
Hyderabad
05 Sep
Hireologist
Hyderabad
Job Description
Job Title: Senior Data Engineer
n
Experience : 5+ years
n
Location : Hyderabad
n
n
About the Role
n
We’re looking for a Data Engineer to join our team and help build the data foundation that powers analytics, experimentation, machine learning, and operational decision-making .
n
n
What You’ll Do
n
n
n
- Build trusted data products: Design and maintain dbt models that produce reliable datasets, metrics, and features used by Data Science, Analytics, ML, and business teams.
n
- Develop scalable data pipelines: Build and operate Databricks and PySpark pipelines that transform raw operational events into high-quality, analysis-ready data.
n
- Understand the business: Develop deep familiarity with company's operations so that the schemas, datasets, and models you build accurately represent how the business works.
n
- Orchestrate workflows: Build and manage end-to-end data workflows using Airflow, with Prefect where appropriate, ensuring reliable SLAs for daily models, dashboards, and operational decisions.
n
- Partner with Data Science: Collaborate with Data Scientists to design, productionize, and maintain feature pipelines and the data infrastructure supporting ML models.
n
- Raise the quality bar: Participate in code reviews and improve data quality, testing, observability, documentation, and engineering standards across the team.
n
- Optimize for scale: Improve dbt and Spark workloads for performance, cost, reliability, and maintainability as data volumes grow.
n
- Learn and adapt: Quickly pick up new technologies and approaches by working with teammates, documentation, experimentation, and hands-on problem solving.
n
- Embrace AI-native engineering: Use modern AI coding assistants and agentic development workflows to improve productivity, engineering quality, and experimentation.
n
n
n
What You Bring
n
n
- 5+ years of experience building, testing, and deploying data engineering systems in production.
n
- Strong SQL skills and production experience with one or more of PySpark, dbt, or Airflow .
n
- Experience working with at least one distributed data system , with a solid understanding of concepts such as consistency, latency, throughput, scalability, and fault tolerance.
n
- Experience with Infrastructure as Code , using technologies such as Terraform, AWS CDK, or Pulumi.
n
- Solid interest in or understanding of supply chain, logistics, or operational data challenges .
n
- A self-starter mindset with the ability to take ownership, move quickly, and ship high-quality solutions.
n
- Strong collaboration skills and enthusiasm for solving complex, novel problems with teammates.
n
- Excellent written and verbal communication skills in English.
n
- Experience with modern data technologies such as:
n
- dbt, Databricks, PySpark / Spark, Airflow, Prefect, SQL, Delta Lake / Apache Iceberg
n
- Familiarity with the broader data stack — including Kinesis, EMR, Sigma, or Pulumi — is a plus.
n
- Comfortable using modern AI coding assistants such as Claude Code, Cursor, GitHub Copilot, or similar tools .
n
- Experience with AI-native engineering workflows, including prompting, agentic tooling, evaluation, retrieval, and AI-assisted development .
n
- A demonstrated interest in or experience with LLMs , evaluating model outputs, or integrating AI into data pipelines and internal engineering tools is a plus.
n
n
n
Nice to Have
n
n
- Experience designing or working with data lakehouse architectures , particularly Delta Lake or Apache Iceberg.
n
- Production experience with Kinesis or other streaming technologies .
n
- Exposure to MLOps or working closely with ML/AI teams.
n
- Experience collaborating with US-based engineering teams across multiple time zones .
n
n
📌 Senior Data Engineer (Hyderabad)
🏢 Hireologist
📍 Hyderabad