21 Aug
|
DeHaat
|
Gurugram
Role Overview:
We are looking for a Data Analyst / Data Engineer Intern to work closely with the Product & Technology team on building and maintaining data pipelines, internal dashboards, reporting systems, and data-driven workflows.
This is a hands-on role suited for someone who is technically strong, comfortable working with databases and data pipelines, and interested in understanding how data is used to build operational products.
Key Responsibilities
- Data Engineering & Data Pipelines A major part of the role will involve managing data flowing across multiple systems and databases.
- Build, maintain, and troubleshoot data pipelines across sources such as AWS RDS, Amazon Redshift, application databases, APIs, Google Sheets, and other internal systems.
- Write and optimize SQL queries for data extraction, transformation, reconciliation, and reporting.
- Develop Python scripts for ETL/ELT workflows, data transformation, validation, and automation.
- Manage and monitor scheduled cron jobs and recurring data workflows.
- Identify pipeline failures, missing records, duplicates, schema inconsistencies, and other data-quality issues.
- Create data validation and reconciliation checks between source systems and downstream dashboards.
- Work with APIs and JSON-based integrations to move data between different applications.
- Support documentation of data sources, schemas, transformations, and dependencies.
- Dashboarding & Frappe-Based Applications The second major responsibility will be building and maintaining operational dashboards and internal tools, primarily using Frappe Framework / ERPNext.
- Develop and maintain dashboards, reports, DocTypes, forms, workflows, and views in Frappe.
- Connect Frappe applications with databases, APIs,
and internal data pipelines.
- Build operational dashboards for Product, Technology, Field, and Business teams.
- Translate business requirements into structured datasets, KPIs, filters, and dashboard components.
- Debug data discrepancies between backend databases and dashboard outputs.
- Improve dashboard usability, performance, and data freshness.
- Support basic frontend customisation using JavaScript, HTML, CSS, and Frappe client/server scripting where required.
Additional Responsibilities
- Use tools such as Postman for API testing and debugging.
- Work with REST APIs, JSON payloads, authentication mechanisms, and webhook-based integrations.
- Perform product and data QA before releases.
- Support workflow automation across internal systems.
- Use AI tools such as ChatGPT, Claude, or similar LLMs for coding assistance, data analysis, documentation, debugging, and workflow automation.
- Prepare technical documentation, data dictionaries, SOPs, and implementation notes.
- Quickly understand new systems and work across multiple technology platforms.
##Preferred Skills
Must-have / Strongly Preferred
- Strong knowledge of SQL, including joins, aggregations, CTEs, subqueries, and data manipulation.
- Good working knowledge of Python, particularly for data processing and automation.
- Understanding of relational databases and data modelling.
- Experience working with ETL/ELT pipelines and scheduled data jobs.
- Strong proficiency in Microsoft Excel / Google Sheets.
- Understanding of REST APIs, JSON, and Postman.
- Strong analytical and logical problem-solving skills.
Good to Have
- Experience with Frappe Framework / ERPNext.
- Exposure to AWS RDS, Amazon Redshift, PostgreSQL, MySQL, or similar databases.
- Familiarity with cron jobs, schedulers, and basic Linux/server environments.
- Knowledge of JavaScript, HTML, and CSS.
- Experience building dashboards or internal reporting tools.
- Familiarity with Git/GitHub.
- Experience using AI coding/productivity tools such as ChatGPT, Claude, Cursor, or similar platforms.
## Candidate Profile
We are looking for someone who:
- Has robust computer science and data fundamentals.
- Enjoys working with databases and solving data problems.
- Can independently investigate why a pipeline, query, API, or dashboard is not working.
- Is comfortable learning unfamiliar technologies quickly.
- Pays close attention to data accuracy and edge cases.
- Can work across engineering, product, and business requirements rather than operating only as a traditional analyst.
- Is comfortable taking ownership of small technical projects from requirement gathering through implementation and testing.
## Internship Details
- Role: Data Analyst / Data Engineer Intern
- Duration: 6–9 months
- Location: Gurugram – Work from Office
- Joining: Immediate
- Team: Product & Technology
The role will provide significant hands-on exposure to production data pipelines, AWS databases, Frappe-based applications, APIs, operational dashboards, workflow automation, and AI-assisted development.
📌 Data Analyst/Data Engineer Intern (Gurugram)
🏢 DeHaat
📍 Gurugram