18 Sep
|
DeHaat
|
Gurugram
Job Description
Role Overview:
/n
We are looking for a Data Analyst / Data Engineer Intern to work closely with the Product & Technology team on building and maintaining data pipelines, internal dashboards, reporting systems, and data-driven workflows.
/n
This is a hands-on role suited for someone who is technically strong, comfortable working with databases and data pipelines, and interested in understanding how data is used to build operational products.
/n
Key Responsibilities
/n
1. Data Engineering & Data Pipelines
/n
A major part of the role will involve managing data flowing across multiple systems and databases.
/n
/n
- Build, maintain, and troubleshoot data pipelines across sources such as **AWS RDS, Amazon Redshift, application databases, APIs, Google Sheets, and other internal systems**.
/n
- Write and optimize SQL queries for data extraction, transformation, reconciliation, and reporting.
/n
- Develop Python scripts for **ETL/ELT workflows, data transformation, validation, and automation**.
/n
- Manage and monitor scheduled **cron jobs and recurring data workflows**.
/n
- Identify pipeline failures, missing records, duplicates, schema inconsistencies, and other data-quality issues.
/n
- Create data validation and reconciliation checks between source systems and downstream dashboards.
/n
- Work with APIs and JSON-based integrations to move data between different applications.
/n
- Support documentation of data sources, schemas, transformations, and dependencies.
/n
/n
1. Dashboarding & Frappe-Based Applications
/n
The second major responsibility will be building and maintaining operational dashboards and internal tools, primarily using Frappe Framework / ERPNext.
/n
/n
- Develop and maintain dashboards, reports, Doc Types, forms, workflows, and views in Frappe.
/n
- Connect Frappe applications with databases, APIs, and internal data pipelines.
/n
- Build operational dashboards for Product, Technology, Field, and Business teams.
/n
- Translate business requirements into structured datasets, KPIs, filters, and dashboard components.
/n
- Debug data discrepancies between backend databases and dashboard outputs.
/n
- Improve dashboard usability, performance, and data freshness.
/n
- Support basic frontend customisation using Java Script, HTML, CSS, and Frappe client/server scripting where required.
/n
/n
Additional Responsibilities
/n
/n
- Use tools such as Postman for API testing and debugging.
/n
- Work with REST APIs, JSON payloads, authentication mechanisms, and webhook-based integrations.
/n
- Perform product and data QA before releases.
/n
- Support workflow automation across internal systems.
/n
- Use AI tools such as ChatGPT, Claude, or similar LLMs for coding assistance, data analysis, documentation, debugging, and workflow automation.
/n
- Prepare technical documentation, data dictionaries, SOPs, and implementation notes.
/n
- Quickly understand new systems and work across multiple technology platforms.
/n
/n
##Preferred Skills
/n
Must-have / Strongly Preferred
/n
/n
- Strong knowledge of SQL, including joins, aggregations, CTEs, subqueries, and data manipulation.
/n
- Good working knowledge of Python, particularly for data processing and automation.
/n
- Understanding of relational databases and data modelling.
/n
- Experience working with ETL/ELT pipelines and scheduled data jobs.
/n
- Strong proficiency in Microsoft Excel / Google Sheets.
/n
- Understanding of REST APIs, JSON, and Postman.
/n
- Robust analytical and logical problem-solving skills.
/n
/n
Good to Have
/n
/n
- Experience with Frappe Framework / ERPNext.
/n
- Exposure to AWS RDS, Amazon Redshift, PostgreSQL, MySQL, or similar databases.
/n
- Familiarity with cron jobs, schedulers, and basic Linux/server environments.
/n
- Knowledge of Java Script, HTML, and CSS.
/n
- Experience building dashboards or internal reporting tools.
/n
- Familiarity with Git/Git Hub.
/n
- Experience using AI coding/productivity tools such as ChatGPT, Claude, Cursor, or similar platforms.
/n
/n
## Candidate Profile
/n
We are looking for someone who:
/n
/n
- Has strong computer science and data fundamentals.
/n
- Enjoys working with databases and solving data problems.
/n
- Can independently investigate why a pipeline, query, API, or dashboard is not working.
/n
- Is comfortable learning unfamiliar technologies quickly.
/n
- Pays close attention to data accuracy and edge cases.
/n
- Can work across engineering, product, and business requirements rather than operating only as a traditional analyst.
/n
- Is comfortable taking ownership of small technical projects from requirement gathering through implementation and testing.
/n
/n
## Internship Details
/n
/n
- Role: Data Analyst / Data Engineer Intern
/n
- Duration: 6–9 months
/n
- Location: Gurugram – Work from Office
/n
- Joining: Immediate
/n
- Team: Product & Technology
/n
/n
The role will provide significant hands-on exposure to production data pipelines, AWS databases, Frappe-based applications, APIs, operational dashboards, workflow automation, and AI-assisted development.
📌 Data Analyst/Data Engineer Intern (Gurugram)
🏢 DeHaat
📍 Gurugram