16 Aug
|
AppInventiv
|
Delhi
Job Overview
Design, build, and maintain robust, scalable, and high-performance data pipelines (batch and streaming).
? Develop and optimize ETL/ELT workflows using contemporary orchestration tools (e.g., Airflow, dbt, ADF, Databricks, Prefect).
? Collaborate with Data Scientists, Analysts, and Business stakeholders to ensure data models support analytical and AI/ML needs.
? Integrate data from multiple sources (SQL/NoSQL databases, APIs, flat files, etc.) into centralized warehouses or lakehouses.
? Implement and enforce data quality, integrity, and governance standards across the pipeline.
? Work with cloud platforms (Azure, AWS, or GCP) for data storage, compute, and orchestration.
? Contribute to data schema design, query optimization, and performance tuning.
? Maintain detailed technical documentation for pipelines, jobs, and data models.
? Actively participate in code reviews, design discussions, and production support.
Job Qualification
- 1-3 years of experience
Programming: Python (Pandas, PySpark, SQLAlchemy), SQL (advanced query writing
and optimization)
? ETL / Orchestration Tools: Apache Airflow, Azure Data Factory, or Prefect
? Data Processing Frameworks: PySpark / Spark
📌 Data Engineer (Delhi)
🏢 AppInventiv
📍 Delhi