07 Aug
|
hackajob
|
Hyderabad
07 Aug
hackajob
Hyderabad
MAKE AN IMPACT
Job Description:
We are looking for a robust Datawarehousing resource with 7 to 8 years of hands-on experience in Dataiku DSS, strong SQL programming skills in PostgreSql or Teradata databases, good hands-on on python programming and basic tableau reporting.
- The candidate will be a core member of the Data Analytics & Reporting Tribe, working on the design, development and operationalization of end to end data solutions using Dataiku DSS.
- The candidate should be able to independently design, build, debug, maintain, and enhance analytics solutions while collaborating effectively with business teams.
- The role focuses on translating business requirements into robust, scalable data pipelines, data modeling and self service analytics while ensuring data quality, governance and performance standards of the BNPParibas group.
- We are looking for top-notch, energetic talent to keep up with the industrys momentum and support the team efforts.
Responsibilities:
- Design, build and maintain data pipelines, recipes and flows in Dataiku DSS to ingest, transform and expose data from source systems (PostgreSQL, Teradata, flat files, etc.) into the enterprise data warehouse / lake.
- Implement data warehousing concepts (dimensional modelling, slowly changing dimensions, star/snowflake schemas, data vault) to support analytical reporting and BI consumption.
- Develop reusable Python code (Pandas, PySpark, SQLAlchemy) within Dataiku plugins, recipes and notebooks for data cleansing, enrichment and advanced analytics.
- Ensure data quality and consistency by defining validation rules, unit tests and automated data quality checks inside Dataiku.
- Collaborate with data architects and business analysts to translate functional requirements into technical specifications, functional designs and implementation plans.
- Document technical and functional artefacts (design specs, data dictionaries, operational run books) and contribute to knowledge base for the Data Tribe.
- Set up CI/CD pipelines (GitLab) for Dataiku projects to automate testing, versioning and deployment across environments (dev / test / prod).
- Monitor and optimise pipeline performance, storage utilisation (Parquet/ORC) and cost on the cloud infrastructure (Kubernetes Engine, IBM Cloud).
- Support production incidents and perform root cause analysis, applying fixes and preventive measures.
- Mentor junior team members and champion best practices in data engineering and Dataiku usage.
Contributing Responsibilities
- Act as a team player and actively participate in squad ceremonies (stand ups, sprint planning, retrospectives).
- Adhere to Group standards (coding conventions, security, data governance, CI/CD).
- Promote a culture of continuous learning, sharing insights about Dataiku, data warehouse trends and emerging technologies.
- Work closely with cross functional teams (Data Scientist, BI, Business Units) to understand data needs and deliver reliable solutions.
- Able to integrate in a French-Indian agile extended team.
Technical & Behavioral Competencies
- Experience of 7-8 years in creating data-warehousing solutions using Dataiku DSS required.
- Strong background in data warehousing concepts, ETL design, and performance tuning.
- Experience building projects, recipes, plugins, APIs and orchestrations in Dataiku.
- Strong SQL programming hands-on experience in creating procedures, functions, writing complex SQL queries. Knowledge on PostreSql / Teradata will be a add-on.
- Programming - Strong hands-on experience in Python, concepts such as OOPs, reusable utilities, exception handling, debugging, and clean code practices and scripting within Dataiku.
- Hands on with GitLab or similar tools to automate builds and deployments.
- Knowledge of Kubernetes, Docker and cloud storage (COS) for scalability.
- Experience with Parquet, ORC and other columnar storage formats.
- Experienced user of ALM/QC, JIRA, Confluence, SharePoint.
- Excellent oral and writing communication skills.
- Experience in working in Agile environment.
- Add On (Nice to have) Apache Airflow experience building DAGs and orchestrating Dataiku pipelines.
- Add On (Nice to have) Data Visualization exposure to tools such as Tableau.
- Add On (Nice to have) Streaming / Search – familiarity with Kafka, Elasticsearch, Kibana.
Specific Qualifications:
- Minimum 6 - 8 years of experience along with Bachelor's Degree - in IT, Computer related field or equivalent experience. Skills Referential (Required knowledge, skills and abilities) Other Skills:
📌 Dataiku Developer (Hyderabad)
🏢 hackajob
📍 Hyderabad