06 Aug
|
Shell Infotech
|
Hyderabad
06 Aug
Shell Infotech
Hyderabad
Hi Everyone,
Hope youre doing well!
We have an urgent opening for a Data Engineer one of our esteemed clients.
Client Location: Hyderabad
Employment Type: Contract to hire(9 to 12Months)
Experience Required: 5 – 10Years
Education: Bachelor's degree in Computer Science, Engineering, Information Systems, or related field.
:
Role Purpose
Build and maintain the pipelines and data models that bring Workday hiring data and other entity-level function data into the Databricks Lakehouse, producing clean, reliable, query-ready datasets for the analyst team to build dashboards on.
Key Responsibilities
- Design, build, and maintain ETL/ELT pipelines that ingest data from Workday (hiring) and the respective source systems used by other entity-level functions into Databricks, on a defined refresh cadence.
- Build and own the Lakehouse architecture on Databricks (bronze/silver/gold layers using Delta Lake) that supports both hiring reporting and broader entity-level dashboarding.
- Develop reusable, well-documented data models (dimensional models / semantic layer) with transparent, documented business logic so analysts can build dashboards without re-deriving metrics.
- Implement data quality checks (row counts, null checks, reconciliation against source systems) and monitoring/alerting so data issues are caught before they reach a dashboard.
- Optimize pipeline performance (Databricks Workflows or equivalent orchestration) for cost, runtime, and reliability, and manage job scheduling and failure handling.
- Collaborate closely with the analyst team to expose curated datasets that map cleanly to Power BI semantic models or native Databricks dashboards.
- Manage data access, security, and compliance (role-based access control, data masking where needed) in partnership with IT/security,
given the sensitivity of HR and hiring data.
- Maintain technical documentation (pipeline architecture, data lineage, source-to-target mappings, runbooks) so the system is maintainable beyond any one individual.
- Support ad hoc data requests from business users and analysts, and support troubleshooting when reported numbers don't reconcile to source.
Required Skills & Experience
- 5–8 years of experience in data engineering, with hands-on experience building pipelines on Databricks (PySpark/Spark SQL) or a comparable Lakehouse platform.
- Experience integrating with Workday (Workday Report-as-a-Service, Workday APIs/Connectors, or middleware such as Workato/Boomi) is strongly preferred given the hiring data source.
- Strong SQL and Python skills, and experience working with Delta Lake or equivalent table formats in a medallion (bronze/silver/gold) architecture.
- Working knowledge of orchestration tools (Databricks Workflows, Airflow, Azure Data Factory, or similar) and CI/CD practices for data pipelines.
- Solid grounding in data modeling (star schema, slowly changing dimensions) suited to BI/reporting consumption, not just raw data storage.
- Familiarity with the relevant cloud platform (Azure, AWS, or GCP, depending on the Databricks deployment) including storage, networking, and cost-management basics.
- Understanding of data governance and PII handling practices, particularly relevant for HR/employee data.
Preferred Qualifications
- Experience building data pipelines for HR/TA-specific reporting (headcount, hiring funnel, attrition) in a prior role.
- Relevant certifications (Databricks Certified Data Engineer, Azure/AWS data certifications) are a plus.
- Bachelor's degree in Computer Science, Engineering, Information Systems, or related field.
📌 Data Engineer (Hyderabad)
🏢 Shell Infotech
📍 Hyderabad