DataHub is an AI & Data Context Platform adopted by over 3,000 enterprises, including Apple, CVS Health, Netflix, and Visa. Innovated jointly with a thriving open-source community of 13,000+ members, DataHub's metadata graph provides in-depth context of AI and data assets with best-in-class scalability and extensibility.
The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos.
The company's enterprise SaaS offering, DataHub Cloud, delivers a fully managed solution with AI-powered discovery, observability, and governance capabilities. Organizations rely on DataHub solutions to accelerate time-to-value from their data investments, ensure AI system reliability, and implement unified governance, enabling AI & data to work together and bring order to data chaos.
In this role, you will
Enhance the Python-based ingestion framework to support ingesting usage statistics, lineage, and operational metadata from systems like Snowflake, Redshift, Kafka, & more
Build connectors for major systems in the up-to-date data and ML stacks
Enable the ingestion framework to run in a cloud native setting
Requirements:
5+ years of engineering experience
Expertise in Python
Familiarity with tools in the modern data and ML ecosystem
Knowledge of distributed systems
Ability to design for scale and fault tolerance
Perks and Perks
We invest in people so they can do their best work and enjoy doing it. Our benefits reflect the way we build: practical, thoughtful, and designed to support long-term growth.
Competitive compensation
We offer salaries that reflect your skills, experience, and the impact you make. You bring value—we make sure you're re
📌 Software Engineer Bengaluru (India)
🏢 Datahub
📍 India