▪ Designs, implements and operates the industrial data lakehouse — Spark, Iceberg, MinIO, InfluxDB, PostgreSQL ▪ Builds end-to-end data pipelines: ingestion, quality, validation and governance
▪ Implements and maintains data gateways for machine and sensor connectivity
▪ Delivers clean, reliable data flows for advanced analytics; optimizes data infrastructure with DevOps and application teams
Responsibilities
- Designs and runs the lakehouse + end-to-end data pipelines.
- Owns ingestion, data quality, validation and governance.
- Stands up and maintains data gateways for machines and sensors.
- Feeds clean,
reliable data to the analytics and dashboard layers.
- Works with DevOps to optimize the data infrastructure.