1. Build distributed data platforms supporting online, batch, and streaming use cases.
2. Develop core libraries in Python and Golang for interaction with internal data stores.
3. Create scalable data ingestion systems and data lakes.
4. Define and maintain internal SLAs for data infrastructure.
Role Responsibilities:
1. Work with technologies like Kafka, Presto, Pinot, Flink, Mongo, Redis, and Spark.
2. Design and develop scalable core services with good abstractions and architecture.
3. Collaborate in a rapid-paced environment to drive improvements in data platforms.
4. Research and implement emerging data technologies to support business growth.