A Senior Data Platform Engineer with 12-14 years of experience to own and build the core data platform for a multi-region CPaa S system supporting multiple digital channels.
Experience & Background
10-15 years of experience in data engineering, platform engineering, or distributed systems
Proven experience building large-scale, streaming-heavy data platforms
Solid hands-on background in real-time analytics and data consistency
Comfortable operating at Principal & Staff Engineer level with full technical ownership
Responsibilities:
Mandatory / Primary Responsibilities
Own the end-to-end data architecture for high-throughput, event-driven CPaa S systems.
Design state-transition and event-sourced data models handling millions of updates per second.
Build and operate streaming data pipelines using Apache Pulsar for ingestion and Apache Flink for stateful processing and real-time aggregation.
Design and maintain a data lake using Apache Hudi to support: Upserts and deletes
Incremental queries
Time-travel and auditability
Build real-time analytics datasets in Click House for dashboards,
tenant metrics, and campaign analysis.
Define and enforce processing semantics, including idempotency, deduplication, ordering, and replay safety.
Own data correctness, completeness, and freshness SLAs across streaming and analytics systems.
Design multi-region data ingestion and aggregation, including replication, failover, and reprocessing.
Lead hands-on development, code reviews, performance tuning, and production troubleshooting.
Secondary Responsibilities
Mentor senior engineers and review architectural designs.
Define data standards, schema evolution practices, and platform guidelines.
Participate in release planning and technical prioritization.
Evaluate new data technologies and patterns through POCs.
Collaborate with product and operations teams on data-driven features.
Skills
Mandatory / Primary Skills
Apache Pulsar (topics, partitions, subscriptions, geo-replication)
Apache Fli