Unified Data Platform Architect (Bengaluru)

Unified Data Platform Architect (Bengaluru)

19 Sep
|
The Networker
|
Bengaluru

19 Sep

The Networker

Bengaluru

Job Summary

We are looking for a Principal Software Engineer to provide technical leadership for the design and evolution of large-scale data platforms and distributed systems. This is a senior individual contributor role with responsibility for solving complex, ambiguous engineering problems and establishing technical direction across multiple systems and teams. You will architect platforms that process and serve data at significant scale, combining large-scale batch and streaming computation with reliable backend services and cloud-native infrastructure.

The ideal candidate has deep expertise in distributed systems and data infrastructure, with strong hands-on experience in Java, Hadoop, Apache Spark, Apache Flink, Apache Airflow, and Kubernetes . Python experience is also important for data processing, platform automation, and engineering productivity. You will remain technically hands-on while influencing architecture, engineering standards, and long-term platform strategy across the organization.

Responsibilities

- Define the architecture and long-term technical direction for large-scale data platforms and distributed processing systems.
- Lead the design of highly scalable, reliable backend and data infrastructure supporting business-critical workloads.
- Architect and build high-performance backend services and platform components primarily in Java , with Python used where appropriate for data processing, orchestration, automation, and tooling.
- Design and evolve large-scale batch and real-time data processing architectures using Apache Spark and Apache Flink.
- Establish architectural patterns for data ingestion, transformation, computation, orchestration, storage, and serving across the data lifecycle.
- Design reliable workflow and dependency-management capabilities using Apache Airflow and related orchestration technologies.
- Define architecture and operational patterns for running large-scale data and backend workloads on Kubernetes .
- Solve complex distributed-systems challenges involving scalability, state management, fault tolerance, consistency, partitioning, backpressure, resource management, and recovery.
- Drive improvements in platform reliability, performance, observability, developer productivity,



and infrastructure efficiency.
- Identify systemic bottlenecks and lead architectural initiatives that improve throughput, latency, availability, and cost at scale.
- Establish technical standards and reusable platform capabilities that enable multiple engineering and data teams.
- Lead architecture reviews and provide technical guidance for high-impact initiatives spanning multiple systems and organizational boundaries.
- Partner with architects, senior engineers, engineering leaders, product teams, data engineers, and infrastructure teams to translate business requirements into long-term technical strategy.
- Mentor senior engineers and raise the technical bar through design reviews, code reviews, technical guidance, and engineering best practices.
- Evaluate emerging technologies and make strategic build-versus-buy and architectural decisions for the data platform.
- Lead complex migrations and modernization initiatives while maintaining production reliability and minimizing disruption to dependent systems.

Minimum Qualifications

- 10+ years of software engineering experience, including significant experience designing and operating large-scale distributed systems or data platforms.
- Deep expertise in Java and strong software engineering fundamentals.
- Proficiency with Python for data engineering, automation, or platform development.
- Extensive experience designing and building production backend services and distributed systems.
- Deep hands-on experience with Apache Spark and large-scale distributed data processing.
- Strong experience with the Hadoop ecosystem , including technologies such as HDFS, Hive, and YARN.
- Experience designing and operating real-time or stateful streaming systems using Apache Flink .
- Experience designing large-scale workflow orchestration using Apache Airflow or comparable technologies.




- Strong production experience with Kubernetes , containers, and cloud-native application architectures.
- Deep understanding of distributed-systems concepts including partitioning, replication, consistency, fault tolerance, distributed state, scheduling, resource management, and failure recovery.
- Strong understanding of both batch and streaming architectures and the tradeoffs between different processing models.
- Demonstrated experience driving architecture and technical decisions across multiple teams or major platform initiatives.
- Proven ability to operate effectively in ambiguous problem spaces and turn broad business or platform requirements into executable technical strategies.
- Track record of mentoring senior engineers and influencing engineering practices beyond an immediate team.

Preferred Qualifications

- Experience architecting platforms processing petabyte-scale datasets and/or billions of events per day .
- Deep knowledge of Apache Spark and Apache Flink .
- Experience with modern data lake and lakehouse technologies such as Apache Iceberg
- Experience designing multi-tenant data platforms, including workload isolation, resource governance, capacity management, and cost optimization.
- Solid knowledge of data formats, partitioning strategies, schema evolution, metadata management, and data lifecycle management.
- Experience with data governance, lineage, data quality, and platform observability.
- Experience designing highly available control-plane or platform services supporting large numbers of internal users and workloads.

Core Technology Areas

- Languages: Java, Python
- Distributed Processing: Apache Spark, Apache Flink
- Data Platform: Hadoop, HDFS, Hive, YARN, Iceberg

- Streaming: Apache Kafka

- Orchestration: Apache Airflow
- Infrastructure: Kubernetes, Docker
- Architecture: Distributed Systems, Batch & Streaming Processing, Data Lake/Lakehouse, Cloud-Native Data Infrastructu

Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.

📌 Unified Data Platform Architect (Bengaluru)
🏢 The Networker
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: unified data platform architect (bengaluru) / bengaluru