Software Engineer II Backend & Data Platform Python Kafka PySpark (Pune)

Software Engineer II Backend & Data Platform Python Kafka PySpark (Pune)

02 Sep
|
HiLabs
|
Pune

02 Sep

HiLabs

Pune

HiLabs builds AI-driven data solutions for US healthcare payers, solving large-scale data quality problems across provider, claims, and clinical datasets. You will join the TMV platform team in Pune.

About the Role

We are hiring a Backend Engineer to build production-grade, scalable services for the TMV platform. The work spans Python backend engineering, distributed data processing with PySpark, event-driven systems on Kafka and Redis, and ML inference integration all within a secure AWS environment. You will work directly with Data Science, Data Engineering, DevOps and Product teams.

What You Will Do

- Design, build and maintain Python backend services and microservices using FastAPI/Flask.
- Build scalable PySpark pipelines for ingestion, classification, linkage, feature engineering and schema generation.
- Develop event-driven and asynchronous services using Kafka and Redis Streams producers, consumers, consumer groups, topics, partitioning, offset management and retry handling.
- Use Redis for caching, quick lookups, TTL-based state and distributed coordination.
- Build and deploy containerized services on AWS EKS / Kubernetes.
- Integrate with AWS S3, EMR, EventBridge, PostgreSQL, Snowflake, Kafka and Redis.
- Deploy batch scoring and ML inference services using versioned model artifacts.
- Implement reliability patterns: idempotency, retries, timeouts, dead-letter queues and failure recovery.
- Design efficient database schemas, indexes,



queries and data-access layers.
- Implement logging, metrics, monitoring and alerting; write unit and integration tests; participate in code reviews.
- Debug issues across application, data, messaging and infrastructure layers.

Must Have (3+ years hands-on)

- Strong Python — OOP, data structures, design principles
- PySpark / Apache Spark, distributed data processing
- Apache Kafka — producers, consumers, consumer groups, partitions, offsets
- Redis — caching, TTL, Pub/Sub or Streams
- FastAPI or Flask, REST APIs, microservices
- AWS — S3, IAM, CloudWatch, and EKS or EMR
- Docker; working knowledge of Kubernetes
- PostgreSQL / SQL — schema design, indexing, query optimization
- Git, CI/CD, unit and integration testing

Good to Have

- Amazon MSK or Kafka on Kubernetes; Redis Cluster or ElastiCache
- AWS EMR and Spark at scale; Snowflake
- Parquet and schema management
- ML inference, model serving, embeddings, risk-scoring pipelines, MLflow
- GPU-based workloads
- Healthcare / claims / clinical data; PHI and HIPAA awareness

Ideal Candidate A strong backend and distributed-systems engineer comfortable across application, data and infrastructure layers, with real ownership of production services — debugging depth, system-design fundamentals, and the instinct to build fault-tolerant, observable systems. Location: Pune (work from office). Candidates open to relocating to Pune may apply.

📌 Software Engineer II Backend & Data Platform Python Kafka PySpark (Pune)
🏢 HiLabs
📍 Pune

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: software engineer ii backend & data platform python kafka pyspark (pune) / pune

Subscribe to this job alert:

Get the latest job offers by email for: software engineer ii backend & data platform python kafka pyspark (pune) / pune