Responsibilities
- Operate and maintain Elasticsearch, MongoDB, and ClickHouse platforms.
- Manage Elasticsearch indices, shards, replicas, allocation, lifecycle policies, upgrades, and performance.
- Operate MongoDB replica sets and sharded clusters, including balancing, backups, restores, and recovery.
- Operate ClickHouse replication, Keeper, distributed tables, storage, ingestion, and query workloads.
- Troubleshoot availability, replication, storage, query performance, and client-integration issues.
- Perform capacity planning based on data growth, ingestion, retention, replication, and query patterns.
- Plan and validate upgrades, failover, backup, recovery, and failure scenarios.
- Lead production incident triage, root-cause analysis, and performance tuning.
- Review data-platform architectures and build operational automation and observability.
- Participate in a periodic on-call rotation to maintain 24/7 reliability and performance of production services.
Role Overview
We are seeking a hands-on Senior Platform Software Engineer with experience operating stateful data, search, and analytics platforms at scale. As a key member of the Platform Services team, you will help ensure the availability, reliability, performance, and scalability of Elasticsearch, MongoDB, and ClickHouse.
This role is based remotely in Pune. Candidates for this position are required to reside within the Pune metropolitan area. Relocation support is not available at this time.
Qualifications
Minimum Qualifications
- 5 years of software, systems, platform, database, DevOps, or SRE experience.
- 3 years of experience operating stateful distributed systems in production.
- Database Platforms: Hands-on operational experience with at least two of the following platforms in production: Elasticsearch, MongoDB, or ClickHouse.
- Incidents Troubleshooting: Experience leading production incident triage, root-cause analysis (RCA), and performance tuning for database availability, replication, and storage issues.
- Resiliency Capacity: Experience executing cluster upgrades, failovers, backups, disaster recovery, and capacity forecasting for data-intensive platforms.
- Software Automation: Experience writing and maintaining infrastructure automation or backend tooling in Java, Go, or Python.
- Technical Collaboration Review: Experience conducting formal architecture reviews, documenting system design proposals (e.g., RFCs/ADRs), and partnering across engineering teams to standardize infrastructure patterns.
Preferred Qualifications
- Experience with the third platform among Elasticsearch, MongoDB, and ClickHouse.
- Experience operating stateful platforms on Kubernetes or cloud infrastructure.
- Experience designing and validating disaster-recovery and failure scenarios.
- Experience with platform automation, observability, GitOps, or infrastructure as code.
- Experience supporting large-scale, business-critical data platforms.
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Senior Platform Software Engineer (Pune)
🏢 Medallia
📍 Pune