Looking For Site Reliability Engineer (Bangalore) (Bengaluru)

Looking For Site Reliability Engineer (Bangalore) (Bengaluru)

24 Sep
|
Tech Mahindra
|
Bengaluru

24 Sep

Tech Mahindra

Bengaluru

Own the reliability and observability of our Kafka and Elasticsearch platforms proactively identifying risks before they become incidents.

Lead the response to production degradations, conducting thorough post-mortems and driving systemic fixes to eliminate repeat failures.

Design and implement Ansible playbooks for automated deployment, configuration management, rolling upgrades, and day-2 operations across Kafka and Elasticsearch clusters.

Build Python tooling to automate health checks, operational workflows, alerting integrations, and pipeline diagnostics — reducing toil and improving team efficiency.

Partner with development teams to optimize producer/consumer configurations, index strategies, and query performance.

Drive capacity planning, cluster scaling, and architecture improvements in collaboration with senior and platform engineers.

Contribute to and continuously improve runbooks, and internal knowledge — raising the bar for how the team operates.

Participate in an on-call rotation with a strong culture of sustainable incident management.

Technical & Behavioral Competencies

1–3 years of hands-on experience operating Apache Kafka and/or Elasticsearch in production environments.

Solid understanding of Kafka internals:



brokers, partitions, offsets, consumer groups, Kafka Connect, and Schema Registry.

Solid understanding of Elasticsearch: clusters, nodes, indices, mappings, shards, ILM, and the broader Elastic Stack.

Robust Linux/Unix fundamentals and confidence diagnosing issues at the infrastructure level.

Experience with observability tooling — Dynatrace, Prometheus, Kibana, etc

Specific Qualifications:

Ansible: Hands-on experience writing and maintaining playbooks, roles, and inventories for infrastructure automation. Familiarity with Ansible Vault for secrets management.

Python: Confident scripting for automation and tooling, with experience using libraries such as kafka-python, confluent-kafka, or the Elasticsearch Python client. Ability to write clean, maintainable scripts suitable for shared operational use.

Experience integrating automation into CI/CD pipelines (Jenkins, Bitbucket, Kubernetes, ) is a plus.

Skills Referential (Required knowledge, skills and abilities)

Technical Skills:

- Apache Kafka
- Elasticsearch
- Ansible
- Python
- CI/CD(Jenkins), Bitbucket and Kubernetes

📌 Looking For Site Reliability Engineer (Bangalore) (Bengaluru)
🏢 Tech Mahindra
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: looking for site reliability engineer (bangalore) (bengaluru) / bengaluru