ELK StackAdministrator / Engineer (Mumbai)

ELK StackAdministrator / Engineer (Mumbai)

06 Sep
|
Kyndryl India
|
Mumbai

06 Sep

Kyndryl India

Mumbai

Experience: 3–6 years overall IT experience, including 2–4 years of hands-on ELK Stack experience

Role Level: L2 Support / Operations

Technology: Elasticsearch, Logstash, Kibana, Beats

Role Summary: Responsible for day-to-day administration, monitoring, troubleshooting, and operational support of the ELK Stack environment. The engineer will handle incidents and service requests, perform health checks, support configuration and deployment activities, troubleshoot log ingestion and search issues, and coordinate with L3/OEM teams for complex problems.

Area Skill / Responsibility :

Elasticsearch Cluster administration, node/index/shard management, cluster health monitoring, allocation troubleshooting and performance monitoring

Logstash Pipeline configuration, troubleshooting input/filter/output issues, parsing problems and pipeline performance. Integration management.

Kibana Dashboard, visualization and index/data-view administration; troubleshooting access and visualization issues

Beats / Agents Working knowledge of Filebeat, Metricbeat and other Elastic agents

Log Management Troubleshoot missing/delayed logs, ingestion failures, parsing errors and data-quality issues. Incident analysis for new integration, working with application team on KPI.

Monitoring Monitor cluster health, disk utilization, JVM/heap, CPU, memory, shard status, ingestion rates and service availability

Incident Management Perform troubleshooting, RCA inputs, log analysis, service restoration and escalation to OEM. Analyse application incident to identify integration, insight management.

Change Management Execute approved configuration changes, deployments, maintenance activities and standard changes

Index Management Index templates, mappings, aliases, rollover, retention and ILM policies

Security Basic understanding of ELK authentication, authorization, roles, certificates/TLS and access management





Linux Strong Linux administration and troubleshooting skills; process, filesystem, network and resource troubleshooting

Networking Basic understanding of TCP/IP, DNS, ports, connectivity, firewall and load-balancing concepts

Scripting Basic Shell/Python scripting for operational activities and automation

ITILWorking knowledge of Incident, Problem, Change and Service Request Management

Documentation Maintain SOPs, runbooks, troubleshooting documents, shift handover and operational reports

Key Responsibilities

- Perform daily health checks and proactive monitoring of ELK infrastructure.
- Monitor Elasticsearch cluster, nodes, indices, shards, JVM heap, storage and overall performance.
- Troubleshoot yellow/red cluster status, unassigned shards, high resource utilization and node availability issues.
- Troubleshoot Logstash pipelines and log ingestion failures.
- Investigate missing, delayed, duplicate or incorrectly parsed logs.
- Support Filebeat/Metricbeat/Elastic Agent configuration and troubleshooting.
- Perform index management, retention and ILM-related operational activities.
- Support Kibana dashboard, visualization and data-access issues.
- Analyze alerts and take corrective actions as per SOP/runbook.
- Perform service restart/recovery and standard operational activities.
- Execute approved production changes and provide implementation/rollback evidence.
- Participate in incident bridges for ELK-related production issues.




- Perform initial RCA and provide technical inputs for Problem Management.
- Escalate complex product/architecture issues to L3/Elastic OEM with appropriate logs and diagnostics.
- Work with application team on incident analysis for new integration & KPI mapping.
- Maintain operational documentation, SOPs and knowledge-base articles.
- Work in a 24×7 production support setting where required.

Mandatory SkillsCandidates should have hands-on knowledge of Elasticsearch, Logstash, Kibana, Filebeat/Metricbeat, Linux administration, JSON, REST APIs, index/shard concepts, ELK monitoring and production troubleshooting. They should be comfortable troubleshooting issues such as cluster RED/YELLOW state, unassigned shards, disk watermark conditions, JVM/heap utilization, failed Logstash pipelines, parsing errors, log ingestion delays, connectivity issues and Kibana access problems.

Preferred SkillsExposure to Elastic Stack upgrades/patching, Elasticsearch snapshots and recovery, TLS/certificate management, LDAP/AD integration, monitoring tools, ServiceNow, Ansible, Shell/Python automation, VMware/cloud platforms and enterprise security environments would be advantageous.

Experience RequirementMinimum: 5 years of IT application dev/support experience with at least 2 -3 years of hands-on ELK Stack support. Having fair knowledge in Operating System, Security.

Preferred: 6+ years overall experience with 3+ years supporting ELK in a large enterprise/production environment,

preferably handling operations involving incidents, changes, monitoring and troubleshooting.

Educational Qualification

- Bachelor's degree in Computer Science, Information Technology, Engineering or equivalent technical qualification. Relevant Elastic certifications are desirable but not mandatory.

Work Location:Navi Mumbai, 5 days working from Client office.

📌 ELK StackAdministrator / Engineer (Mumbai)
🏢 Kyndryl India
📍 Mumbai

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: elk stackadministrator / engineer (mumbai) / mumbai

Subscribe to this job alert:

Get the latest job offers by email for: elk stackadministrator / engineer (mumbai) / mumbai