06 Sep
|
Allied Globetech
|
Mumbai
06 Sep
Allied Globetech
Mumbai
Experience: 6–10+ Years
Location: [Mumbai / Navi Mumbai / Hyderabad / Pan India]
Employment Type: Full-Time
Work Mode: [On-site / Hybrid / Remote]
About the Role
We are looking for an experienced Monitoring & Observability Engineer to join our team and take ownership of enterprise-scale monitoring, application performance management, log analytics, and observability solutions.
The ideal candidate will have strong hands-on experience with the ELK Stack, Kibana/Grafana, AppDynamics, and enterprise monitoring environments. You will be responsible for developing dashboards, analysing application and infrastructure performance, troubleshooting production issues, and integrating monitoring solutions with enterprise applications.
This role requires someone who can go beyond monitoring alerts — identify patterns, analyse root causes, drive performance improvements, and build meaningful observability solutions.
Key Responsibilities
- Design, develop, and maintain monitoring and observability solutions for enterprise applications and infrastructure.
- Administer and manage Elasticsearch, Logstash, and Kibana (ELK) environments.
- Develop and maintain Kibana and Grafana dashboards, visualizations, alerts, and monitoring reports.
- Perform log analysis, application monitoring, and performance troubleshooting across production environments.
- Administer and monitor applications using AppDynamics, including application performance monitoring (APM), health monitoring, alerts, and troubleshooting.
- Monitor application, server, API, database, and infrastructure performance.
- Analyse application logs, metrics, traces, and performance indicators to identify anomalies and potential issues.
- Troubleshoot production incidents and perform Root Cause Analysis (RCA) and performance analysis.
- Define and implement meaningful KPIs, SLIs, SLOs, alerts, and monitoring thresholds.
- Integrate monitoring and observability tools with enterprise applications and IT infrastructure.
- Collaborate with Application,
Infrastructure, DevOps, Cloud, API, and Security teams to improve system reliability and visibility.
- Identify opportunities to automate monitoring, alerting, reporting, and operational workflows.
- Participate in production support, incident management, and performance optimization activities.
- Maintain monitoring documentation, dashboards, alert configurations, and operational procedures.
Required Technical Skills
ELK Stack
- Strong hands-on experience with:
- Elasticsearch
- Logstash
- Kibana
- Experience with log ingestion, parsing, filtering, indexing, searching, and visualization.
- Knowledge of Elasticsearch indices, mappings, queries, clusters, and performance considerations.
Dashboard & Visualization
- Strong experience developing Kibana dashboards and visualizations.
- Experience with Grafana dashboard development, data sources, panels, alerts, and monitoring views.
- Ability to translate operational requirements into meaningful dashboards and actionable insights.
AppDynamics
- Hands-on experience with AppDynamics administration and monitoring.
- Experience with application performance monitoring, business transactions, health rules, alerts, and performance analysis.
- Ability to troubleshoot application performance issues using AppDynamics metrics and diagnostics.
Monitoring & Observability
- Strong understanding of Application Performance Monitoring (APM) and infrastructure monitoring.
- Experience with log, metric, and event-based monitoring.
- Understanding of application and infrastructure observability concepts.
- Strong troubleshooting, analytical,
and problem-solving skills.
- Experience working with enterprise-scale production environments.
Good to Have
- Experience with Cloud Environments (AWS / Azure / GCP).
- Knowledge of APIs, microservices, containers, Kubernetes, and distributed applications.
- Exposure to DevOps and CI/CD environments.
- Understanding of networking, databases, operating systems, and application architecture.
- Experience integrating monitoring platforms with enterprise applications and ITSM tools.
- Exposure to other observability platforms such as Prometheus, OpenTelemetry, Splunk, Dynatrace, or New Relic.
Experience Requirements
- 10+ years of relevant experience in monitoring, application performance management, log analytics, or observability solutions.
- Candidates with 6–8 years of robust relevant experience may also be considered based on technical expertise and hands-on experience.
- Proven experience supporting enterprise-scale production environments.
- Strong experience in incident troubleshooting, performance analysis, and root-cause identification.
What We're Looking For
We are looking for a technically strong professional who can operate effectively across monitoring, observability, application performance, and production support.
The right candidate should be comfortable working with large-scale environments, investigating complex production issues, and converting monitoring data into actionable insights.
If you have a strong observability mindset and enjoy solving performance and production challenges, we'd like to hear from you.
.
.
.
.
.
.
.
.
.
.
.
.
Key Skills / Keywords
ELK | Elasticsearch | Logstash | Kibana | Grafana | AppDynamics | APM | Application Monitoring | Infrastructure Monitoring | Observability | Log Analysis | Performance Monitoring | Dashboard Development | Production Support | Troubleshooting | Root Cause Analysis | Cloud Monitoring | Enterprise Monitoring
📌 Senior Monitoring & Observability Engineer (Mumbai)
🏢 Allied Globetech
📍 Mumbai