Role & responsibilities-
- Implement and support enterprise observability solutions using Dynatrace, SCOM, and ServiceNow ITOM/Event Management.
- Configure monitoring for infrastructure, applications, databases, cloud, and hybrid environments.
- Improve alert quality through threshold tuning, anomaly detection, dashboarding, and alert optimization.
- Drive alert noise reduction, event correlation, deduplication, and monitoring standardization.
- Manage integrations between Dynatrace, SCOM, and ServiceNow Event Management platforms.
- Develop automation solutions using Python, PowerShell, REST APIs, YAML/JSON, and CI/CD pipelines.
- Support observability-as-code initiatives and automated monitoring onboarding.
- Partner with application, infrastructure, and cloud teams to onboard services into monitoring platforms.
- Apply ITIL practices to improve incident quality, operational response, and service reliability.
Required skills-
- 10+ years of overall IT experience with strong expertise in Observability, Monitoring, Event Management, Infrastructure Operations, Application Support, or Platform Engineering.
- Hands-on experience with Dynatrace including OneAgent, dashboards, alerting, management zones, synthetic monitoring, service flow, problem detection, tagging, and Davis AI.
- Solid experience with Microsoft SCOM, including management packs, alert configuration, rule tuning, agent monitoring, and infrastructure monitoring.
- Working knowledge of ServiceNow ITOM/Event Management, including event ingestion, correlation, deduplication, enrichment, and incident integration.
- Experience integrating monitoring and observability platforms with ServiceNow or other ITSM tools.
- Solid understanding of observability concepts such as metrics, logs, traces, events, topology mapping, synthetic monitoring, and full-stack monitoring.
- Strong automation and scripting skills using Python, PowerShell, REST APIs, Shell Scripting, YAML/JSON.
- Experience with CI/CD pipelines, source control, automation frameworks, and observability-as-code practices.
- Exposure to Azure, AWS, VMware, Windows, Linux, databases, middleware, and hybrid cloud environments.
- Good understanding of ITIL processes, including Incident, Problem, Change, and Event Management.
- Strong analytical and troubleshooting skills with the ability to identify monitoring gaps, alert quality issues, and event-flow problems.
- Ability to work independently while collaborating effectively with application, infrastructure, engineering, and support teams.
Interested in building next-generation observability solutions? Visit our Careers Page and apply for this opportunity- Northern Trust | Careers
If you have relevant experience or know someone who would be a outstanding fit, please share your profile/referral at
[email protected].
📌 Senior Lead | Observability & Dynatrace | Pune
🏢 Northern Trust
📍 Pune