10 Sep
|
GSPANN
|
Hyderabad
Role & responsibilities
- Administer, configure, govern, and optimize the Dynatrace SaaS platform across enterprise applications, infrastructure, cloud, and containerized environments.
- Deploy, upgrade, configure, and troubleshoot Dynatrace OneAgents and ActiveGates to ensure comprehensive monitoring coverage.
- Configure and manage Alerting Profiles, Problem Notifications, Metric Events, Anomaly Detection, and Davis AI-based alerting capabilities.
- Design, implement, and maintain Browser and HTTP Synthetic Monitoring solutions to proactively identify application performance issues.
- Support and enhance Real User Monitoring (RUM) capabilities to provide visibility into end-user application experience.
- Create and manage Management Zones, tagging strategies, auto-tagging policies, naming conventions, and monitoring governance standards.
- Develop and maintain Dynatrace Dashboards, Notebooks, Workflows, and advanced DQL-based analytics using Grail.
- Onboard new applications, APIs, cloud resources, infrastructure components, and business services into the Dynatrace monitoring ecosystem.
- Configure service detection, process grouping, dependency mapping, and automated monitoring policies.
- Manage and optimize monitoring integrations across Azure, AWS, and hybrid cloud environments.
- Implement and support log ingestion, log monitoring, custom metrics, business event monitoring, and enterprise observability solutions.
- Analyze application topology, service flow, distributed tracing, dependencies, and PurePath diagnostics to identify performance bottlenecks.
- Utilize Data Explorer, Grail, and DQL for troubleshooting, performance optimization, capacity planning, and operational insights.
- Perform platform health assessments and drive implementation of enterprise monitoring and observability best practices.
- Support the adoption of modern Dynatrace capabilities including Grail, DQL, Notebooks, Workflows, and next-generation Dashboard frameworks.
- Validate monitoring coverage and observability readiness for new releases, deployments, infrastructure changes, and cloud migrations.
- Configure maintenance windows and monitoring policies to support planned outages and maintenance activities.
- Manage user access, RBAC roles, permissions, licensing, compliance requirements, and platform security standards.
- Integrate Dynatrace with ServiceNow, Microsoft Teams, and other enterprise ITSM, collaboration, and notification platforms.
- Support incident management, service restoration, Major Incident processes, Root Cause Analysis (RCA), and continuous reliability improvement initiatives.
- Define, implement, and monitor SLIs, SLOs, and SLAs to improve platform reliability and operational excellence.
- Develop automation solutions and collaborate closely with Application, Infrastructure, Cloud, Security, and DevOps teams to improve operational efficiency.
- • Drive observability, monitoring maturity, and Site Reliability Engineering (SRE) initiatives across the organization.
Required Skills
- 5+ years of experience in Dynatrace Administration, Observability Engineering, Application Performance Monitoring (APM), or Site Reliability Engineering (SRE).
- Hands-on expertise with Dynatrace SaaS Platform Administration.
- Strong experience with Dynatrace OneAgent and ActiveGate deployment, configuration, troubleshooting, and maintenance.
- Expertise in Synthetic Monitoring, Real User Monitoring (RUM), Application Monitoring, Infrastructure Monitoring, and Cloud Monitoring.
- Strong knowledge of Dynatrace Grail, DQL (Dynatrace Query Language), Data Explorer, Dashboards, Notebooks, and Workflows.
- Experience configuring Alerting Profiles, Problem Detection, Metric Events, Anomaly Detection, and Davis AI.
- Knowledge of Management Zones, Tagging Strategies, Naming Rules, Auto-Tagging Policies, and Monitoring Governance.
- Hands-on experience with Distributed Tracing, Service Flow Analysis, Application Topology Mapping, and PurePath Diagnostics.
- Experience integrating Dynatrace with cloud platforms such as Azure, AWS, and hybrid cloud environments.
- Strong understanding of Log Monitoring, Log Ingestion, Custom Metrics, Business Events, and Enterprise Observability.
- Experience integrating Dynatrace with ServiceNow, Microsoft Teams, and other enterprise monitoring ecosystems.
- Strong understanding of Site Reliability Engineering (SRE), Incident Management, Problem Management, RCA, SLIs, SLOs, and SLAs.
- Experience with automation, scripting, and DevOps practices to improve monitoring efficiency and platform operations.
- Excellent troubleshooting, analytical, and problem-solving skills.
- Strong communication and stakeholder management skills with the ability to work across cross-functional teams.
- Experience supporting large-scale, mission-critical enterprise applications and infrastructure environments.
- Knowledge of ITIL processes and enterprise monitoring best practices is preferred.
- Ability to work in a fast-paced production setting and support critical business applications.
📌 Dynatrace Administrator (Hyderabad)
🏢 GSPANN
📍 Hyderabad