04 Oct
|
Religent Systems
|
Hyderabad
04 Oct
Religent Systems
Hyderabad
Key Responsibilities
- Lead enterprise observability strategy across applications, API platforms, infrastructure, and business-critical services.
- Assess observability maturity across logs, metrics, traces, events, dashboards, and alerts.
- Define and establish Observability CoE standards, including:
- Instrumentation guidelines
- Naming conventions
- Tagging and metadata standards
- Dashboard design principles
- Synthetic monitoring
- Alerting and escalation standards
- Identify observability gaps and develop structured improvement roadmaps aligned with business priorities and operational risks.
- Design and implement end-to-end observability for critical business workflows.
- Configure request attributes to enrich traces with business context such as transaction IDs, process identifiers, and service-specific metadata.
- Trace distributed transactions across microservices, APIs, MQ, mainframes, and external systems using Dynatrace PurePath.
- Define and manage Dynatrace Segmentation for multi-team, multi-application, and multi-environment environments.
- Use DQL (Dynatrace Query Language) for metrics analysis, operational analysis, dashboard development, and reporting.
- Design effective alerting strategies using baselines, anomaly detection, SLO-driven thresholds, and severity-based alerting to reduce alert noise.
- Build audience-specific dashboards for executives, operations, engineering, and application teams.
- Leverage Dynatrace Smartscape and Davis AI for topology-aware root-cause analysis and incident investigation.
- Define and manage SLIs, SLOs, and error budgets for critical business services.
- Integrate SLOs and observability practices into release and operational processes.
- Support integration of Dynatrace with external platforms such as Moogsoft, ITSM tools, and event management solutions.
- Collaborate with engineering, operations, infrastructure, application, and business teams to drive observability adoption and continuous improvement.
- Support observability across OpenStack-based infrastructure and cloud environments.
Required Skills
- Strong hands-on experience with Dynatrace Observability Platform.
- Solid knowledge of Dynatrace OneAgent, PurePath, Smartscape, Davis AI, DQL, dashboards, alerting, and monitoring.
- Experience with distributed tracing and application performance monitoring (APM).
- Strong understanding of logs, metrics, traces, events, synthetic monitoring, and infrastructure monitoring.
- Experience defining SLI, SLO, SLA, and error-budget frameworks.
- Experience with observability standards, governance, and Observability CoE initiatives.
- Strong knowledge of microservices, APIs, MQ, distributed systems, and enterprise applications.
- Experience with OpenStack infrastructure and cloud environments.
- Experience with ITSM/event management integrations.
- Strong analytical and troubleshooting skills with experience in root-cause analysis and incident investigation.
- Ability to work with multiple application, infrastructure, engineering, and business teams.
📌 Dynatrace Observability Engineeror Architect OpenStack Infrastructure (Hyderabad)
🏢 Religent Systems
📍 Hyderabad