17 Sep
|
Maneva Consulting
|
Hyderabad
17 Sep
Maneva Consulting
Hyderabad
Greetings from Maneva!
Job Description
Senior SRE / Observability Engineer
Location: Hyderabad / Chennai / Bangalore
Experience: 8 - 15 Years
Notice Period: Immediate to 15 Days
Requirements:
- Provide SRE and Dynatrace subject matter expertise to mature observability across the Core Banking Vault ecosystem, supporting a single pane of glass view across services, dependencies and operational health.
- Work with upstream and downstream neighbours to strengthen end-to-end monitoring, ensuring consistent telemetry ingestion, dashboards, alerts and actionable insight across critical service journeys.
- Expand monitoring capability across Istio, Kafka, synthetic monitoring and cloud cost observability, with a focus on standardisation, automation, incident reduction and operational resilience.
- Establish and embed SLI/SLO practices, including service level indicators, service level objectives, error budgets, burn-rate alerting and evidence-based reliability improvements.
- Support ongoing training and knowledge transfer to build internal capability and improve consistent use of Dynatrace and observability practices across the platform. Essential Skills:-
- Strong experience in Site Reliability Engineering, Observability Engineering, Platform Engineering or DevOps within cloud-native and business-critical environments.
- Hands-on experience with Dynatrace, including dashboarding, alerting, telemetry analysis, synthetic monitoring and configuration management.
- Solid understanding of observability principles across metrics, logs, traces, events, service dependencies and user journeys.
- Experience applying observability configuration using self-service or automated platforms, with good governance and standardisation.
- Strong working knowledge of Kubernetes platforms, preferably GKE, and service mesh technologies such as Istio.
- Experience using OpenTelemetry data, including standardised ingestion, reporting, dashboards and alert generation.
- Strong working knowledge of Kafka observability, including monitoring brokers, topics, partitions, consumers, lag, throughput, errors and resilience indicators.
- Experience defining and implementing SLI/SLO frameworks, including burn-rate alerting, error budgets and reliability dashboards.
Required Skills
- Strong experience refining alerts to reduce false positives and improve actionable alerting for genuine service issues.
- Valuable understanding of incident management, problem management, operational readiness and continuous improvement practices.
- Banking or core banking platform experience.
- Experience working with Thought Machine Vault.
- Experience supporting observability for large-scale distributed systems and regulated production platforms.
- Experience integrating observability data with cloud cost data to support FinOps insight and cost optimisation.
- Experience with GCP cost management, FinOps practices, cloud consumption reporting or cost-to-service mapping.
- Experience designing training material, running enablement sessions and building internal knowledge across engineering teams.
- Knowledge of operational resilience frameworks, service health reporting and evidence-based reliability improvement.
📌 Senior SRE / Observability Engineer (Hyderabad)
🏢 Maneva Consulting
📍 Hyderabad