06 Aug
|
Tranzeal
|
Chennai
Senior Observability/AIOps Automation Engineer or Architect The client is looking for a Senior Observability/AIOps Engineer (10+ years) who can design and automate enterprise monitoring using Current Relic, LogicMonitor, Terraform, and AIOps, replacing manual alert configuration with intelligent, automated observability solutions.
- Experience: 10+ years
- Location: Pune / PAN India (Hybrid)
- Duration: 6+ months
- Role Level: Senior Engineer / Lead / Architect
JD: We have below requirement in Observability AIOps. Customer's primary need is automation in the Observability space. they are looking for someone who can lead automation of observability configuration and help eliminate manual alert configuration and automate alert aggregation instead of completing manual tasks.
We're looking for a candidate with ~ 10 + years of Experience and a strong background in Observability AIOps. He should have a proven track record of managing and optimizing observability tools, a deep understanding of standardization & automation, and the ability to work effectively across different teams to drive business value & outcome.
Skills -
Observability Platform Management: Provide expert operational support for a range of enterprise monitoring tools, including New Relic, Logic Monitor with a specific emphasis on Digital Employee Experience (DEX) platforms like Nexthink. Drive new features and capabilities by leading Proofs of Concept (POCs)
Observability-as-Code & Automation:
Champion the 'Observability-as-Code' paradigm using terraform / equivalent by integrating monitoring configuration directly into CI/CD pipelines & automate the action
AIOps for Proactive Insights: Utilize AIOps to integrate machine learning and AI into our monitoring systems, automating the analysis of data to predict issues, automate incident detection and event correlation with focus how to reduce MTTR, increase SLA & shift mindset of observability
Incident & Alert Management: Lead incident management by creating, refining, and automating monitoring alerts to ensure proactive issues detection and minimize downtime.
Proactive Problem-Solving: Use Observability platform to proactively identify and resolve employee-impacting issues like slow logins, application crashes, and network latency before they escalate.
Strategic Collaboration & Enablement: Partner with business and development teams to understand their requirements and define comprehensive monitoring .Act as a key collaborator by providing compelling demos and presentations to internal teams, showcasing the value and insights gained from our observability & AIops.
Data-Driven Insights: Work with various teams to deliver best-in-class tools for visualizing Real User Monitoring (RUM), Synthetics, log data and performance data.
Analytics & Reporting: Develop comprehensive health and performance reports, create AIOps rules, design custom dashboards, and create business values out of Observability.
📌 Observability Lead (Chennai)
🏢 Tranzeal
📍 Chennai