- experience in Production Operations, SRE, Observability, Application Support, or Operations Engineering roles.
- Experience with observability, logging, alert management, alert correlation, and alert quality improvement practices using Splunk ITSI, Splunk Observability Cloud & Splunk Enterprise logging.
- Experience implementing and supporting SLI, SLO, Error Budget tracking, service health monitoring, reliability reporting, and continuous improvement initiatives.
- Experience with Incident Management, Problem Management, Troubleshooting, Change Management, Root Cause Analysis (RCA), and Blameless Postmortem practices.
- Experience supporting business applications running on VM and container platforms.
- Experience with cloud platforms such as Azure, AWS, and GCP.
- Experience with automation technologies to reduce the operational toil (Ansible, Python, RPA and other automation platforms)
- Experience with ITSM processes and tools such as ServiceNow.
- Robust analytical, troubleshooting, communication, and collaboration skills."
Please fill the below form to apply for this job.
SRE -Drive- 5th Sep 26 – Fill out form
📌 Site Reliability Engineer - Chennai-Virtual drive -5th Sep-Saturday
🏢 HCLTech
📍 Chennai
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.