Shift Timings: Night Shift (5:00 PM 2:00 AM IST) / US EST Primetime
Role Overview
We are looking for a Service Assurance Technical Lead with 9+ years of experience to manage service health monitoring, operational stability, and incident response across multi-cloud applications. You will connect up-to-date DevOps practices with IT service monitoring and lead operational escalations during US hours.
Key Responsibilities
Governance & Standards: Establish service assurance guidelines, incident protocols, and operational stability practices. Mentor team engineers and improve incident lifecycle operations.
Observability & Delivery: Oversee central monitoring systems and multi-project deployments. Help automate container deployments and service mesh routing.
Client Engagement: Keep client stakeholders updated on system health, open incident statuses, and post-incident improvements.
Incident Escalation:
Direct the incident management lifecycle for production outages, coordinating cross-functional teams to resolve issues and finish detailed RCAs.
Required Technical & Skilled Skills
Experience: 9+ years managing IT service operations, network/service monitoring, and cloud-native application support.
Monitoring Systems: Hands-on experience with IBM Tivoli Netcool/Omnibus, Prometheus, Grafana, and Splunk.
CI/CD & GitOps: Direct experience using Jenkins (declarative/pipeline setups), ArgoCD, Git, and GitHub.
Containers & Mesh: Working knowledge of Docker, Kubernetes, Helm charts, and Istio service mesh setups.
IaC & Administration: Practical knowledge of Terraform, Linux administration, and Shell scripting for task automation.
Preferred Certifications
KodeKloud DevOps Engineer Certification
Certified Kubernetes Administrator (CKA)
ITIL Foundation / Service Operation Certification
📌 Service Assurance Technical Lead Hyderabad
🏢 Softility Tech
📍 Hyderabad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.