19 Sep
|
NOKIA
|
Bengaluru
Your responsibilities
- Lead L4 support operations for cloud-native observability applications and distributed systems, ensuring high availability, reliability, and operational excellence across production environments.
- Drive resolution of complex production incidents, lead service recovery efforts, and leverage AI-assisted troubleshooting and analytics tools to accelerate diagnosis, impact assessment, and restoration of services.
- Lead deep root cause analysis (RCA) investigations for critical incidents and recurring issues, driving corrective and preventive actions to improve platform stability, resilience, and customer experience.
- Participate in a 24x7 rotational support organization, ensuring SLA compliance, continuous service availability, and effective management of critical incidents and escalations across global customer deployments.
Your skills and experience
- Engineering degree with 6-12 years of experience in production support, incident management, and technical leadership for large-scale cloud-native and distributed systems.
- Hands-on expertise in Kubernetes/OpenShift, microservices, Linux, and containerized application platforms, including troubleshooting complex issues across application, platform, and infrastructure layers.
- Proven experience in Root Cause Analysis (RCA), performance troubleshooting, service recovery, and resolution of critical production incidents, with the ability to act as the highest technical escalation point for customer-critical issues.
- Experience troubleshooting complex networking, databases, Kafka/messaging systems, and distributed application architectures in large-scale production environments.
- Exposure to 4G/5G Core Networks, telecom network functions, and 3GPP standards.
- Proven ability to collaborate with Engineering, Product, Operations, and Customer teams to drive operational improvements, platform reliability, and continuous service enhancement.
📌 4LS Support (Bengaluru)
🏢 NOKIA
📍 Bengaluru