09 Sep
|
F5 Networks
|
Bengaluru
09 Sep
F5 Networks
Bengaluru
- Be the Force Behind Observability & Stability
- Drive end-to-end Observability (Logs, Metrics, and Alerts) across our hybrid SaaS stack, spanning cloud, edge, and physical network devices.
- Take ownership of Alerting strategy, cutting through noise while ensuring actionable, high-fidelity alerts.
- Implement intelligent automation to reduce operational toil and enhance real-time visibility.
- Own & Automate Operations
- Design, build, and manage automation for self-healing infrastructure across cloud + global PoPs.
- Develop automation for Kubernetes, ArgoCD, Helm Charts, Golang-based services, AWS, GCP, Terraform.
- Improve networking observability, ensuring our routers, switches, and firewalls are monitored at scale.
- Continuously eliminate manual ops work through automation and platform improvements.
- Lead Incident Response & Operational Excellence
- Participate in on-call rotations, ensuring rapid incident response across our cloud + edge stack.
- Drive incident response automation, reducing MTTR and increasing system resilience.
- Ensure security, compliance, and best practices in observability & automation.
- Collaborate & Mentor
- Work closely with application teams, network engineers, and SREs to improve reliability and performance.
- Mentor junior engineers, fostering a culture of automation-first thinking and deep observability.
Skills: Kubernetes, Terraform, Incident Response, Automation
Experience: 7.00-11.00 Years
📌 Sr. Principal Site Reliability Engineer (Bengaluru)
🏢 F5 Networks
📍 Bengaluru