05 Aug
|
HCL Technologies
|
Noida
05 Aug
HCL Technologies
Noida
Team Lead - Monitoring Tools, Event Monitoring
Experience: Not Available to Not Available years
Location: Gautam Buddha Nagar, India
Skills: Prometheus, SNMP, GCP, Kubernetes, REST, Python, Go, Linux, Docker
Job Summary
Role Requirements:
• Experience with Network Monitoring and Observability solutions of CNFs/VNFs
• Have good knowledge of Prometheus, SNMP monitoring platforms, GCP, and K8s.
• Should be able to assess the integration of these tools into alerting/on-call systems.
• Experience in Integration/Automation
• REST based integrations for monitoring/alerting tools with external systems.
• Should be able to visualize the end to end workflow (alerting to ticketing).
• Should have experience in automation (Python/Go) for integrations and RESTful API requests for custom jobs.
• Experience in Incident Management Workflow
• Defining escalation policies.
• Ensuring alerts/incident prioritization/closures.
• GAP analysis on the monitoring infrastructure and tools
• Missing functionalities in the current implementation.
• Evaluation of potential replacement tools.
• Should be able to drive a POC on the potential replacement tool.
• DevOps experience
• Linux, Docker (CREs), networking.
• Documentation and Communications
• Should be able to document and maintain the configurations, alert mapping, WOW, error handling.
• Should be able to communicate and convince the non technical stakeholders.
Key Responsibilities
Role Requirements:
• Experience with Network Monitoring and Observability solutions of CNFs/VNFs
• Have valuable knowledge of Prometheus, SNMP monitoring platforms, GCP, and K8s.
• Should be able to assess the integration of these tools into alerting/on-call systems.
• Experience in Integration/Automation
• REST based integrations for monitoring/alerting tools with external systems.
• Should be able to visualize the end to end workflow (alerting to ticketing).
• Should have experience in automation (Python/Go) for integrations and RESTful API requests for custom jobs.
• Experience in Incident Management Workflow
• Defining escalation policies.
• Ensuring alerts/incident prioritization/closures.
• GAP analysis on the monitoring infrastructure and tools
• Missing functionalities in the current implementation.
• Evaluation of potential replacement tools.
• Should be able to drive a POC on the potential replacement tool.
• DevOps experience
• Linux, Docker (CREs), networking.
• Documentation and Communications
• Should be able to document and maintain the configurations, alert mapping, WOW, error handling.
• Should be able to communicate and convince the non technical stakeholders.
Skill Requirements
Role Requirements:
• Experience with Network Monitoring and Observability solutions of CNFs/VNFs
• Have good knowledge of Prometheus, SNMP monitoring platforms, GCP, and K8s.
• Should be able to assess the integration of these tools into alerting/on-call systems.
• Experience in Integration/Automation
• REST based integrations for monitoring/alerting tools with external systems.
• Should be able to visualize the end to end workflow (alerting to ticketing).
• Should have experience in automation (Python/Go) for integrations and RESTful API requests for custom jobs.
• Experience in Incident Management Workflow
• Defining escalation policies.
• Ensuring alerts/incident prioritization/closures.
• GAP analysis on the monitoring infrastructure and tools
• Missing functionalities in the current implementation.
• Evaluation of potential replacement tools.
• Should be able to drive a POC on the potential replacement tool.
• DevOps experience
• Linux, Docker (CREs), networking.
• Documentation and Communications
• Should be able to document and maintain the configurations, alert mapping, WOW, error handling.
• Should be able to communicate and convince the non technical stakeholders.
Other Requirements
Role Requirements:
• Experience with Network Monitoring and Observability solutions of CNFs/VNFs
• Have good knowledge of Prometheus, SNMP monitoring platforms, GCP, and K8s.
• Should be able to assess the integration of these tools into alerting/on-call systems.
• Experience in Integration/Automation
• REST based integrations for monitoring/alerting tools with external systems.
• Should be able to visualize the end to end workflow (alerting to ticketing).
• Should have experience in automation (Python/Go) for integrations and RESTful API requests for custom jobs.
• Experience in Incident Management Workflow
• Defining escalation policies.
• Ensuring alerts/incident prioritization/closures.
• GAP analysis on the monitoring infrastructure and tools
• Missing functionalities in the current implementation.
• Evaluation of potential replacement tools.
• Should be able to drive a POC on the potential replacement tool.
• DevOps experience
• Linux, Docker (CREs), networking.
• Documentation and Communications
• Should be able to document and maintain the configurations, alert mapping, WOW, error handling.
• Should be able to communicate and convince the non technical stakeholders.
📌 Team Lead Monitoring Tools, Event Monitoring (Noida)
🏢 HCL Technologies
📍 Noida