Noc Engineer (Chennai)

Noc Engineer (Chennai)

29 Sep
|
Tanla Platforms
|
Chennai

29 Sep

Tanla Platforms

Chennai

Network Operations Center(NOC) Infrastructure Support

The NOC Team shall operate on a 247 basis to monitor, support, troubleshoot, and maintain all Karix Infrastructure services, including Application, Database, Network Monitoring, Top Customer Monitoring, and Telco Monitoring services used by customers.

Technical Skills & Experience

- Strong hands-on experience in Linux/Unix administration and troubleshooting is mandatory.
- Good knowledge of Linux server monitoring, health checks, performance monitoring, capacity utilization, service availability, log monitoring, and issue troubleshooting in production environments.
- Experience with Linux infrastructure monitoring tools such as Checkmk, Nagios, Zabbix, SolarWinds, Prometheus, Grafana, or equivalent monitoring platforms.
- Experience in installation, configuration, maintenance, administration, and troubleshooting of Infrastructure and Application Monitoring tools such as ServiceNow ITOM, Dynatrace, Checkmk, Nagios, SolarWinds, Prometheus, Grafana, and equivalent tools.
- Working knowledge of holistic application and infrastructure monitoring, with the ability to drill down across Network, Linux/Infrastructure, Middleware, Application, and Database layers.
- Should have implemented 23 projects using commercially available monitoring/observability tools such as Dynatrace, SolarWinds, ServiceNow, Checkmk, ManageEngine, or equivalent platforms.
- Should be able to design and implement monitoring solutions, configure monitoring parameters, and create smart/intelligent alerts to proactively identify failures, anomalies, and performance issues.




- Robust working knowledge of SQL and databases, including SQL queries, data validation, database health checks, monitoring, and troubleshooting of database-related incidents.
- 2+ years of experience in scripting/programming, including Python, Shell Scripting, and PowerShell, for automation, monitoring, troubleshooting, and operational activities.
- Good understanding of modern software development methodologies and Object-Oriented Programming (OOP) concepts.
- Hands-on experience with Open Telemetry and modern observability practices.
- Ability to correlate Linux, application, database, network metrics, logs, and traces to identify the root cause of incidents.
- Good understanding of monitoring architecture, alert management, event correlation, threshold tuning, dashboards, and proactive monitoring.

NOC Responsibilities

1. Monitor infrastructure, applications, databases, networks, and alerts through monitoring and observability tools.

2. Monitor Linux infrastructure for availability, performance, capacity, utilization, service health, and operational issues.

3. Configure and maintain monitoring platforms such as Checkmk, Nagios, SolarWinds, Dynatrace, Prometheus, Grafana, and ServiceNow ITOM.

4.



Perform initial incident triage, troubleshooting, impact assessment, and escalation.

5. Open and manage bridge calls for all Critical/Major incidents.

6. Notify relevant stakeholders for Critical/Major issues, including Support Teams, L2, IT, Sales, BizOps, DM teams, and other relevant teams.

7. Perform and coordinate Root Cause Analysis (RCA) for platform and service-related incidents.

8. Perform SQL-based validation and troubleshooting for database-related incidents.

9. Monitor the health, performance, capacity, and utilization of platforms.

10. Publish platform-wise health dashboards and operational reports.

11. Monitor and publish customer KPIs and service-performance metrics.

12. Perform alert optimization, threshold tuning, event correlation, and monitoring-tool configuration.

13. Develop and maintain Python, Shell, and PowerShell scripts for automation and operational efficiency.

14. Support implementation and enhancement of OpenTelemetry-based observability.

15. Identify opportunities for continuous improvement, automation, and proactive monitoring.

16. Ensure ticket/ticker closure within defined SLAs.

17. Build and maintain SOPs, runbooks, troubleshooting guides, and operational documentation.

18. Maintain accurate records and log all operational activities, troubleshooting steps, and incident updates.

19. Proactively identify recurring issues and recommend improvements to enhance platform availability, reliability, performance, and customer experience.

📌 Noc Engineer (Chennai)
🏢 Tanla Platforms
📍 Chennai

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: noc engineer (chennai) / chennai

Subscribe to this job alert:

Get the latest job offers by email for: noc engineer (chennai) / chennai