- Must be able to work on 24*7 shift model (Rotational Basis).
- Serve as the first level engagement & support for incidents, escalations for Home Value Stream.
- Understand the problem using alerting and monitoring dashboards and escalate to other teams,
management, vendors or service provider to resolve issue.
- Track Alerts/Incident over Jira and Teams for end to end closure.
- Monitor Application & Infrastructure Alerting
- Proper Handover to different shift members
- Work closely with Incident Manager.
- Monitor Dashboards of Application & Infrastructure, Work with Platform and Engineering team to
make Dashboards more effective.
- Handling alerts and incidents to ensure minimal disruption of the service.
- Proper Handover to different shift members
Required Skills and Experience
- AWS knowledge is must (EC2, IAM, VPC, S3, IAM, Route53, CloudWatch)
- Linux fundamental knowledge is required
- Docker & Kubernetes basics is required
- Monitoring Tools (Dynatrace, Grafana, Prometheus, DataDog, Alert-manager, ELK) knowledge is