16 Sep
|
Yantran
|
Chennai
Seeking an SRE to improve reliability, scalability, availability, monitoring, and performance of production systems.
Responsibilities:
- Monitor production environments.
- Develop automation for operational tasks.
- Improve system reliability and availability.
- Define SLIs, SLOs, and SLAs.
- Perform incident management.
- Conduct root-cause analysis.
- Implement observability.
- Optimize infrastructure and applications.
Required Skills:
- SRE.
- Linux.
- Kubernetes.
- Cloud platforms.
- Python/Go.
- Prometheus.
- Grafana.
- CI/CD.
- Infrastructure as Code.
📌 Site Reliability Engineer (Chennai)
🏢 Yantran
📍 Chennai