17 Sep
|
Yantran
|
Chennai
Seeking an SRE to improve reliability, scalability, availability, monitoring, and performance of production systems.
Responsibilities:
Monitor production settings.
Develop automation for operational tasks.
Improve system reliability and availability.
Define SLIs, SLOs, and SLAs.
Perform incident management.
Conduct root-cause analysis.
Implement observability.
Optimize infrastructure and applications.
Required Skills:
SRE.
Linux.
Kubernetes.
Cloud platforms.
Python/Go.
Prometheus.
Grafana.
CI/CD.
Infrastructure as Code.
📌 Site Reliability Engineer Chennai
🏢 Yantran
📍 Chennai