01 Oct
|
Talentzo Delhi
|
Bengaluru
01 Oct
Talentzo Delhi
Bengaluru
Responsibilities
- Manage and maintain production infrastructure on AWS.
- Handle incidents and prepare Root Cause Analysis (RCA) reports.
- Build and improve CI/CD pipelines and automation.
- Implement infrastructure security and compliance.
- Manage monitoring, dashboards, alerts, and logs.
- Monitor cloud costs and improve resource usage.
- Build disaster recovery, backups, and failover processes.
Preferred candidate profile
- Strong experience with AWS and Terraform.
- Proficiency in Python and Bash; Django is a plus.
- Experience with Prometheus, Grafana, VictoriaMetrics, and ELK/OpenSearch.
- Positive knowledge of databases and RabbitMQ.
- Understanding of SSO, secrets management, and security practices.
- Hands-on experience with Kubernetes and AWS EKS.
- Experience managing Kubernetes clusters across multiple environments.
📌 Site Reliability Engineer II at A FinTech Company (Bengaluru)
🏢 Talentzo Delhi
📍 Bengaluru