04 Oct
|
Cloudxtreme
|
Hyderabad
04 Oct
Cloudxtreme
Hyderabad
Role & responsibilities
e are looking for a Site Reliability Engineer to ensure the reliability, scalability, and performance of critical applications in 24x7 production environments. The role involves proactive monitoring, incident management, and automation of operational tasks using tools like Splunk, AppDynamics, Datadog, and Grafana. Hands-on expertise in cloud platforms (GCP, AWS, Azure, PCF), CI/CD pipelines, Linux administration, and scripting (Python, Go, Java, Shell) is essential. Robust knowledge of ITIL processes, DevOps practices, and debugging distributed systems with Java/PostgreSQL is required. The candidate should excel in stakeholder communication, team collaboration, and driving service improvements through automation and process initiatives.
📌 Production support (Hyderabad)
🏢 Cloudxtreme
📍 Hyderabad