e are looking for a Site Reliability Engineer to ensure the reliability, scalability, and performance of critical applications in 24x7 production settings. The role involves proactive monitoring, incident management, and automation of operational tasks using tools like Splunk, AppDynamics, Datadog, and Grafana. Hands-on expertise in cloud platforms (GCP, AWS, Azure, PCF), CI/CD pipelines, Linux administration, and scripting (Python, Go, Java, Shell) is essential. Robust knowledge of ITIL processes, DevOps practices, and debugging distributed systems with Java/PostgreSQL is required. The candidate should excel in stakeholder communication, team collaboration, and driving service improvements through automation and process initiatives.
📌 Production Support Hyderabad
🏢 Cloudxtreme
📍 Hyderabad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.