Role Description
UST is seeking an experienced Site Reliability Engineer (SRE) to support one of the leading financial services organizations in the US and Australia. The ideal candidate will have solid expertise in Google Cloud Platform (GCP), CI/CD, DevOps, Automation, Containerization, and Cloud Operations. The role involves building reliable, scalable, and secure cloud infrastructures while ensuring high availability, performance, and operational excellence across enterprise applications. Key Responsibilities Continuously improve GCP deployment practices to enhance reliability, scalability, security, and operational efficiency. Collaborate with development teams, DevOps engineers, architects, and business stakeholders to deliver robust cloud solutions. Manage the complete service lifecycle, including design, deployment, monitoring, operations, and continuous improvement. Support application and infrastructure readiness through architecture reviews, capacity planning, launch assessments, and operational governance. Monitor and maintain service health, availability,
latency, performance, and reliability. Implement automation solutions to streamline infrastructure provisioning, deployments, and operational activities. Drive reliability engineering practices through proactive monitoring, incident management, root cause analysis, and blameless postmortems. Develop scripts and automation tools for cloud infrastructure management and configuration. Build, maintain, and enhance CI/CD pipelines using tools such as Jenkins. Manage containerized workloads using Docker and Kubernetes. Troubleshoot production issues through log analysis, monitoring tools, and problem-solving techniques. Create and maintain deployment documentation, architecture diagrams, and operational runbooks. Work closely with Quality Engineering, Software Development, and Infrastructure teams to ensure smooth application delivery.
Skills
GCP, Terraform,Docker,Devops
📌 GCP DevOps (Chennai)
🏢 UST
📍 Chennai