17 Sep
|
Tata Consultancy Services
|
Bengaluru
17 Sep
Tata Consultancy Services
Bengaluru
Role : GCP Site Reliability Engineer (SRE )
Exp range: 6-10 years
Locations: Chennai, Bengaluru, Hyderabad, Pune, Kochi, Kolkata, NCR, Delhi
Virtual Interview: 19th Sep
Responsibilities
We are seeking a skilled GCP Site Reliability Engineer (SRE) with strong hands-on expertise in Google Cloud Platform, Kubernetes, Terraform, and Java Spring Boot applications. The candidate will ensure the reliability, scalability, automation, security, and operational excellence of cloud-native platforms and business-critical applications running on GCP.
Key Responsibilities
- Manage and support cloud infrastructure hosted on Google Cloud Platform (GCP).
- Design, implement, and maintain Kubernetes-based platform solutions, preferably using Google Kubernetes Engine (GKE).
- Develop and manage Infrastructure as Code using Terraform.
- Monitor, troubleshoot, and optimize Java Spring Boot applications running in cloud environments.
- Implement reliability engineering practices, including SLIs, SLOs, and error budget management.
- Automate deployment, monitoring, infrastructure provisioning, and operational activities.
- Support production incidents, perform root cause analysis, and drive corrective and preventive actions.
- Collaborate with development, support, security, and DevOps teams to improve platform stability.
- Build and maintain CI/CD pipelines for application and infrastructure deployments.
- Ensure security, compliance, scalability, resilience,
and high availability of cloud platforms.
Skill Requirements
- 7-15 years of experience in Site Reliability Engineering, DevOps, Cloud Operations, or Production Support.
- Strong hands-on experience with Google Cloud Platform services.
- Extensive practical experience with Kubernetes and containerized deployments, with GKE preferred.
- Strong expertise in Terraform modules, remote state management, reusable infrastructure patterns, and automation.
- Experience supporting or developing applications using Java and Spring Boot.
- Knowledge of Linux administration, networking, shell scripting, and automation tools.
- Experience with monitoring and observability tools such as Google Cloud Operations Suite, Prometheus, Grafana, Current Relic, or Datadog.
- Strong understanding of cloud networking, load balancing, DNS, IAM, secrets management, and security controls.
- Experience in incident management, problem management, production support, and root cause analysis.
- Excellent troubleshooting, communication, documentation, and stakeholder management skills.
Must Have Skills
- Google Cloud Platform (GCP)
- Kubernetes / Google Kubernetes Engine (GKE)
- Terraform
- Java Spring Boot Application Support
- CI/CD Pipelines
- Docker and Containerization
Note: If you are comfortable with the above JD, kindly apply for the same. I will connect with you to collect the required details and schedule the interview.
📌 GCP Site Reliability Engineer (SRE)_TCS Virtual Interview_19th Sep (Bengaluru)
🏢 Tata Consultancy Services
📍 Bengaluru