Position: Kubernetes Site Reliability Engineer (SRE)
Location: Bangalore (Work From Office)- Hybrid
About the Role
We are looking for a highly skilled Kubernetes SRE with strong expertise in Kubernetes administration, automation, and software development using Python or Golang. The ideal candidate will play a key role in building and supporting scalable, reliable, and secure cloud-native platforms. Experience with OpenShift is preferred but not mandatory.
Key Responsibilities
- Manage, maintain, and optimize Kubernetes-based container platforms.
- Drive reliability, scalability, performance, and availability of production environments.
- Design and support microservices-based architectures in hybrid cloud environments.
- Develop automation solutions and operational tooling using Golang or Python.
- Implement and maintain CI/CD and GitOps workflows.
- Troubleshoot complex platform and application issues, perform root cause analysis, and drive preventive improvements.
- Collaborate with development, infrastructure, and DevOps teams to ensure seamless deployment and operations.
- Champion SRE best practices, observability, automation, and continuous improvement initiatives.
Required Experience
- 5+ years of hands-on experience with Kubernetes in production environments.
- Strong background in SRE, DevOps, platform engineering, or cloud operations.
- Experience in continuous delivery, deployment automation, and infrastructure operations.
Must-Have Skills
Expert-level Kubernetes experience
Strong programming experience in Golang or Python
Experience with containerized and microservices-based architectures
Preferred Skills
- OpenShift administration and support (Good to Have / Optional)
- GitOps practices and tools
- Terraform
- Ansible Automation
- Jenkins
- Hybrid Cloud environments
Required Competencies
- Robust analytical and problem-solving abilities.
- Excellent communication and stakeholder management skills.
- Ability to work effectively with cross-functio