Responsibilities
- Proven experience of enterprise grade Kubernetes clusters across development staging and production environments
- Implement infrastructureascode IaC solutions using tools such as Terraform Ansible or CloudFormation
- Manage container registry infrastructure and deployment pipelines
- Establish standardized Kubernetes configurations and best practices
- Monitor and maintain Kubernetes cluster health performance and uptime
- Manage cluster upgrades patches and lifecycle management with zerodowntime deployment strategies
- Implement comprehensive logging monitoring and observability solutions Prometheus ELK stack Grafana etc
- Troubleshoot and resolve complex infrastructure and application deployment issues
- Optimize resource allocation autoscaling policies and cost efficiency
- Establish runbooks and operational procedures for incident response
- Lead rootcause analysis on platform incidents and implement preventative measures
Required Skills
- 3 years of handson DevOps engineering experience with container orchestration platforms
- 3 years of proven productiongrade Kubernetes deployment and management experience
- Deep knowledge of Kubernetes architecture components and API objects Deployments StatefulSets DaemonSets Services Ingress etc
- Demonstrated expertise in rolling out and managing multinode Kubernetes clusters
- Proficiency with container technologies Docker containerd CRI
- Experience with Kubernetes networking storage and service mesh technologies Istio Cilium Calico
- Solid understanding of Linux Unix system administration and networking TCPIP DNS routing firewalls
- Proven track record implementing and maintaining secure Kubernetes environments Knowledge of container security best practices and vulnerability scanning
- Experience with RBAC network policies and pod security standards
- Understanding of encryption secrets management and identity protocols
- Handson experience with major cloud providers
- oAWS EKS EC2