- Build, configure, and maintain AWS cloud infrastructure.
- Build, administer, and optimize Kubernetes clusters and cloud-native platforms.
- Support the implementation and maintenance of CI/CD automation pipelines.
- Manage GitHub repositories, branching strategies, and release workflows.
- Monitor infrastructure health, performance, and availability.
- Troubleshoot and resolve production, platform, and infrastructure incidents.
- Ensure compliance with security, patching, governance, and operational standards.
- Administer Linux-based environments, preferably SUSE Linux Enterprise Server (SLES).
- Collaborate with development, security, and operations teams to improve system reliability and operational efficiency.
- Apply Site Reliability Engineering (SRE) principles and practices to improve system reliability, scalability, availability, performance, observability, and operational efficiency across the AWS landscape.