We are seeking a Principal SRE to serve as a technical authority, shaping the platform strategy, reliability, and security practices of our cloud infrastructure. This is a hands-on leadership role combining deep cloud-native architecture with mentorship and regional team growth.
Core Responsibilities:
- Platform Architecture: Define multi-account AWS standards, zero-trust network boundaries, and resilience/disaster recovery strategies.
- Kubernetes & Automation: Standardize and upgrade enterprise Amazon EKS clusters; build reusable provisioning automation using Terraform, Python, or Go .
- CI/CD Modernization: Own the evolution of enterprise CI/CD systems, migrating pipelines to modern GitOps practices with Argo CD and Helm .
- Observability & Response: Establish organization-wide SLO/SLI policies using Datadog/Prometheus/Grafana and lead critical incident root-cause analyses.
- Governance: Partner with security for SOC 2/ISO compliance and drive rigorous cloud cost-optimization strategies.
Technical Requirements:
- Cloud/K8s: Deep expertise in AWS ecosystems and enterprise-scale Amazon EKS.
- IaC & Tooling: Robust skills in Terraform, CloudFormation, Ansible, and Jenkins (Pipeline-as-code).
- Coding/Scripting: Proficiency in Python, Go, Groovy, or Shell scripting.
- Education: B.E / B.Tech / M.Tech / MCA in Computer Science or IT.
If you are ready to raise the engineering standard and build a resilient, secure, and cost-optimized ecosystem, let’s talk! Write to
[email protected] to get connecetd!
📌 Principal Site Reliability Engineer | AWS | EKS | Kubernetes | Terraform | GitOps | Platform Architecture (Bengaluru)
🏢 CareerXperts Consulting
📍 Bengaluru