02 Oct
|
TalentStack
|
Mumbai
02 Oct
TalentStack
Mumbai
- Own the AWS estate end to end, including ECS, RDS, networking, and infrastructure provisioned through
Terraform.
- Build and maintain GitHub Actions pipelines to build, test, and promote code from UAT to production, with
deployment gates and reliable rollback strategies.
- Define and run versioning, release testing, and rollout processes across shared, dedicated, on-premise, and air-
gapped customer environments.
- Monitor production, respond to incidents and after-hours alerts, and resolve L2 escalations that L1 support
cannot address. Improve reliability to prevent repeat failures.
- Run vulnerability scanning across dependencies, containers, and code; assess CVE relevance and prioritize
remediation. Own secrets, access controls, container hardening, and key rotation.
- Lead technical responses to customer security due diligence, explaining implemented controls and documenting
incident response procedures.
- Own backups and disaster recovery for every customer setting,
including recovery approaches for on-
premise and air-gapped deployments without cloud access.
- Track and optimize cloud spend through right-sizing, cost visibility, and early detection of unexpected usage;
explain infrastructure costs and their drivers.
What you will bring
- Several years of production AWS experience and hands-on Terraform or equivalent infrastructure-as-code
expertise.
- Experience building and maintaining CI/CD pipelines, ideally GitHub Actions, with deployment gating and
rollback strategies.
- Working application security knowledge covering dependency scanning, secret management, container
hardening, and practical CVE assessment.
- Experience owning cloud cost accountability and making independent architectural decisions in a small team.
📌 Principal Platform Reliability & Security Engineer (Mumbai)
🏢 TalentStack
📍 Mumbai