Job Summary
Own the path from a merged pull request to a running service, and the pager that rings when it goes wrong. We are looking for someone who treats infrastructure as a product with users, and those users are our engineers. Today we run a mix of AWS and managed hosting. Some of it is beautifully automated. Some of it is not, and fixing that is a large part of this job.
If your idea of a positive week is deleting a manual runbook because you automated it, we should talk.
Responsibilities
- CI and CD pipelines across multiple product teams
- Infrastructure as code using Terraform, and container workloads on ECS or Kubernetes
- Monitoring, alerting and the incident process, including honest postmortems
- Cost visibility, because unreviewed cloud spend grows on its own
Requirements
- Four or more years in DevOps, platform or SRE work
- Strong Linux fundamentals and scripting in Bash or Python
- Real Terraform experience, not just having read about it
- Judgement about what deserves an alert at three in the morning and what does not
- Security instincts around secrets, network boundaries and least privilege
Skills
- AWS
- Terraform
- Docker
- Kubernetes
- Linux
- Bash
- Python
- Prometheus
- Grafana
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 DevOps / Site Reliability Engineer (Pune)
🏢 Moonfire
📍 Pune