Seeking a Senior DevOps / AI Platform Engineer with expertise in cloud infrastructure, Kubernetes, DevOps, AI platform operations, and Site Reliability Engineering (SRE). The ideal candidate should have hands-on experience building secure, scalable, and highly available AI platforms across Azure, AWS, and GCP, with strong skills in Infrastructure as Code, CI/CD, DevSecOps, AI FinOps, and cloud cost optimization.Mandatory Skills
- 8–15+ years of DevOps/Platform Engineering experience
- Azure, AWS &
- Google Cloud Platform (GCP)
- Kubernetes (AKS, GKE, EKS) &
- Docker
- Infrastructure as Code (Terraform, Bicep)
- CI/CD (GitHub Actions, Azure DevOps, GitOps)
- Site Reliability Engineering (SRE)
- Monitoring &
- Observability (Prometheus, Grafana, Azure Monitor, Cloud Monitoring, Datadog)
- AI Platform Operations &
- Model Health Monitoring
- DevSecOps &
- AI Governance
- AI FinOps, GPU Optimization &
- Cloud Cost Optimization
- Identity &
- Access Management (Entra ID / IAM)
- Python, Go, Java,
or C#
Preferred Skills
- Enterprise AI platform implementation
- Production support for AI workloads
- High Availability &
- Disaster Recovery
- Incident, Problem &
- Change Management
- Drift Detection &
- Operational Dashboards
- Knowledge transfer and mentoring
- Hybrid Cloud &
- Landing Zones
Valuable Profile Indicators
- Experience managing enterprise AI platforms on Azure, AWS, or GCP
- Solid Kubernetes administration (AKS, EKS, GKE)
- Expertise in Terraform, Bicep, and GitOps
- Hands-on with monitoring tools such as Prometheus, Grafana, and Datadog
- Experience supporting production AI/ML workloads with defined SLAs
- Proven background in cloud security, AI FinOps, automation, and disaster recovery
Pay: ₹2,000,000.00 - ₹2,500,000.00 per year
Work Location: In person
📌 Senior DevOps / AI Platform Engineer (Pune)
🏢 Mahi-Sys
📍 Pune