18 Sep
|
MediaMint
|
Hyderabad
18 Sep
MediaMint
Hyderabad
- Must be flexible to work in night shift
- Prefer local to Hyderabad or who can attend in-person interview.
: As a DevOps Engineer (On-call) at MediaMint AI Platform, you'll be the guardian of our platform's 24/7 reliability. You'll maintain critical infrastructure, respond to incidents, and continuously improve our deployment pipelines and observability systems. This role requires someone comfortable with on-call responsibilities who thrives under pressure and views incidents as learning opportunities.
You'll work closely with engineering teams to ensure our AI agents run smoothly and our clients experience zero disruption.
Responsibilities:
- Participate in on-call rotation to respond to production incidents and ensure rapid resolution
- Maintain and optimize cloud infrastructure (AWS/GCP/Azure) for high availability and performance
- Automate deployment pipelines using CI/CD tools for fast, reliable releases
- Monitor system performance, set up alerts, and proactively address potential issues
- Handle escalations from support and engineering teams during critical outages
- Improve observability through logging, monitoring, and distributed tracing systems
- Implement disaster recovery procedures and conduct regular backup/restore testing
- Manage containerized workloads using Docker and Kubernetes
- Conduct capacity planning and cost optimization for cloud infrastructure
Must-Have Requirements:
- 7+ years of DevOps/SRE experience in production environments
- Solid experience with at least one major cloud platform (AWS/GCP/Azure)
- Hands-on experience with Docker and Kubernetes for container orchestration
- Proficiency with CI/CD tools (Jenkins, GitLab CI, GitHub Actions, CircleCI)
- Proven incident management skills with experience in on-call rotations Strong scripting skills in Python, Bash, or similar languages
- Comfortable working in on-call rotation (including nights and weekends as needed)
Nice-to-Have Requirements:
- Experience with SRE practices and error budgets
- Infrastructure as Code experience with Terraform or CloudFormation
- Hands-on experience with observability tools (Datadog, Grafana, Prometheus, New Relic)
- Experience with PagerDuty or similar incident management platforms
- Knowledge of service mesh technologies (Istio, Linkerd)
- Familiarity with chaos engineering and resilience testing
📌 Senior Devops Engineer (Hyderabad)
🏢 MediaMint
📍 Hyderabad