Job Summary
EXL is looking for a skilled Managed Service Engineer with strong hands-on experience in AWS Cloud, DevOps, Infrastructure Monitoring, Automation, Incident Management, and Production Support.
The ideal candidate will be responsible for ensuring the stability, availability, performance, and security of cloud-native platforms. This role requires expertise in managing cloud infrastructure, CI/CD pipelines, monitoring solutions, incident response, and operational support processes while collaborating with cross-functional teams to deliver high-quality managed services.
Key Responsibilities
Managed Services & Production Support
- Provide day-to-day operational support for cloud infrastructure, applications, and data platforms.
- Monitor system health, availability, performance, and security alerts across production environments.
- Manage incidents, service requests, change requests, and problem management activities.
- Perform root cause analysis (RCA) and implement preventive and corrective actions.
- Ensure compliance with SLAs, OLAs, and support governance frameworks.
- Create and maintain runbooks, SOPs, knowledge articles, and operational documentation.
- Support BAU operations while continuously identifying opportunities for service improvements.
DevOps Engineering
- Manage and support CI/CD pipelines for application and infrastructure deployments.
- Automate repetitive operational and deployment tasks using scripting and DevOps tools.
- Support build, release, and deployment activities across development, testing, and production environments.
- Collaborate with development teams to troubleshoot deployment and environment-related issues.
- Implement Infrastructure as Code (IaC) solutions for environment provisioning and configuration management.
- Ensure deployment reliability, scalability, and operational efficiency.
AWS Cloud Operations Manage and support AWS services including:
- EC2
- S3
- IAM
- VPC
- Lambda
- RDS
- CloudWatch
- CloudTrail
- ELB / ALB
- Auto Scaling
- ECS / EKS (preferred)
Additional responsibilities include:
- Monitor and optimize AWS resource utilization.
- Support cloud security, patch management, backup and recovery processes.
- Manage access controls, permissions, and IAM policies.
- Assist in cloud cost optimization initiatives and governance activities.
- Ensure highly available and secure cloud environments.
Monitoring, Incident & Problem Management
- Configure and maintain monitoring dashboards, alerts, and logging solutions.
- Respond to incidents and alerts based on defined severity levels and escalation procedures.
- Coordinate with application, infrastructure, security, and business stakeholders during outages.
- Conduct proactive health checks and preventive maintenance activities.
- Support post-incident reviews, problem management, and continuous service improvement initiatives.
- Ensure timely resolution of production issues and operational risks.
Security & Compliance
- Follow AWS and enterprise security best practices for access management, network security, and data protection.
- Manage IAM roles, permissions, key rotation, and vulnerability remediation activities.
- Support audit requests, compliance reviews, and operational risk assessments.
- Ensure adherence to organizational security, compliance, and change management policies.
- Participate in disaster recovery planning and operational readiness activities.
Required Technical Skills
Cloud Technologies
- Strong hands-on experience with AWS Cloud Services.
- Working knowledge of cloud operations, networking, and infrastructure management.
DevOps & Automation
Experience with:
- Jenkins
- GitHub Actions
- GitLab CI/CD
- Azure DevOps
- Git / Bitbucket
- Terraform
- AWS CloudFormation
- Docker
- Kubernetes
- Ansible (preferred)
Operating Systems & Scripting
- Linux/Unix Administration
- Shell Scripting (Bash)
- Python scripting (preferred)
Monitoring & Observability
Experience with:
- AWS CloudWatch
- Splunk
- Datadog
- Grafana
- Prometheus
- ELK Stack (Elasticsearch, Logstash, Kibana)
Networking
Strong understanding of:
- VPC
- Subnets
- Route Tables
- Security Groups
- Load Balancers
- DNS
- VPN Connectivity
IT Service Management
Experience with:
- ServiceNow
- Jira
- Other ITSM tools
Security & Governance
- Identity & Access Management (IAM)
- Backup & Disaster Recovery
- Vulnerability Management
- Cloud Governance and Compliance
Required Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related discipline.
- AWS certifications such as AWS Cloud Practitioner, AWS Solutions Architect Associate, or AWS SysOps Administrator are preferred.
- Additional certifications in DevOps, Kubernetes, Terraform, or ITIL are a plus.
Experience Requirements
- 3 to 8 years of experience in Cloud Operations, DevOps, Production Support, IT Operations, or Managed Services.
- Minimum 2+ years of hands-on experience working with AWS environments.
- Proven experience supporting production applications and cloud infrastructure.
- Experience handling incidents, change management, service requests, and operational support activities.
- Exposure to 24x7 support environments, on-call rotations, or shift-based operations is preferred.
Preferred Competencies
- Strong analytical and problem-solving skills.
- Excellent troubleshooting and incident management capabilities.
- Ability to work independently in a rapid-paced environment.
- Strong communication and stakeholder management skills.
- Continuous improvement mindset with focus on automation and operational excellence.
- Ability to manage multiple priorities and critical production issues effectively.
📌 Aws Devops Engineer (Pune)
🏢 EXL
📍 Pune