Senior Site Reliability Engineer (Chennai)

Senior Site Reliability Engineer (Chennai)

05 Aug
|
Important Group
|
Chennai

05 Aug

Important Group

Chennai

Description

This role is for one of the Weekday's clients

Salary range: Rs 3000000 - Rs 4500000 (ie INR 30-45 LPA)

Min Experience: 7+ years

Location: Chennai

JobType: full-time

The Senior SRE is responsible for deployment, updates, and operational support for environments hosting our leading client’s cloud-based solutions. This role ensures operational excellence, a seamless client experience, and continuous improvement across infrastructure and delivery processes. The ideal candidate combines strong technical capabilities with the ability to lead delivery through influence and hands-on engineering expertise.

Requirements

Key Responsibilities

- Manage deployments, upgrades, maintenance, and operational support for cloud environments.
- Ensure high availability, scalability, performance, and reliability of production systems.
- Define, monitor, and improve SLAs, SLOs, and SLIs.
- Drive automation initiatives and Infrastructure as Code (IaC) adoption.
- Perform Root Cause Analysis (RCA) and implement preventive actions.
- Optimize cloud infrastructure, operational efficiency, and costs.
- Enhance monitoring, observability, security, and deployment processes.
- Collaborate with Engineering, Project Management, Customer Success, and cross-functional teams to deliver reliable services.

Required Skills

- Strong hands-on experience with AWS cloud platforms.
- Expertise in Kubernetes for container orchestration and cluster management.




- Experience with Terraform for Infrastructure as Code (IaC).
- Proficiency in Ansible for configuration management and automation.
- Hands-on experience with Helm for Kubernetes application deployments.
- Experience managing MariaDB and MongoDB databases in production environments.
- Strong understanding of CI/CD pipelines, deployment automation, and DevOps practices.
- Experience with monitoring, observability, logging, and alerting tools (e.g., Prometheus, Grafana, ELK, CloudWatch, Azure Monitor).
- Valuable knowledge of Linux administration, networking, DNS, load balancing, and cloud security best practices.
- Experience troubleshooting production environments, conducting Root Cause Analysis (RCA), and improving platform reliability.
- Understanding of SRE principles, including SLAs, SLOs, and SLIs.
- Scripting experience using Bash, Python, or Shell for automation.
- Excellent problem-solving, communication, and stakeholder management skills.

Preferred Experience

- Experience working in Site Reliability Engineering, DevOps, or Cloud Operations roles.
- Experience managing large-scale, production cloud environments.
- Ability to thrive in a fast-paced, customer-focused setting.
- Strong analytical mindset with a proactive approach to continuous improvement.

Must-have skills

AWS, Kubernetes, Site Reliability Engineering

Good-to-have skills

Helm Charts, IaC, monitoring

📌 Senior Site Reliability Engineer (Chennai)
🏢 Important Group
📍 Chennai

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer (chennai) / chennai

Subscribe to this job alert:

Get the latest job offers by email for: senior site reliability engineer (chennai) / chennai