20 Aug
|
HCLTech
|
Chennai
Chennai, Tamil Nadu
Job Summary
Key Responsibilities
1. Collaborate within team to identify and address critical system issues.
2. Implement automation tools and processes to improve system reliability and efficiency.
3. Perform regular system audits to ensure compliance with security standards and best practices.
4. Respond to and resolve escalated technical support issues in a timely manner.
5. Develop and maintain documentation related to system configurations, processes, and procedures.
Skill Requirements
1. Proficiency in programming languages such as python, java, or go.
2. Handson experience with cloud platforms like aws, azure, or google cloud.
3. Strong knowledge of containerization tools like docker and orchestration tools like kubernetes.
4. Familiarity with monitoring and logging tools such as prometheus, grafana, elk stack.
5.
Good problem-solving skills and the ability to work under pressure in a quick paced environment.
Other Requirements
5+ years in SRE/DevOps or L2 operations for cloud-native stacks; strong AWS production experience.
Proven incident/change/problem management in 24×7 environments; adept at RCA and postmortems.
Hands-on with observability tooling and operational automation; excellent collaboration and documentation skills.
body.unify div.unify-button-container.unify-apply-now: focus, #body.unify div.unify-button-container.unify-apply-#body.unify div.unify-button-container.unify-apply-now: focus, #body.unify div.unify-button-container.unify-apply
📌 Senior Site Reliability Engineer (Chennai)
🏢 HCLTech
📍 Chennai