SeniorAdministrator - Kubernetes, Terraform (India)

SeniorAdministrator - Kubernetes, Terraform (India)

18 Sep
|
HCLTech
|
India

18 Sep

HCLTech

India

SeniorAdministrator - Kubernetes, Terraform

Chennai, Tamil Nadu

Job Summary

AWS SRE Admin: The SRE Administrator is responsible for maintaining the reliability, availability, performance, and operational stability of applications and infrastructure hosted on AWS. The role focuses on monitoring, incident management, platform administration, observability, automation, and operational excellence to ensure seamless business services across the technology estate. The AWS SRE Administrator works closely with application, cloud, infrastructure, security, and operations teams to proactively identify issues, optimize platform performance, and drive continuous service improvements. Key Responsibilities: Monitor AWS-hosted applications, infrastructure, and services to ensure high availability and performance. Perform proactive health checks, alert monitoring, incident triage, and operational support activities. Manage AWS services including EC2, EKS, ECS, Lambda, RDS, S3, Route 53, CloudWatch, and related platform components. Investigate production incidents, perform root cause analysis, and coordinate resolution with support and engineering teams. Administer observability platforms and monitoring tools, including dashboard maintenance, alert tuning, and reporting. Support production releases by performing deployment validation, smoke testing, and post-change monitoring activities. Develop and maintain operational runbooks, standard operating procedures, and knowledge documentation. Automate repetitive operational tasks using AWS native services, scripting, and infrastructure-as-code tools. Monitor platform capacity, utilization trends, and system reliability metrics to support capacity planning. Collaborate with security, infrastructure, and application teams to ensure compliance with operational and security standards. Generate operational reports, SLA/KPI dashboards, and service health updates for stakeholders. Drive continuous improvement initiatives focused on reliability, automation, operational efficiency, and reduction of manual effort.

Key Responsibilities

AWS SRE Admin: The SRE Administrator is responsible for maintaining the reliability, availability, performance, and operational stability of applications and infrastructure hosted on AWS. The role focuses on monitoring, incident management, platform administration, observability, automation, and operational excellence to ensure seamless business services across the technology estate. The AWS SRE Administrator works closely with application, cloud, infrastructure, security, and operations teams to proactively identify issues, optimize platform performance, and drive continuous service improvements. Key Responsibilities: Monitor AWS-hosted applications, infrastructure, and services to ensure high availability and performance. Perform proactive health checks,



alert monitoring, incident triage, and operational support activities. Manage AWS services including EC2, EKS, ECS, Lambda, RDS, S3, Route 53, CloudWatch, and related platform components. Investigate production incidents, perform root cause analysis, and coordinate resolution with support and engineering teams. Administer observability platforms and monitoring tools, including dashboard maintenance, alert tuning, and reporting. Support production releases by performing deployment validation, smoke testing, and post-change monitoring activities. Develop and maintain operational runbooks, standard operating procedures, and knowledge documentation. Automate repetitive operational tasks using AWS native services, scripting, and infrastructure-as-code tools. Monitor platform capacity, utilization trends, and system reliability metrics to support capacity planning. Collaborate with security, infrastructure, and application teams to ensure compliance with operational and security standards. Generate operational reports, SLA/KPI dashboards, and service health updates for stakeholders. Drive continuous improvement initiatives focused on reliability, automation, operational efficiency, and reduction of manual effort.

Skill Requirements

AWS SRE Admin: The SRE Administrator is responsible for maintaining the reliability, availability, performance, and operational stability of applications and infrastructure hosted on AWS. The role focuses on monitoring, incident management, platform administration, observability, automation, and operational excellence to ensure seamless business services across the technology estate. The AWS SRE Administrator works closely with application, cloud, infrastructure, security, and operations teams to proactively identify issues, optimize platform performance, and drive continuous service improvements. Key Responsibilities: Monitor AWS-hosted applications, infrastructure, and services to ensure high availability and performance. Perform proactive health checks, alert monitoring, incident triage, and operational support activities. Manage AWS services including EC2, EKS, ECS, Lambda, RDS, S3, Route 53, CloudWatch, and related platform components. Investigate production incidents, perform root cause analysis, and coordinate resolution with support and engineering teams. Administer observability platforms and monitoring tools, including dashboard maintenance, alert tuning, and reporting.



Support production releases by performing deployment validation, smoke testing, and post-change monitoring activities. Develop and maintain operational runbooks, standard operating procedures, and knowledge documentation. Automate repetitive operational tasks using AWS native services, scripting, and infrastructure-as-code tools. Monitor platform capacity, utilization trends, and system reliability metrics to support capacity planning. Collaborate with security, infrastructure, and application teams to ensure compliance with operational and security standards. Generate operational reports, SLA/KPI dashboards, and service health updates for stakeholders. Drive continuous improvement initiatives focused on reliability, automation, operational efficiency, and reduction of manual effort.

Other Requirements

AWS SRE Admin: The SRE Administrator is responsible for maintaining the reliability, availability, performance, and operational stability of applications and infrastructure hosted on AWS. The role focuses on monitoring, incident management, platform administration, observability, automation, and operational excellence to ensure seamless business services across the technology estate. The AWS SRE Administrator works closely with application, cloud, infrastructure, security, and operations teams to proactively identify issues, optimize platform performance, and drive continuous service improvements. Key Responsibilities: Monitor AWS-hosted applications, infrastructure, and services to ensure high availability and performance. Perform proactive health checks, alert monitoring, incident triage, and operational support activities. Manage AWS services including EC2, EKS, ECS, Lambda, RDS, S3, Route 53, CloudWatch, and related platform components. Investigate production incidents, perform root cause analysis, and coordinate resolution with support and engineering teams. Administer observability platforms and monitoring tools, including dashboard maintenance, alert tuning, and reporting. Support production releases by performing deployment validation, smoke testing, and post-change monitoring activities. Develop and maintain operational runbooks, standard operating procedures, and knowledge documentation. Automate repetitive operational tasks using AWS native services, scripting, and infrastructure-as-code tools. Monitor platform capacity, utilization trends, and system reliability metrics to support capacity planning. Collaborate with security, infrastructure, and application teams to ensure compliance with operational and security standards. Generate operational reports, SLA/KPI dashboards, and service health updates for stakeholders. Drive continuous improvement initiatives focused on reliability, automation, operational efficiency, and reduction of manual effort.

📌 SeniorAdministrator - Kubernetes, Terraform (India)
🏢 HCLTech
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senioradministrator - kubernetes, terraform (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: senioradministrator - kubernetes, terraform (india) / india