05 Aug
|
HCL Technologies
|
Bengaluru
05 Aug
HCL Technologies
Bengaluru
Senior Administrator - Ansible, Terraform, GitHub
Experience: Not Available to Not Available years
Location: Bangalore, India
Skills: Linux, Python, Go, Java, Bash, AWS, Azure, GCP, Kubernetes, Docker, Terraform, Ansible, CI/CD, Splunk, moongsoft
Job Summary
To be effective, he must possess a blend of software engineering and systems administration skills. Automation & Toil Reduction: Identifying repetitive manual tasks (known as "toil") and writing code to automate them.
Incident Response: Acting as a "first responder" to production outages, diagnosing root causes, and implementing fixes to prevent recurrence.
Monitoring & Alerting: Setting up dashboards (Splunk) and monitoring tools (like moongsoft) to track system health and performance metrics.
Defining Reliability Targets: Establishing Service Level Objectives (SLOs) and tracking Service Level Indicators (SLIs) to measure whether a service meets user expectations.
Post-Mortems: Conducting "blameless" reviews after major incidents to document lessons learned and systematic improvements.
Key Responsibilities
Core:
• Linux expertise
• Programming/Scripting: Proficiency in languages like Python, Go, or Java for tool building and Bash for automation.
• Cloud Platforms: Hands-on experience managing infrastructure on AWS, Azure, or GCP.
• Containerization: Expertise with Kubernetes and Docker to manage and orchestrate distributed systems.
• Infrastructure as Code (IaC): Using tools like Terraform or Ansible to provision and manage resources consistently.
• CI/CD:
Understanding continuous integration and deployment pipelines to ensure protected and frequent software releases.
Skill Requirements
• Automation & Toil Reduction: Identifying repetitive manual tasks (known as "toil") and writing code to automate them.
• Incident Response: Acting as a "first responder" to production outages, diagnosing root causes, and implementing fixes to prevent recurrence.
• Monitoring & Alerting: Setting up dashboards (Splunk) and monitoring tools (like moongsoft) to track system health and performance metrics.
• Defining Reliability Targets: Establishing Service Level Objectives (SLOs) and tracking Service Level Indicators (SLIs) to measure whether a service meets user expectations.
• Post-Mortems: Conducting "blameless" reviews after major incidents to document lessons learned and systematic improvements.
Other Requirements
• Automation & Toil Reduction: Identifying repetitive manual tasks (known as "toil") and writing code to automate them.
• Incident Response: Acting as a "first responder" to production outages, diagnosing root causes, and implementing fixes to prevent recurrence.
• Monitoring & Alerting: Setting up dashboards (Splunk) and monitoring tools (like moongsoft) to track system health and performance metrics.
• Defining Reliability Targets: Establishing Service Level Objectives (SLOs) and tracking Service Level Indicators (SLIs) to measure whether a service meets user expectations.
• Post-Mortems: Conducting "blameless" reviews after major incidents to document lessons learned and systematic improvements.
📌 Senior Administrator Ansible, Terraform, GitHub (Bengaluru)
🏢 HCL Technologies
📍 Bengaluru