13 Sep
|
Ericsson
|
Noida
Join our Team:
About this chance:
Join a cross-functional infrastructure team responsible for the end-to-end administration, operations, and resilience of cloud-native infrastructure and container platforms.
Own the health, capacity, lifecycle and disaster recovery readiness of Cloud Container Distribution (CCD) and related CNIS components across on-prem and cloud environments.
Drive automation, runbook maturity, and operational excellence through Infrastructure-as-Code, observability, and disciplined change/incident processes.
What you will do:
Administer and operate CNIS platforms including Kubernetes clusters, container runtimes, and cloud-native infrastructure components; handle provisioning, upgrades, backups, restores and lifecycle tasks for CCD.
Perform regular CCD health checks, configuration drift analysis, capacity tracking, workplace validation, and readiness testing across clusters.
Lead CCD and cluster-level disaster recovery planning, periodic DR execution, and post-incident RCA for major outages.
Implement and maintain Infrastructure-as-Code automation (Ansible) for repeatable CNIS/CCD operations and on-boarding workflows.
Monitor system performance and availability using tools such as Prometheus and Grafana; tune alerts and dashboards and drive performance remediation.
Maintain and continuously improve operational documentation: runbooks, MOPs, checklists, and DR playbooks.
Troubleshoot complex cross-domain issues spanning SDI, storage, compute, OS, and networking layers; coordinate fixes and communicate root causes.
The skills you bring:
Solid experience administering Kubernetes (cluster ops, upgrades, backup/restore, lifecycle) and container runtimes in production.
Hands-on with Cloud Container Distribution (CCD) operations and cluster validation practices.
Proficient in Infrastructure-as-Code and automation using Ansible; comfortable writing and maintaining playbooks for infra tasks.
Solid observability experi
📌 It System Expert 5 Noida
🏢 Ericsson
📍 Noida