SENIOR DEVOPS ENGINEER — INFRASTRUCTURE & KUBERNETES OPERATIONS
ABOUT THE ROLE
We are looking for a Senior DevOps Engineer to join our Infrastructure team, focused on Kubernetes operations and automation at scale. You will drive the strategy and own the health, scalability, and reliability of our Kubernetes clusters, while architecting and leading the development of tooling that platform and service teams rely on to manage infrastructure safely and efficiently. You will work closely with service owners, platform engineers, and tooling teams — and mentor junior engineers — to keep clusters running smoothly and to automate away manual, error-prone operational work.
WHAT YOU'LL DO
- Own and drive the roadmap for Kubernetes cluster operations — lead cluster upgrades, manage node groups (scaling, draining, replacement), and maintain overall cluster health across environments at scale.
- Lead troubleshooting of complex production and non-production issues — use kubectl, logs, metrics, and other diagnostics to identify root cause and resolve workload, node, and networking failures,
often across multiple interdependent systems.
- Architect and lead development of internal tooling — design, implement, and maintain automation (scripts, CLIs, controllers, operators) that helps teams manage infrastructure more reliably and with less manual effort.
- Drive automation strategy for operational workflows — replace manual runbooks with scripts and tools that handle upgrades, remediation, scaling, and routine maintenance.
- Diagnose and resolve complex infrastructure blockers — debug deployment failures, node/pod scheduling issues, resource constraints, and misconfigurations across clusters.
- Define and improve observability strategy — instrument logging, metrics, and alerting to increase visibility into cluster health and reduce time-to-detect/resolve.
- Set standards for issue tracking and triage — author detailed bug reports capturing root cause, repro steps, and impact; drive
📌 Senior DevOps Engineer (India)
🏢 Artech
📍 India