Solid expertise in Linux, Kubernetes, Docker and Big Data with deep technical proficiency.
Support production stability by monitoring, troubleshooting, and resolving service issues across distributed systems.
Perform initial triage and coordination for producton problems, following defined procedures and escalation paths.
Support Kubernetes-based workloads, including pod health, service availability, resource usage, and basic networking issues.
Assist with deployment and change activities, validating configurations, monitoring rollouts, and ensuring minimal disruption.
Support application build and delivery workflows, including basic code compilation, container image builds, and artifact validation.
Assist in troubleshooting data platforms and pipelines, including batch and streaming workloads.
Provide oncall and troubleshooting support for big data platforms and services, including Flink, Hive, Yarn, Spark, and related data processing workflows.
Maintain transparent operational communication, documentation, and shift handoffs while collaborating with engineering and platform teams.