Role: Senior SRE
Experience: 8+ Years
Location: Bangalore
Work Mode: Hybrid
Key Responsibilities
Linux Expertise:
Extensive experience working with Linux distributions such as RHEL and CentOS, including strong knowledge of shells, filesystems, and system utilities.
Automation:
Lead the design, development, and maintenance of automation scripts using Python, Ansible, and Shell to streamline infrastructure provisioning, configuration, monitoring, and operational processes.
Infrastructure as Code (IaC):
Implement and manage infrastructure automation using Terraform, Ansible, or similar frameworks to enable repeatable, scalable, and efficient deployments.
Container Orchestration:
Hands-on experience with container orchestration platforms such as Kubernetes (on-premises and Rancher), with a deep understanding of Kubernetes objects, configurations, and workloads.
Storage Management:
Robust experience with storage solutions, preferably ONTAP, including management of volumes, aggregates, backups, and disaster recovery (DR) planning.
Monitoring and Observability:
Implement and manage monitoring solutions using tools like Dynatrace, Apica, and Grafana, along with developing custom monitoring scripts using cron or Airflow.
Database Management:
Knowledge of SQL and NoSQL databases, with hands-on experience in database administration, performance tuning, and optimization.
CI/CD Pipelines:
Design, build, automate, and maintain CI/CD pipelines to ensure smooth, reliable, and efficient application deployment and delivery.
Cloud Platforms:
Strong expertise in cloud platforms, especially AWS, with the ability to manage cloud-native applications and services effectively.
Incident Management:
Experience in handling incidents, conducting root cause analysis (RCA), and implementing problem management practices to improve system stability and reliability.
Gen AI / MCP Integration:
Collaborate with cross-functional teams to design, develop, and deploy Generative AI solutions
📌 SRE_ Lead II (Bengaluru)
🏢 UST
📍 Bengaluru