29 Aug
|
Xoriant
|
Bengaluru
Dear Applicants,
We are looking for SRE role Immediate Joiner, Rotational shift.
Role Summary
We are seeking an experienced SRE-II (Site Reliability Engineer) with 5+ years of experience in Production Support, Site Reliability Engineering, Monitoring & Observability, Kubernetes, and Cloud Operations. The ideal candidate should possess robust troubleshooting skills, a proactive mindset toward reliability and automation, and be comfortable working in rotational shifts, night shifts, and on-call support schedules.
The candidate will be responsible for ensuring platform reliability, availability, performance monitoring, incident management, and operational excellence across cloud-native environments.
Key Responsibilities
- Monitor and maintain production systems to ensure high availability and reliability.
- Investigate and resolve incidents, performance bottlenecks, and infrastructure issues.
- Participate in rotational shifts, night shifts, and scheduled on-call support.
- Create, maintain, and enhance monitoring dashboards, alerting mechanisms,
and observability solutions.
- Troubleshoot Kubernetes clusters, pods, node-level issues, and resource constraints.
- Analyze system metrics, logs, and traces to identify operational risks and performance issues.
- Support cloud-based applications running on Google Cloud Platform (GCP).
- Drive operational improvements through automation and process optimization.
- Collaborate with development, platform, and infrastructure teams to improve system reliability.
- Participate in incident response, root cause analysis (RCA), and postmortem reviews.
- Track and improve key reliability metrics including SLA, SLO, SLI, MTTR, and MTTA.
Required Technical Skills
Operating Systems
Networking Fundamentals
Monitoring & Observability
Kubernetes
Google Cloud Platform (GCP)
SRE Core Concepts
Database Fundamentals
Kafka
Infrastructure as Code (IaC).
Regards.
📌 Job Opportunity For SRE Role With Maplelabs ( A Unit Of Xoriant) (Bengaluru)
🏢 Xoriant
📍 Bengaluru