Greetings from Maneva!
Job Title - Site Reliability Engineer (SRE) Lead
Experience - 7 – 10 Years
Location - PAN India
Notice - Immediate Joiner
Requirements:
A Senior SRE drives enterprise-wide reliability, observability, automation, and operational excellence across large-scale hybrid environments. This level requires architectural judgement, leadership in high-severity incidents, and the ability to mature SRE practices. operational functions performed by Site Reliability Engineering teams to ensure availability, reliability, performance, and resilience of applications and infrastructure.
Key Responsibilities
- Reliability Engineering & Service Governance
- Define SLIs/SLOs/SLAs and govern error budgets.
- Lead architecture reviews for reliability and resilience.
- Drive SRE adoption and reduce toil.
- Maintain and improve service availability, resilience, and SLO/SLI governance.
- Apply SRE principles (risk management, error budget tracking, elimination of toil).
- Evaluate application readiness and architecture for reliability standards.
- Observability & AIOps
- Architect observability platforms (AppDynamics, Datadog, Prometheus, Grafana, Splunk, Dynatrace, ELK).
- Implement logging, metrics, and tracing.
- Improve visibility, reduce noise, and optimize MTTD/MTTI/MTTR.
3. Incident,
Problem & Change Management
- Lead major incident response and RCA.
- Execute capacity, security, and change control processes.
- Improve operational processes (capacity, security, change).
- Automation & Platform Engineering
- Build automation for deployments, monitoring, remediation.
- Design CI/CD pipelines and IaC.
- Drive AIOps adoption.
- Leadership & Collaboration
- Mentor teams and lead reliability initiatives.
- Influence roadmaps and enforce reliability standards.
Required Technical Expertise
- Solid hands-on experience in Unix/Linux, Shell scripting, Python/Java/Go/NodeJS.
- Deep expertise in monitoring & observability tools
- Distributed systems knowledge.
- Incident leadership and RCA.
- Proven CI/CD pipeline and IaC experience.
- Dashboard/KPI/metric creation and tracking.
Experience Requirements
- 5+ years in SRE/DevOps/Production Engineering.
- Experience with mission-critical, large-scale systems.
- Ability to work with architects and leadership.
- Experience operating large-scale, distributed systems using SRE practices.
If you are excited to grab this opportunity, please apply directly or share your CV at
[email protected] and
[email protected].
📌 Site Reliability Engineer (SRE) Lead (Bengaluru)
🏢 Wipro
📍 Bengaluru