Job Title: SRE Practice Lead – Engineering Services
Skills: SRE, Reliability Engineering Frameworks, SLI / SLO / Error Budget Management, Incident Management & RCA, Service Resilience & Availability Engineering, Capacity Planning, Cloud & Platform Engineering, DevOps & Automation, Observability & AIOps, COE Setup, Competency Development
Experience: 18+ Years
Location: Greater Noida
Job Summary:
We are seeking a seasoned SRE Practice Lead to drive Reliability Engineering transformation across enterprise clients through a combination of consulting, solutioning, presales, and delivery leadership.
This role requires a leader who has built and scaled SRE capabilities within an engineering services or system integration environment and has successfully partnered with clients to modernize operations through cloud-native engineering, observability, automation, platform engineering, and AI-enabled reliability practices. The ideal candidate should possess solid customer-facing experience, hands-on solution architecture capabilities, and the ability to lead end-to-end transformation initiatives, from presales and consulting through execution and value realization.
Note: Candidates from engineering services, digital engineering, cloud transformation, infrastructure modernization, or system integration organizations will be strongly preferred. Candidates whose experience is primarily limited to governance, enterprise architecture, or captive/shared-services environments may not be suitable unless they demonstrate substantial presales and delivery ownership.
Key Responsibilities:
Practice Leadership:
- Define and scale enterprise-wide SRE strategies, frameworks, operating models, and best practices
- Lead transformation engagements that move clients from traditional support-centric operations to engineering-led, automation-first reliability models
- Establish reusable assets, accelerators,
reference architectures, and service offerings aligned to modern managed services and digital engineering
- Drive adoption of AI-enabled SRE practices, predictive operations, intelligent automation, and self-healing platforms
Solutioning & Presales Leadership:
- Partner with sales and account teams to develop winning SRE, observability, platform engineering, and cloud transformation solutions
- Lead customer workshops, discovery sessions, assessments, and executive-level discussions
- Own solution architecture, effort estimation, commercials, response to RFPs/RFIs, and technical proposals
- Present solution strategies, transformation roadmaps, business cases, and value realization models to customer stakeholders
- Support deal pursuits through Proof of Concepts, demonstrations, solution reviews, and technical due diligence
Engineering & Delivery Leadership:
- Lead large-scale SRE and platform engineering programs across cloud-native and distributed environments
- Design highly resilient, scalable, self-healing systems leveraging Kubernetes, containers, cloud services, observability platforms, and modern DevOps toolchains
- Drive implementation of reliability metrics including SLIs, SLOs, Error Budgets, MTTR reduction, and operational excellence initiatives
- Govern delivery quality, transformation outcomes, stakeholder management, and customer success
- Drive automation across provisioning, deployment, incident response, monitoring, and operational workflows
Customer Consulting:
- Advise customers on SRE maturity assessments, operating model transformation, cloud adoption, observability strategy, and reliability roadmaps
- Articulate business outcomes around availability, performance, productivity, operational efficiency, and cost optimization
- Act as a trusted advisor to engineering leadership, architecture teams, and executive stakeholders
People & Capability Development:
- Build and mentor high-performing SRE and platform engineering teams
- Define capability roadmaps, learning paths, certifications, and engineering standards
- Foster a culture of reliability, automation, innovation, and continuous improvement
Role Competencies:
- 15+ years of experience across SRE, Platform Engineering, DevOps, Cloud Engineering, or Infrastructure Transformation
- Strong experience working within global engineering services, digital engineering, or system integration organizations
- Proven experience leading client-facing solutioning, architecture discussions, consulting engagements, and presales pursuits
- Demonstrated ownership of large transformation programs from proposal stage through delivery execution
- Strong expertise in AWS, Azure, and/or GCP environments
- Deep understanding of SRE principles including SLIs, SLOs, Error Budgets, Incident Management, Observability, and Reliability Engineering
- Experience designing cloud-native platforms using Kubernetes, containers, IaC, CI/CD, and automation frameworks
Preferred Skills
- Terraform, Ansible, GitOps, CI/CD toolchains
- Prometheus, Grafana, Dynatrace, Datadog, Splunk, ELK, OpenTelemetry
- Platform Engineering and Internal Developer Platform experience
- AI-driven operations, AIOps, predictive monitoring, and intelligent automation
- Executive stakeholder management and consulting skills
📌 SRE Practice Lead (Noida)
🏢 Coforge
📍 Noida