Role Summary:
"We’re seeking a seasoned Site Reliability Engineer (SRE) who thrives at the intersection of software engineering, infrastructure, and AI systems. You’ll help ensure our platforms are scalable, reliable, and secure—while also contributing code, automation, and architectural improvements that support both traditional services and AI-driven workloads.
This role is ideal for someone who thinks like a developer, understands AI infrastructure, and is passionate about reliability, observability, and operational excellence."
Key Roles & Responsibilities:
• Design, build, and maintain scalable infrastructure and automation tools for both traditional and AI-based systems.
• Develop software solutions to improve system reliability and reduce manual toil.
• Implement and manage CI/CD pipelines, including model deployment workflows.
• Monitor system performance, availability, and security using modern observability tools.
• Collaborate with data science and ML engineering teams to support AI/ML model training, serving, and lifecycle management.
• Lead incident response, root cause analysis, and postmortem processes.
• Advocate for SRE principles across engineering and AI teams.
Certifications (if any): Azure Foundation Certification and any SRE related.
Work Location: Bangalore/Pune/Delhi/Hyderabad
Additional Notes / Business Justification:
Building a solid SRE team for preparedness for upcoming Workplace cloud migration and currently ongoing BNFT migrations
📌 Web SRE (Bengaluru)
🏢 Voya India
📍 Bengaluru
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.