02 Aug
|
Oracle
|
Bengaluru
About the Role
Join Autonomous Recovery Service (RCV)—a mission-critical cloud service that protects customer data and ensures reliable backup and recovery at cloud scale.
This is a hands-on technical leadership role, not a traditional people-management position. You will lead a high-performing engineering team while remaining deeply involved in production operations, technical escalations, incident response, service reliability, architecture reviews, automation, and continuous improvement.
You'll work with cutting-edge technologies including Oracle Database, RMAN, Zero Data Loss Recovery Appliance (ZDLRA), Exadata, OCI, Linux, Python, Terraform, and Site Reliability Engineering (SRE) to build and operate highly available cloud services.
What You'll Do
• Lead the engineering team responsible for Oracle Autonomous Recovery Service (RCV).
• Stay hands-on by participating in production support, incident response, technical escalations, troubleshooting, and architecture reviews.
• Drive service reliability, scalability, security, observability, and operational excellence.
• Own the end-to-end lifecycle of production services, including provisioning, deployments, upgrades, patching, backup, recovery, and maintenance.
• Lead root cause analysis (RCA), postmortems, and long-term corrective actions.
• Improve monitoring, alerting, automation, self-service tooling, and operational efficiency.
• Guide capacity planning, performance tuning, and infrastructure scalability.
• Partner closely with Software Engineering,
Database Engineering, Infrastructure, Security, and Product teams.
• Mentor engineers, conduct technical reviews, develop talent, and foster a robust engineering culture.
• Drive adoption of automation, Infrastructure as Code, and AI-assisted engineering practices to reduce operational toil.
Required Technical Skills
• Oracle Database Administration
• Oracle Recovery Manager (RMAN)
• Zero Data Loss Recovery Appliance (ZDLRA)
• Oracle Exadata
• Oracle Cloud Infrastructure (OCI)
• Linux Administration
• Site Reliability Engineering (SRE)
• Backup, Restore & Disaster Recovery
• Production Operations
• Incident Management
• Root Cause Analysis
• Capacity Planning
• Monitoring & Observability
• Python
• Bash / Shell Scripting
• SQL
• Terraform (Infrastructure as Code)
• CI/CD
• Distributed Systems
• Automation & Orchestration
Preferred Qualifications
• Experience leading technical teams while remaining hands-on.
• Strong production support and cloud operations experience.
• Expertise operating mission-critical distributed systems.
• Experience improving reliability, scalability, and operational excellence.
• Passion for mentoring engineers and building high-performing teams.
Why Join Oracle
You'll help build and operate one of Oracle Cloud Infrastructure's most critical services, protecting customer data at global scale while working alongside world-class engineers on cutting-edge cloud technologies.
📌 Hands-on Technical Manager-Oracle Database/SRE/Exadata (Bengaluru)
🏢 Oracle
📍 Bengaluru