: POSTGRESQL,JAVA SPRING BOOT,INCIDENT MANAGEMENT,REST APIS
Job Description:
Role Summary
Own end-to-end production stability — from hotfix delivery to root-cause closure. Serve as the final escalation point for critical incidents, with deep system knowledge to diagnose quick and ship safe fixes under pressure.
Key Responsibilities
- Production Hotfixes: Own the full diagnosis-to-deployment lifecycle; coordinate emergency releases with zero ambiguity.
- Deep-Dive Troubleshooting: Investigate complex, intermittent, and data-related issues across services, databases, and infra.
- SLA Ownership: Ensure P1/P2 tickets are resolved within agreed SLAs; escalate proactively when needed.
- Code & Patch Delivery: Write, review, and deploy targeted hotfix code to production-grade quality standards.
- RCA & Post-Mortems: Document thorough root-cause analyses with structured preventive action items.
- Change Advisory: Evaluate risk of hotfixes; liaise with release and CAB processes.
- Runbooks & SOPs: Maintain and improve operational runbooks for known failure patterns.