- Monitor application health, consumption, and performance; propose and implement optimizations.
- Perform regular functional deliveries, releases, and post deployment after care.
- Maintain an up to date knowledge base and capture lessons learned from incidents.
- Triage, prioritize, and track incidents, problems, and service requests in ServiceNow & Jira.
- Act as escalation point for complex issues; coordinate with development, architecture and third party vendors.
- Lead root cause analysis (RCA) and post incident reviews; drive corrective and preventive actions.
- Deploy releases using CI/CD pipelines (Bitbucket, Git, Jenkins, Ansible, Terraform, Kubernetes, Docker).