Key Responsibilities:
- 1. Database Provisioning & Infrastructure-as
- Code
- Deploy and configure current database instances (dev, test, prod) in accordance with system and storage requirements
- Build and maintain automated provisioning pipelines using Terraform, Ansible, and CI/CD tools
- Enforce infrastructure-as-code standards, audit logging, and reporting
- Validate automation through test runs and integrate database changes into CI/CD workflows
- 2. Backup, Recovery & Disaster Recovery
- Design and schedule daily, weekly, and monthly backups; configure and validate backup jobs in native DBMS tools
- Test recovery scenarios quarterly and maintain up-to-date DR runbooks
- Monitor backup job success/failure; escalate and remediate missed or failed backups
- Implement and validate log shipping, Always On Availability Groups, Oracle RAC or clustering for DR
- 3. Patch Management & Upgrades
- Plan and coordinate quarterly patch reviews and database version upgrades
- Apply vendor-released patches, perform post-upgrade validation, and resolve any functional regressions
- 4. Performance Monitoring, Tuning & Capacity Planning
- Monitor CPU, memory, storage usage, query performance, waits and latches using native and third-party tools
- Run baseline health checks post-maintenance, resolve anomalies, and document changes
- Forecast resource growth (1224 months), identify workload spikes, and plan infrastructure scaling
- Recommend and implement index and query optimizations in collaboration with application owners
- 5. High Availability & Replication
- Design, configure, and validate HA solutions including clustering, Always On, log shipping, and replication
- Monitor replication jobs, troubleshoot latency/errors, and conduct periodic failover drills
- 6. Security, Compliance & Audit
- Implement database security standards: encryption, data masking, role-based access control,
schema and user setup
- Approve or deny access requests, remove orphaned accounts, and audit user activity (failed logins, suspicious behavior)
- Generate compliance reports, track policy adherence, and escalate any violations
- 7. Incident Management & Root Cause Analysis
- Monitor for deadlocks, transaction failures, data corruption, and alert conditions
- Troubleshoot and resolve incidents in partnership with application teams or escalate to vendors
- Document root-cause analyses, corrective actions, and update runbooks
- 8. Reporting & Dashboarding
- Produce weekly performance and availability dashboards
- Provide insights and recommendations for tuning, optimization, and capacity upgrades
- 9. Scripting & Task Automation
- Develop and maintain scripts in Shell, PowerShell, PL/SQL, T-SQL, etc. and store them in version control
- Automate routine tasks such as patching, backups, user provisioning, and environment deployments
- Test scripts in lower environments and deploy to production via automated pipelines
Preferred Knowledge/Skills:
- Strong expertise with Terraform, Ansible, Git, Jenkins/GitLab CI or equivalent CI/CD tools
- Deep understanding of HA/DR architectures, backup/restore processes, and disaster recovery testing
- Proven ability to monitor, troubleshoot, and tune database performance at scale
- Familiarity with security best practices, compliance frameworks, and audit reporting
- Scripting proficiency (Shell, PowerShell, PL/SQL, T-SQL) and infrastructure-as-code principles
- Excellent communication, documentation, and collaboration skills
- Should possess hands-on knowledge about following tools -
- A. Oracle (i.e., Oracle RAC, Data Guard, Log Shipping)
- B. MySQL
- C. MariaDB
- D. MongoDB
- E. PostgreSQL
- F. PowerShell
- G. AWS RDS, Aurora, NoSQL
- H. Terraform
- I. Solarwinds SQL Diagnostics
- J. DBArtisan for Oracle DB Management
- K. Ansible
📌 Database Administrator (Gurugram)
🏢 PwC
📍 Gurugram