14 Aug
|
Dhruv Technology Solutions
|
Bengaluru
14 Aug
Dhruv Technology Solutions
Bengaluru
About the Role
We are seeking a Database Reliability Engineer to join our team. In this critical role, you will bridge the gap between traditional database administration and site reliability engineering. You will be responsible for the architecture, deployment, performance tuning, and extreme high availability of our database infrastructure—treating operational excellence as a core feature of the system.
You will manage a hybrid environment consisting of self-hosted Microsoft SQL Server instances on
Windows/EC2 and fully managed AWS services (RDS and Aurora). We are looking for an engineer who combines deep database knowledge with strong infrastructure automation skills (IaC) to drive reliability, security, and efficiency across our cloud networking and infrastructure.
Key Responsibilities
Database Administration & Performance Optimization
Administer self-hosted Microsoft SQL Server Always On Clusters.
Execute advanced SQL concepts, including server-level DDL/DML operations, query performance tuning, and proactive maintenance to prevent outages.
High Availability & Disaster Recovery (HA/DR)
Design, plan, and execute robust HA/DR strategies across RDS, Aurora, and self-hosted environments.
Manage switchovers, failovers, and disaster recovery rebuilds with minimal recovery time objectives (RTO/RPO).
Backup & Recovery
Implement and oversee bulletproof backup/recovery strategies to ensure data integrity and business continuity.
Automation & Infrastructure as Code (IaC)
Develop and maintain automation scripts utilizing PowerShell and Python.
Write, review,
and maintain Terraform scripts for cloud infrastructure provisioning (EC2, VPC, S3,
etc.).
Security & Access Management
Implement and manage database authentication and access controls utilizing Active Directory
(AD) and AWS IAM.
Observability & Documentation
Author and maintain comprehensive technical documentation, Standard Operating Procedures
(SOPs), and operational runbooks. Establish SLOs/SLIs for database performance and uptime.
Baseline Requirements
Experience: 3 to 5 years in Database Engineering, Site Reliability Engineering, or a similar hybrid operations role.
Database Expertise: Deep understanding of self-hosted Microsoft SQL Server Always On Clusters in Windows environments.
Disaster Recovery: Proven experience with advanced DR concepts, including switchover, failover,
and restoration activities.
AWS & Technical Competencies
AWS Ecosystem: Robust working knowledge of RDS, Aurora, EC2, VPC, S3, Lambda, Step
Functions, Systems Manager (SSM), Route 53, CloudWatch, and IAM.
Scripting & Automation: Strong proficiency in PowerShell and Python.
IaC: Solid experience writing and maintaining Terraform scripts.
Security: Solid grasp of Active Directory (AD)-based authentication and access management concepts.
Preferred Qualifications
Proven track record of supporting mission-critical production database environments.
Hands-on exposure to broader automation and Infrastructure-as-Code (IaC) practices.
Excellent troubleshooting and analytical problem-solving skills under pressure.
Strong communication skills with the ability to collaborate across infrastructure, application, and cloud teams.
📌 Database Engineer (Bengaluru)
🏢 Dhruv Technology Solutions
📍 Bengaluru