Job Description
About the Role:
n
We are seeking a highly skilled AWS Data Platform Engineer (Exp: 3 - 6 yrs) to manage, operate, and optimize enterprise-scale AWS data platforms across production environments. The role focuses on ensuring reliability, scalability, security, backup compliance, and operational excellence across relational databases, object storage, data lakes, and analytics services.
n
The ideal candidate should possess robust operational expertise in Amazon RDS/Aurora, Amazon S3, AWS Backup, and AWS-native analytics platforms, with hands-on experience in performance optimization, disaster recovery, governance, monitoring, and automation.
n
n
Key Responsibilities:
n
AWS Database Platform Operation
n
n
- Manage and operate Amazon RDS and Aurora database environments across development, staging, and production
n
- Configure and maintain Multi-AZ deployments, read replicas, replication strategies, and failover mechanisms
n
- Perform database patching, upgrades, maintenance scheduling, and parameter tuning
n
- Monitor database performance, replication lag, storage utilization, and query efficiency
n
- Troubleshoot SQL performance bottlenecks, deadlocks, connectivity issues, and operational failures
n
- Implement backup, recovery, snapshot management, and Point-in-Time Recovery (PITR) strategies
n
- Conduct database restoration validation and disaster recovery drills
n
- Support schema migration and production deployment activities
n
- Perform capacity planning for compute, storage, IOPS, and throughput requirements
n
n
n
AWS Storage & Data Lake Operations
n
n
- Manage enterprise-scale Amazon S3 storage environments for application, analytics, and archival workloads
n
- Configure bucket policies, lifecycle policies, replication, versioning, and archival strategies
n
- Implement storage tiering and archival workflows using Glacier and Intelligent Tiering
n
- Monitor storage growth, access patterns, and optimize storage costs
n
- Configure and manage cross-region replication (CRR) and same-region replication (SRR)
n
- Ensure secure access management using IAM policies, bucket policies, VPC endpoints, and encryption controls
n
- Support operational management of data lake environments using AWS Glue, Athena, and Lake Formation
n
n
Backup, Recovery & Disaster Recover
n
n
- Configure and manage AWS Backup policies across databases, storage, and related services
n
- Implement centralized backup governance and retention policies
n
- Monitor backup health, recovery points, and backup job success/failure
n
- Design and validate disaster recovery procedures aligned to defined RPO/RTO objectives
n
- Conduct periodic recovery and restoration testing
n
- Manage cross-account and cross-region backup strategies for business continuity
n
n
Monitoring, Reliability & Operations
n
n
- Configure monitoring, alerting, and observability using CloudWatch, OpenSearch, and related AWS services
n
- Develop operational dashboards and automated alerting for critical infrastructure metricss
n
- Ensure high availability, scalability, reliability, and operational stability of data platform
n
- Troubleshoot production incidents, storage failures, pipeline disruptions, and service degradation
n
- Participate in on-call support and incident management processes
n
- Create operational runbooks, SOPs, and recovery procedures
n
n
Security, Governance & Compliance
n
n
- Implement encryption at rest and in transit using AWS KMS and AWS-native security controls
n
- Manage secrets, credentials, and database authentication mechanisms
n
- Enforce governance, access control, backup retention, and compliance standards
n
- Support audit readiness and operational compliance reporting
n
n
Automation & Cost Optimization
n
n
- Automate infrastructure provisioning and operational workflows using Terraform, CloudFormation, Lambda, or scripting too
n
- Optimize database and storage costs through lifecycle management, tiering, reserved capacity, and utilization analysis
n
- Continuously improve operational efficiency, monitoring coverage, and platform reliability
n
n
Required Skills & Experience:
n
Must-Have AWS Services
n
n
- Amazon RDS / Aurora
n
- Amazon S3
n
- AWS Backup
n
- Amazon DynamoDB
n
- Amazon CloudWatch
n
- AWS IAM
n
- AWS KMS
n
n
Core Technical Skills:
n
n
- Strong understanding of relational database concepts, backup/recovery, replication, and high availability
n
- Experience with S3 lifecycle management, replication, archival, and storage optimization
n
- Hands-on experience in disaster recovery planning and restoration testing for databases
n
- Experience troubleshooting database and storage performance issues
n
- Knowledge of monitoring, alerting, and operational observability
n
- Familiarity with Infrastructure as Code (Terraform or CloudFormation)
n
- Understanding of security best practices and governance controls in AWS
n
n
n
Note: This is not a Data Engineer position. Candidates with primarily Data Engineer/AWS Data Engineer experience are requested not to apply.
📌 Cloud Database Administrator (Noida)
🏢 Zarthi
📍 Noida