SME Disaster Recovery Solution, Disaster Recovery (Noida)

SME Disaster Recovery Solution, Disaster Recovery (Noida)

05 Aug
|
HCL Technologies
|
Noida

05 Aug

HCL Technologies

Noida

SME - Disaster Recovery Solution, Disaster Recovery

Experience: 5 to Not Available years

Location: Noida, India

Skills: AWS, Azure, GCP, Terraform, Ansible, PowerShell, Python, Boto3, Veeam, Commvault, Zerto, Azure Site Recovery, ISO 27001, NIST, FFIEC, SOC 2, Microsoft Azure Administrator, Certified Business Continuity Professional (CBCP), Disaster Recovery Certified Specialist (DRCS)

Job Summary
As a Cloud Disaster Recovery Engineer, you will play a critical role in ensuring the resilience and continuity of our clients’ technology infrastructures. You will design, implement, and maintain robust disaster recovery (DR) and high availability (HA) solutions across on-premises, cloud, and hybrid environments. Your expertise will directly contribute to minimizing business disruption, enhancing system reliability, and supporting organizational objectives for technology resilience and risk mitigation.

Key Responsibilities
Design, implement, and maintain highly available (HA) and fault-tolerant architectures across cloud platforms (AWS, Azure, GCP) and on-premises environments.
• Develop and manage disaster recovery (DR) solutions in alignment with business-defined Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO).
• Automate disaster recovery and failover processes using Infrastructure as Code (IaC) and scripting tools such as Terraform, Ansible, PowerShell, and Python.
• Integrate resilience best practices into system design, deployment, and operations by collaborating with IT infrastructure and application teams.
• Conduct failure mode analysis (FMA) to proactively identify and address system vulnerabilities.
• Design and implement automated disaster recovery runbooks and one-click/event-triggered failover mechanisms using Ansible and Python.
• Develop and execute automated recovery verification, DR drills, and failover testing,



continuously identifying and closing gaps.
• Monitor and optimize system failover and recovery performance, ensuring swift response and recovery times.
• Lead technical response during disruptions or disaster recovery events, ensuring rapid restoration of services.
• Collaborate with cybersecurity teams to integrate DR solutions with cyber resilience strategies, including rapid restoration from ransomware or cyberattacks.
• Support post-incident analysis and recommend improvements to organizational resilience strategies.
• Implement and maintain resilience monitoring tools, ensuring ongoing system availability and DR readiness.
• Ensure compliance with industry standards and regulatory requirements (e.g., ISO 27001, NIST, FFIEC, SOC 2); provide technical input for audits and assessments.
• Generate comprehensive reports on resilience testing, failover performance, and risk mitigation activities.
• Collaborate with business continuity professionals and cloud architects to align resilience strategies with business needs.
• Provide training and technical guidance on DR best practices, system failover, and resilience processes; contribute to the development of technical documentation and playbooks.

Skill Requirements
Bachelor’s degree in Computer Science, Information Technology, Cybersecurity, or a related field.
• Minimum 5 years of experience in IT infrastructure, cloud engineering, disaster recovery, or resilience engineering.




• Proven expertise in disaster recovery planning and high-availability (HA) solution design.
• Advanced hands-on experience with cloud resilience strategies and cloud-native DR tools (AWS, Azure, GCP).
• Demonstrable experience automating AWS disaster recovery, including orchestration with Boto3 (EC2, RDS, S3, Route 53, IAM, AWS Backup, EDRS), cross-region replication, failover strategies, and DR patterns (pilot light, warm standby, multi-region active/active).
• Proficiency in automation engineering: enterprise-scale Ansible automation frameworks (roles, collections, dynamic inventories, AAP/AWX), Python automation (Boto3), and event-driven failover/recovery workflows.
• Experience implementing idempotent Infrastructure as Code (IaC) and integrating automation into CI/CD pipelines (e.g., GitHub Actions, Jenkins).
• In-depth knowledge of backup, replication, and data protection solutions (e.g., Veeam, Commvault, Zerto, Azure Site Recovery).
• Strong understanding of networking, storage, virtualization, and hybrid-cloud architectures.
• Extensive coding and scripting expertise for cloud resource automation, with a focus on AWS and multi-cloud environments.
• Certifications such as AWS Certified Solutions Architect, Red Hat Certified Engineer,
Other Requirements
Microsoft Azure Administrator, Certified Business Continuity Professional (CBCP), or Disaster Recovery Certified Specialist (DRCS) are highly desirable.
• Familiarity with regulatory frameworks and audit processes for technology resilience.
• Experience working in large-scale, multinational environments.
• Excellent communication, collaboration, and technical documentation skills.
• Ability to mentor and train IT staff in DR best practices and current technologies.

📌 SME Disaster Recovery Solution, Disaster Recovery (Noida)
🏢 HCL Technologies
📍 Noida

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: sme disaster recovery solution, disaster recovery (noida) / noida

Subscribe to this job alert:

Get the latest job offers by email for: sme disaster recovery solution, disaster recovery (noida) / noida