Dev Ops Engineer Kubernetes and Cloud Infrastructure
Role Summary
We are seeking an experienced Dev Ops Engineer to manage and scale enterprise Kubernetes platforms across AWS and hybrid infrastructure. In this critical role, you will enable secure automated and compliant delivery of mission critical banking applications. You will design and operate highly available Kubernetes clusters implement robust disaster recovery solutions and drive infrastructure automation using modern Dev Ops practices and tools.
Key Responsibilities
Kubernetes Platform Operations Operate manage and scale Rancher managed Kubernetes clusters RKE2 and EKS in production environments
High Availability and Architecture Design and implement High Availability HA architectures across multi cluster Kubernetes environments
Backup and Disaster Recovery Establish manage and validate Kubernetes backup and recovery solutions using tools such as Portworx Longhorn Velero or Kasten
Disaster Recovery Strategy Own and validate comprehensive DR strategies including failover procedures testing protocols and recovery time objectives RTO and RPO
Infrastructure as Code Drive Infrastructure as Code Terraform implementations for standardized auditable and reproducible infrastructure builds
CI and CD and Git Ops Implement optimize and maintain CI and CD pipelines and Git Ops workflows ArgoCD Flux for automated application delivery
Platform Reliability Ensure platform resiliency availability and performance against defined Service Level Agreements SLAs
Security and Compliance Enforce security controls compliance standards and audit requirements including IAM policies RBAC secrets management and regulatory frameworks
Incident Management Support incident response root cause analysis RCA and maintain regulatory readiness for banking and compliance audits
Monitoring and Observability Deploy and optimize monitoring logging and alerting solutions Grafana Splunk Prometheus to ensure visibility and troubleshooting capability
Networking and Security Configure and maintain secure networking VPC management and implement security best practices across cloud and on premises infrastructure
Core Requirements
Experience 7 years of hands on experience in Kubernetes Dev Ops or platform engineering roles
Education Bachelor's degree in Computer Science Information Technology or related field
Kubernetes Expertise Deep knowledge of Kubernetes architecture cluster management and production grade operations
Rancher and Container Orchestration Proven experience managing Rancher based Kubernetes clusters RKE2 and EKS
AWS Cloud Proficiency Strong expertise in AWS services including EKS VPC IAM EC2 and related infrastructure components
High Availability and Disaster Recovery Demonstrated experience designing HA backup and DR solutions for Kubernetes environments
Infrastructure Automation Hands on proficiency with Terraform for infrastructure provisioning and automation
CI and CD Engineering Proven experience designing and implementing CI and CD pipelines and deployment automation
Production Critical Environments Experience supporting mission critical regulated or high availability production systems
Security and Networking Deep knowledge of Kubernetes security networking RBAC IAM and compliance controls
Observability Hands on experience with monitoring logging and troubleshooting tools Grafana Prometheus Splunk or similar
Mandatory Skills Dev Ops Kubernetes AWS EKS VPC IAM Terraform On Premises Infrastructure and Containerization
Preferred Qualifications
Banking and Regulated Environment Experience Prior experience working in banking financial services or other heavily regulated industries
Compliance and Audit Familiarity with DR drills audit controls compliance frameworks e.g. SOX PCI DSS regulatory standards
Advanced Backup Tools Hands on experience with Velero Kasten or Portworx for enterprise backup and recovery
Git Ops Platforms Experience with ArgoCD or Flux for declarative application delivery
Kubernetes Ecosystem Proficiency with Helm Prometheus and advanced Kubernetes tooling
Hybrid Infrastructure Experience managing Kubernetes across hybrid environments on premises and cloud
AWS Certified Solutions Architect Professional or AWS Certified Dev Ops Engineer
Hashi Corp Certified Terraform Associate
Work Preferences
Work Location Onsite 5 days per week full time office presence required
Shift Night Shift
Note Candidate acknowledgment email confirming acceptance of night shift and onsite requirements is mandatory
Candidate Profile
We are currently seeking Male candidates for this position.
What We Value
- Strong problem-solving and troubleshooting mindset
- Ability to work independently and as part of a collaborative team
- Excellent communication skills and documentation practices
- Proactive approach to learning and staying current with Dev Ops trends
- Commitment to operational excellence, security, and regulatory compliance
- Experience in fast-paced, mission-critical environments
Value Proposition
This role offers the chance to design, build, and operate a highly available, resilient, and compliant Kubernetes platform with enterprise-grade backup and disaster recovery capabilities. You will work on mission-critical banking applications, ensuring alignment with regulatory standards and operational excellence. Your expertise will directly impact the reliability and security of our platform infrastructure, supporting thousands of users and critical business operations.
Equal Opportunity Statement
Our organization is committed to creating an inclusive workplace. We welcome applications from qualified candidates and provide equal opportunities regardless of protected characteristics, in accordance with applicable laws.