10 Sep
|
NST Cyber - Your Trusted Enterprise CTEM Partner
|
Agra
10 Sep
NST Cyber - Your Trusted Enterprise CTEM Partner
Agra
Lead DevOps Engineer -Cloud & Platform Infrastructure
COMPANY:NST Cyber
LOCATION:Remote - India
EMPLOYMENT:Full-time
EXPERIENCE:8+ years
PRIMARY CLOUD:AWS
ROLE TYPE:Hands-on leadership
ROLE SUMMARY
NST Cyber is seeking a hands-on Lead DevOps Engineer to architect, implement, operate and secure the infrastructure
supporting its cybersecurity products. The role owns AWS, Amazon EKS, Terraform, CI/CD, GitOps, observability, reliability,
security and MongoDB Atlas. It also covers private-cloud environments, smaller or regional cloud providers, dedicated and
virtual server platforms, and third-party Infrastructure as a Service (IaaS), Platform as a Service (PaaS) and Software as a
Service (SaaS) systems. The successful candidate will remain directly involved in design, automation, production
troubleshooting and technical reviews while mentoring engineers and establishing platform standards.
KEY RESPONSIBILITIES
AWS Cloud Architecture and Operations
• Design and operate secure, highly available and cost-efficient AWS environments across networking, compute, storage, IAM, load
balancing, encryption, DNS and monitoring.
• Implement least-privilege access, environment isolation, capacity planning, cost governance and strong operational controls.
• Diagnose and resolve complex infrastructure, networking, identity and access-control issues.
Amazon EKS and Kubernetes Platform
• Provision, upgrade and operate production-grade EKS clusters, add-ons, networking, ingress, storage, DNS and workload identity.
• Implement HPA-based workload scaling and Karpenter-based node provisioning, with appropriate on-demand capacity, isolation
and disruption controls for critical workloads.
• Secure and troubleshoot Kubernetes using RBAC, Pod Security Admission, network policies, admission controls, secret management and strong deployment standards.
Broader Infrastructure and Platform Operations
• Operate and manage private-cloud platforms, smaller or regional cloud service providers, dedicated or virtual server environments,
and third-party IaaS, PaaS and SaaS systems.
• Establish consistent standards for access, configuration, patching, monitoring, backup, recovery, security, incident handling, vendor
management and cost control across heterogeneous platforms.
• Integrate these environments with central identity, secrets management, logging, alerting, automation, governance and operational
reporting wherever technically feasible.
Infrastructure as Code, CI/CD and GitOps
• Design reusable,
versioned and secure Terraform modules; manage remote state, locking, imports, drift, provider upgrades and
controlled recovery.
• Build and maintain CI/CD pipelines using Jenkins, GitLab CI, GitHub Actions or equivalent, with GitOps deployments through
ArgoCD, Flux or a comparable platform.
• Implement automated validation, security scanning, plan review, environment promotion and safe rollback or progressive
deployment practices.
Reliability, Security and Operations
• Implement metrics, logs and traces using Prometheus, Grafana, CloudWatch and OpenTelemetry; define actionable dashboards,
alerts and reliability objectives.
• Lead incident response, root-cause analysis, corrective actions, runbooks and automated remediation.
• Own backup validation, disaster-recovery planning and recovery exercises against agreed RTO and RPO targets.
• Establish secrets management, vulnerability scanning, WAF, audit logging and software-supply-chain controls.
MongoDB Atlas and Automation
• Administer MongoDB Atlas clusters covering private connectivity, access control, scaling, alerts, backups, restoration, upgrades and
performance monitoring.
• Build operational automation and platform integrations using Python and Bash; use PowerShell where required.
Technical Leadership
• Review Terraform, Kubernetes, pipeline and platform changes; define reusable engineering standards and production-readiness
expectations.
• Mentor engineers, strengthen technical ownership and translate platform risks into a prioritised roadmap.
• Partner with Engineering, Product, Security and external platform vendors on scalability, reliability, security, cost and service
continuity.
REQUIRED SKILLS AND EXPERIENCE
• Experience: 8+ years in DevOps, SRE, Cloud Engineering or Platform Engineering, including at least 4 years of substantial hands-on work with AWS, Kubernetes/EKS and Terraform.
• AWS: Deep production experience with VPC design, IAM, networking, load balancing, S3, CloudWatch, KMS, availability, security and cost optimisation.
• EKS/Kubernetes: Hands-on ownership of production clusters, upgrades, add-ons, ingress, storage, scheduling, HPA,
Karpenter,RBAC and workload security.
• Terraform: Advanced HCL, reusable module design, remote-state management, versioning, drift management, imports, code
review and CI/CD integration.
• Platform Operations: Experience operating heterogeneous environments such as private cloud, non-hyperscaler CSPs,dedicated/virtual servers and IaaS, PaaS or SaaS platforms, including identity, monitoring, backup, security and vendor coordination.
• CI/CD and GitOps: Strong experience with at least one CI platform and hands-on use of ArgoCD, Flux or an equivalent GitOps
approach.
• Observability and Reliability: Practical Prometheus, Grafana and CloudWatch experience, including production incidents, alerting,
capacity, backup and disaster recovery.
• Automation: Robust Python or Bash scripting capability for infrastructure and operational automation.
• MongoDB Atlas: Hands-on administration covering security, private connectivity, scaling, backup/restore, alerts and performance monitoring.
• Leadership: Demonstrated technical leadership, production ownership, infrastructure reviews and mentoring while remaining hands-on.
GOOD TO HAVE
• AWS Solutions Architect Professional, AWS DevOps Engineer Professional or CKA certification.
• OpenStack, VMware or equivalent private-cloud/virtualisation platforms; Azure, GCP, CloudFormation or Bicep exposure.
• Multi-account AWS, multi-cluster or multi-region operations, hybrid networking and internal developer platforms.
• HashiCorp Vault, OPA Gatekeeper, Kyverno, SIEM integration, policy-as-code or Zero Trust practices.
• Experience supporting ISO 27001, SOC 2 or comparable security-control frameworks.
• LangTrace or AI-assisted operational automation.
QUALIFICATIONS
• Bachelor's degree in Computer Science, Engineering or a related discipline, or equivalent practical experience.
• Proven ownership of business-critical infrastructure with measurable improvements in reliability, recovery, deployment efficiency,
capacity or cost.
• Clear written and verbal communication and experience working with distributed Agile engineering teams and external technology
providers.
WHY JOIN NST CYBER?
Build and operate the secure infrastructure behind a mission-critical cybersecurity product. You will have direct ownership of cloud, Kubernetes, automation, reliability and platform engineering standards across a diverse technology estate, with visible and measurable impact.
NST Cyber is an equal-opportunity employer.
📌 Lead DevOps Engineer (Agra)
🏢 NST Cyber - Your Trusted Enterprise CTEM Partner
📍 Agra