06 Aug
|
NST Cyber - Your Trusted Enterprise CTEM Partner
|
India
06 Aug
NST Cyber - Your Trusted Enterprise CTEM Partner
India
Lead DevOps Engineer -Cloud & Platform Infrastructure COMPANY:NST Cyber
LOCATION:Remote - India
EMPLOYMENT:Full time
EXPERIENCE:8+ years
PRIMARY CLOUD:AWS
ROLE TYPE:Hands-on leadership ROLE SUMMARY NST Cyber is seeking a hands-on Lead DevOps Engineer to architect, implement, operate and secure the infrastructure supporting its cybersecurity products. The role owns AWS, Amazon EKS, Terraform, CI/CD, GitOps, observability, reliability,
security and MongoDB Atlas. It also covers private-cloud environments, smaller or regional cloud providers, dedicated and virtual server platforms, and third-party Infrastructure as a Service (IaaS), Platform as a Service (PaaS) and Software as a
Service (SaaS) systems. The successful candidate will remain directly involved in design, automation, production troubleshooting and technical reviews while mentoring engineers and establishing platform standards.
KEY RESPONSIBILITIES AWS Cloud Architecture and Operations • Design and operate secure, highly available and cost-efficient AWS environments across networking, compute, storage, IAM, load balancing, encryption, DNS and monitoring.
- Implement least-privilege access, environment isolation, capacity planning, cost governance and strong operational controls.
- Diagnose and resolve complex infrastructure, networking, identity and access-control issues.
Amazon EKS and Kubernetes Platform
- Provision, upgrade and operate production-grade EKS clusters, add-ons, networking, ingress, storage, DNS and workload identity.
- Implement HPA-based workload scaling and Karpenter-based node provisioning, with appropriate on-demand capacity, isolation and disruption controls for critical workloads.
- Secure and troubleshoot Kubernetes using RBAC, Pod Security Admission, network policies, admission controls, secret management and strong deployment standards.
Broader
Infrastructure and Platform Operations • Operate and manage private-cloud platforms, smaller or regional cloud service providers, dedicated or virtual server environments,
and third-party IaaS, PaaS and SaaS systems.
- Establish consistent standards for access, configuration, patching, monitoring, backup, recovery, security, incident handling, vendor management and cost control across heterogeneous platforms.
- Integrate these environments with central identity, secrets management, logging, alerting, automation, governance and operational reporting wherever technically feasible.
Infrastructure as Code, CI/CD and GitOps
- Design reusable,
versioned and secure Terraform modules; manage remote state, locking, imports, drift, provider upgrades and controlled recovery.
- Build and maintain CI/CD pipelines using Jenkins, GitLab CI, GitHub Actions or equivalent, with GitOps deployments through
ArgoCD, Flux or a comparable platform.
- Implement automated validation, security scanning, plan review, environment promotion and safe rollback or progressive deployment practices.
Reliability, Security and Operations
- Implement metrics, logs and traces using Prometheus, Grafana, CloudWatch and OpenTelemetry; define actionable dashboards,
alerts and reliability objectives.
- Lead incident response, root-cause analysis, corrective actions, runbooks and automated remediation.
- Own backup validation, disaster-recovery planning and recovery exercises against agreed RTO and RPO targets.
- Establish secrets management, vulnerability scanning, WAF, audit logging and software-supply-chain controls.
MongoDB Atlas and Automation
- Administer MongoDB Atlas clusters covering private connectivity, access control, scaling, alerts, backups, restoration, upgrades and performance monitoring.
- Build operational automation and platform integrations using Python and Bash; use PowerShell where required.
Technical Leadership
- Review Terraform, Kubernetes, pipeline and platform changes; define reusable engineering standards and production-readiness expectations.
- Mentor engineers, strengthen technical ownership and translate platform risks into a prioritised roadmap.
- Partner with Engineering, Product, Security and external platform vendors on scalability, reliability, security, cost and service continuity.
REQUIRED SKILLS AND EXPERIENCE • Experience: 8+ years in DevOps, SRE, Cloud Engineering or Platform Engineering, including at least 4 years of substantial hands-on work with AWS, Kubernetes/EKS and Terraform.
- AWS: Deep production experience with VPC design, IAM, networking, load balancing, S3, CloudWatch, KMS, availability, security and cost optimisation.
- EKS/Kubernetes: Hands-on ownership of production clusters, upgrades, add-ons, ingress, storage, scheduling, HPA,
Karpenter,RBAC and workload security.
- Terraform: Advanced HCL, reusable module design, remote-state management, versioning, drift management, imports, code review and CI/CD integration.
- Platform Operations: Experience operating heterogeneous environments such as private cloud, non-hyperscaler CSPs,dedicated/virtual servers and IaaS, PaaS or SaaS platforms, including identity, monitoring, backup, security and vendor coordination.
- CI/CD and GitOps: Robust experience with at least one CI platform and hands-on use of ArgoCD, Flux or an equivalent GitOps approach.
- Observability and Reliability: Practical Prometheus, Grafana and CloudWatch experience, including production incidents, alerting,
capacity, backup and disaster recovery.
- Automation: Strong Python or Bash scripting capability for infrastructure and operational automation.
- MongoDB Atlas: Hands-on administration covering security, private connectivity, scaling, backup/restore, alerts and performance monitoring.
- Leadership: Demonstrated technical leadership, production ownership, infrastructure reviews and mentoring while remaining hands-on. GOOD TO HAVE • AWS Solutions Architect Professional, AWS DevOps Engineer Professional or CKA certification.
- OpenStack, VMware or equivalent private-cloud/virtualisation platforms; Azure, GCP, CloudFormation or Bicep exposure.
- Multi-account AWS, multi-cluster or multi-region operations, hybrid networking and internal developer platforms.
- HashiCorp Vault, OPA Gatekeeper, Kyverno, SIEM integration, policy-as-code or Zero Trust practices.
- Experience supporting ISO 27001, SOC 2 or comparable security-control frameworks.
- LangTrace or AI-assisted operational automation.
QUALIFICATIONS • Bachelor's degree in Computer Science, Engineering or a related discipline, or equivalent practical experience.
- Proven ownership of business-critical infrastructure with measurable improvements in reliability, recovery, deployment efficiency,
capacity or cost.
- Transparent written and verbal communication and experience working with distributed Agile engineering teams and external technology providers. WHY JOIN NST CYBER?
Build and operate the secure infrastructure behind a mission-critical cybersecurity product. You will have direct ownership of cloud, Kubernetes, automation, reliability and platform engineering standards across a diverse technology estate, with visible and measurable impact. NST Cyber is an equal-opportunity employer.
📌 Lead DevOps Engineer (India)
🏢 NST Cyber - Your Trusted Enterprise CTEM Partner
📍 India