At OneSpan, we specialize in digital identity and anti-fraud solutions that create exceptional and secure experiences.
We are looking for a Site Reliability Engineer to join our growing platform team in Delhi NCR. You will help build and maintain highly available, scalable, and observable systems. You will bridge the gap between development and operations, embedding reliability into every layer of our stack.
RESPONSIBILITIES
• Design, build, and maintain CI/CD pipelines and deployment automation using GitLab CI (experience with GitHub Actions is also valued).
• Monitor production systems using observability stacks (Logz.io, Grafana, etc.) and respond to incidents as part of an on-call rotation (one weekend per month plus weekday daytime coverage in a follow-the-sun model with the team in Canada).
• Define and track SLIs, SLOs, and error budgets in collaboration with product and engineering teams.
• Automate toil: identify repetitive manual tasks and eliminate them through scripting and tooling (Python, Bash, Go).
• Manage cloud infrastructure on AWS/Azure using IaC tools such as Terraform.
• Participate in blameless post-mortems and drive root cause analysis for production incidents.
• Collaborate with software engineers to review system designs for reliability, scalability, and fault tolerance.
REQUIREMENTS
• 5-7 years of experience in SRE, DevOps, or a related infrastructure/platform engineering role.
• Proficiency in at least one scripting/programming language (Python, Java, Go, or Bash).
• Hands-on experience running workloads on Amazon EKS in production, including node groups, IAM roles for service accounts (IRSA), and cluster upgrades.
• Working knowledge of Istio service mesh: traffic management, mTLS, VirtualServices, DestinationRules, and sidecar injection.
• Experience deploying and configuring the OpenTelemetry (OTEL) Collector — pipelines, receivers, exporters, and processors for metrics, logs, and traces.
• Solid experience with AWS core
📌 Site Reliability Engineer (Noida)
🏢 OneSpan
📍 Noida