Site Reliability Engineer (Noida)

Site Reliability Engineer (Noida)

11 Sep
|
Infitoo Systems
|
Noida

11 Sep

Infitoo Systems

Noida

Full time role with Client (no third party)

Role : SRE (Site Reliability Engineer)

Location: Noida

Exp-5 to 10 years

Employment type- Full time (Onsite)

Note- Immediate joiners preferred

Reliability & Operations (SRE Core)

- Own the availability, latency, performance, and capacity of production workloads.
- Define and track SLIs, SLOs, error budgets, and reliability KPIs.
- Lead incident management, including on-call support, triage, root cause analysis (RCA), postmortems, and preventive actions.
- Improve Mean Time to Recovery (MTTR) through automation, runbooks, and self-healing mechanisms.

Kubernetes & Platform Engineering (EKS)

- Operate and continuously improve AWS EKS clusters across multiple namespaces and environments.
- Manage deployments using Helm and Kustomize, enabling protected rollout strategies such as Blue/Green and Canary deployments.
- Configure and manage:
- AWS ALB/NLB Ingress
- Service Discovery
- Horizontal Pod Autoscaler (HPA)
- Vertical Pod Autoscaler (VPA)
- Cluster Autoscaler
- Enforce Kubernetes security using:
- RBAC
- Network Policies
- Pod Security Standards
- Secrets Management





CI/CD & Release Engineering

- Build and maintain CI/CD pipelines using GitHub Actions, Jenkins, or GitLab CI (based on organizational standards).
- Promote an "Everything as Code" approach by storing infrastructure, application configurations, and deployment manifests in Git.

Infrastructure as Code (IaC) & Cloud Automation

- Provision and manage cloud infrastructure using Terraform, AWS CloudFormation, or AWS CDK.
- Manage multi-account AWS environments using IAM roles, federation, and AWS SSO.
- Implement cloud cost optimization practices, including:
- Resource tagging standards
- Budget alerts
- Right-sizing recommendations
- Savings Plans guidance

Observability (Monitoring, Logging & Tracing):

- CloudWatch, ELK, or OpenSearch.
- Enable distributed tracing using OpenTelemetry, AWS X-Ray, or Jaeger, where applicable.

Security & Governance (DevSecOps):

- Implement application and edge security using AWS WAF and CloudFront, ensuring controlled and secure rollouts.

📌 Site Reliability Engineer (Noida)
🏢 Infitoo Systems
📍 Noida

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: site reliability engineer (noida) / noida

Subscribe to this job alert:

Get the latest job offers by email for: site reliability engineer (noida) / noida