11 Sep
|
Persistent
|
Bengaluru
11 Sep
Persistent
Bengaluru
Hands-on experience in Linux administration, AWS cloud services, Datacenter operations, Networking, and Observability tools.
The role involves monitoring, maintaining, and improving the reliability and performance of our cloud and on-premise infrastructure.
Key Responsibilities Cloud & System Operations Manage and support cloud infrastructure primarily on AWS (EC2, S3, IAM, VPC, CloudWatch, EKS, RDS, etc.).
Deploy, maintain, and troubleshoot Linux-based systems in cloud and datacenter environments.
Ensure availability, performance, and scalability of critical applications and infrastructure.
Monitoring & Observability Configure and maintain monitoring systems using observability tools such as Prometheus, Grafana, CloudWatch, ELK/EFK stack, Datadog, Recent Relic, etc.
Set up alerts, dashboards, and health checks to proactively identify issues. Perform root cause analysis (RCA)
for system and application incidents.
Networking & Security Troubleshoot and manage network components including DNS, DHCP, VPN, VLANs, routing, firewalls, and load balancers.
Work with cloud networking services such as AWS VPC, Security Groups, NACLs, Route53, etc.
Implement system hardening and follow security best practices.
Datacenter & Infrastructure Support Provide hands-on support for datacenter activities such as server deployment, hardware maintenance, cabling, racking, and troubleshooting.
Support virtualization platforms (e.g., VMware, KVM) if required.
Operational Excellence Participate in 24/7 rotational on-call support for critical systems. Document procedures, configurations, and incident reports.
📌 Observability Engineer (Bengaluru)
🏢 Persistent
📍 Bengaluru