27 Sep
|
ZenoCloud
|
Noida
Location: Noida
Job Type: Full time
Industry: Web Hosting / Cloud Infrastructure / SaaS
About the Role
We are looking for an experienced and self-driven Senior Cloud Support Engineer to maintain, optimize, and scale production infrastructure for our global enterprise clients. In this role, you will act as a senior technical escalation point, driving infrastructure reliability, troubleshooting complex distributed system issues, and executing cloud migrations.
The ideal candidate brings deep operational knowledge of Linux internals, production AWS architectures, and practical automation skills to minimize downtime and eliminate manual operational overhead.
Key Responsibilities
- L2/L3 Escalation & Troubleshooting: Act as the primary technical point of escalation for complex server outages, performance bottlenecks, high I/O wait, memory leaks, kernel panics, and network routing degradation.
- Linux System Administration: Build, optimize, and administer heterogeneous Linux environments (Ubuntu, CentOS/Rocky/RHEL, Debian) running high-traffic enterprise workloads.
- Stack Tuning & Architecture: Deploy, secure, and tune high-performance LAMP/LEMP stacks (Nginx, Apache, PHP-FPM, MySQL/MariaDB/Percona), configuring reverse proxies, caching layers, and database query optimization.
- Cloud Infrastructure Management: Provision, architect, and manage multi-region AWS environments utilizing EC2, VPC, ALB/NLB, Auto Scaling groups, S3, RDS, and CloudFront.
- Automation & Configuration Management: Write reusable Bash/Python scripts and Ansible playbooks to automate recurring maintenance, deployment pipelines, patching, and configuration drift checks.
- Proactive Observability & Incident Response:
Maintain monitoring and APM stacks (Zabbix, Prometheus, Grafana, CloudWatch), fine-tune alerting thresholds to eliminate noise, and lead Root Cause Analysis (RCA) investigations.
- Security & Compliance: Implement OS-level hardening (CIS benchmarks, kernel tuning), enforce firewall policies (iptables, nftables, firewalld), manage SSL/TLS lifecycles, and coordinate vulnerability remediation.
- Control Panel & Web Hosting Management: Manage advanced configurations, account migrations, and debugging across major control panels (cPanel/WHM, Plesk, Webmin).
- Client Advisory & Escalated Support: Provide high-touch support for critical client incidents, collaborating directly via ticketing, chat, and high-priority incident calls while mentoring junior system administrators.
Candidate Requirements
- 3–5 years of dedicated experience as a Linux System Administrator, Cloud Support Engineer, or Infrastructure Operations Engineer in a 24×7 production or SaaS environment.
- Demonstrated track record diagnosing root causes across the full operating system, web stack, database, and network boundary.
- Hands-on experience architecting and managing production workloads in AWS using best-practice security and reliability frameworks.
- Experience authoring structured Root Cause Analysis (RCA) reports and system runbooks.
- Strong written and verbal communication skills, with experience managing escalations directly with technical and non-technical stakeholders.
- Flexibility to support rotational shifts or critical incident on-call rotations when required.
- Preferred Certifications (Bonus): AWS Certified Solutions Architect / SysOps Administrator, RHCSA, or RHCE.
📌 Cloud Support Engineer (Noida)
🏢 ZenoCloud
📍 Noida