Senior Cloud Reliability Engineer (Thiruvananthapuram)

Senior Cloud Reliability Engineer (Thiruvananthapuram)

05 Oct
|
ThoughtSpot
|
Thiruvananthapuram

05 Oct

ThoughtSpot

Thiruvananthapuram

Senior Cloud Reliability Engineer
We are seeking a Senior Cloud Reliability Engineer with deep enterprise SaaS operations expertise to own the availability, reliability, security, and efficiency of our Multi-Cloud (AWS, GCP) production SaaS platform. The ideal candidate brings hands-on experience running highly available, large-scale Kubernetes-based control and data planes, a robust bias toward automation and AI-augmented operations, and a proven track record in production security, capacity management, and cloud-native data infrastructure.
Responsibilities
• Operate a high-scale, multi-cloud (AWS, GCP) SaaS platform — ensuring reliability, performance, and uptime for business-critical production workloads.
• Embed AI and Agentic workflows into SRE practice: leverage AI Ops platforms and LLM-powered autonomous agents for anomaly detection, automated triage, runbook execution, and incident summarization to reduce MTTR.
• Drive capacity planning and scaling operations — proactively model growth, right-size infrastructure,



and implement horizontal/vertical autoscaling strategies to support SaaS growth without reliability regression.
• Architect and operate Kubernetes controller frameworks governing both control plane and data plane services; define and enforce operational standards for cluster lifecycle, workload scheduling, autoscaling, and failover.
• Own operations of high-scale cloud-native databases and data infrastructure: PostgreSQL/RDS, DynamoDB, MySQL, Elasticsearch/OpenSearch, ElastiCache (Redis/Memcached) on AWS and GCP — including performance tuning, backup/recovery, and incident response.
• Lead incident response and blameless post-mortems for P0/P1 events; drive root cause analysis to permanent resolution and prevention — eliminating repeat incidents through systemic fixes, not workarounds.
• Define and enforce a culture of automation-first: identify and eliminate toil through self-healing systems, automated remediation pipelines, and infrastructure-as-c

📌 Senior Cloud Reliability Engineer (Thiruvananthapuram)
🏢 ThoughtSpot
📍 Thiruvananthapuram

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior cloud reliability engineer (thiruvananthapuram) / thiruvananthapuram

Subscribe to this job alert:

Get the latest job offers by email for: senior cloud reliability engineer (thiruvananthapuram) / thiruvananthapuram