Manager-Site Reliability Engineering (Kochi)

Manager-Site Reliability Engineering (Kochi)

01 Sep
|
Practicesuite
|
Kochi

01 Sep

Practicesuite

Kochi

About the Role

Opportunity to build the first Site Reliability Engineering function from the ground up at a company modernizing how it delivers healthcare technology. You won't be joining an established team and learning its playbook — you'll be writing it. You'll define what reliability means for our systems, build the practices and tooling that make production stable and predictable, and hire and shape the very first engineer on your team.

This role offers technical ownership and people leadership early in a function's life, direct visibility with senior technical leadership, and the opportunity to build lasting influence over how a growing organization runs its most critical systems.

Responsibilities

- Own on-call and incidents: US-hours coverage, severity, incident command, write-ups, and follow-through
- Drive production reliability: availability, latency, and tested backup/restore practices across current and emerging platforms
- Reduce operational toil by converting recurring manual work into monitoring and automation, in partnership with your team
- Implement strong operational standards: logging, access control, deployment evidence, and restore testing
- Build the team: hire the first SRE individual contributor within 90 days of your start

Requirements

- 2+ managing or team lead responsibilities for site reliability engineering and/ or similar production operations (DevOps) roles




- Strong core Linux fundamentals — process/memory management, filesystems, networking, shell scripting, and troubleshooting in production environments
- Experience with monitoring/observability tooling (e.g., Greylog, Prometheus/Grafana) and CI/CD pipelines
- Database experience — Oracle strongly preferred, Postgres a plus
- Experience with virtualization platforms — Proxmox VE experience strongly preferred; general hypervisor/VM management concepts (KVM, VMware, or similar) a plus
- Working knowledge of AWS core services (EC2, VPC, IAM, S3, RDS, CloudWatch) and cloud infrastructure concepts, with the ability to reason about cost, scaling, and resilience trade-offs
- Hands-on Docker experience — building, running, and troubleshooting containerized services in production; container orchestration (e.g., ECS, Kubernetes) a plus
- Understanding of security standards and practices (access control/IAM, encryption at rest and in transit, patching cadence, vulnerability management) familiarity with frameworks like HIPAA, SOC 2, or ISO 27001 a plus given the regulated setting
- Ability to run effective collaboration across time zones with US-based stakeholders

Location: Kochi, India (US-hours overlap required) Reports to: Director of Infrastructure & Reliability

Department: Engineering & System Operations

📌 Manager-Site Reliability Engineering (Kochi)
🏢 Practicesuite
📍 Kochi

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: manager-site reliability engineering (kochi) / kochi

Subscribe to this job alert:

Get the latest job offers by email for: manager-site reliability engineering (kochi) / kochi