Senior AI Platform / MLOps Engineer (Chennai)

Senior AI Platform / MLOps Engineer (Chennai)

26 Sep
|
Ampera Technologies
|
Chennai

26 Sep

Ampera Technologies

Chennai

Title : Senior AI Platform / MLOps Engineer

Experience : 6+ years

Work type : Chennai - Work from Office/other locations - Remote

Employment Type : Full Time

Notice Period : Immediate

Work Day :Mon to Fri

Key Responsibilities:

- Install, configure and operate OpenShift, NVIDIA GPU operator, OpenShift AI, and NIM microservices on 12× RTX PRO 6000 across two servers; single-node and HA control-plane topologies
- Serving configuration and tuning: quantized model deployment (FP8/FP4), replica balancing, batching, KV-cache and context management
- Azure GPU build environments: provisioning, cost control, parity with the on-prem stack via pinned container/model versions; cloud-to-factory migration with parity regression
- GitOps CI/CD, container registry, artifact/model versioning, environment promotion; observability and audit wiring (Splunk, Prometheus/Grafana)
- Benchmark automation: load harness, p50/p95/p99 latency, tokens/sec, GPU utilization; the capacity report data pipeline
- Platform upgrade procedure with evaluation-regression gates; deployment runbook as a first-class deliverable

Technical Skills:

- 6+ years infrastructure/platform engineering with 3+ years production Kubernetes; OpenShift experience strongly preferred
- Hands-on GPU inference serving in production: NIM, Triton, vLLM, or TensorRT-LLM — you have sized, deployed, and tuned LLM serving on real GPUs and can talk memory-bandwidth trade-offs
- GitOps fluency (ArgoCD/Flux), infrastructure-as-code,



container internals; comfortable in air-gapped/proxy-restricted enterprise networks
- Observability depth: metrics, traces, log pipelines; has built performance test harnesses, not just run them

- Azure or AWS GPU compute operations experience

Strongly preferred

- NVIDIA GPU operator and AI Enterprise stack specifics; KServe; Milvus or pgvector operations; VAST/NFS/S3 storage integration; banking or other regulated-environment delivery

About Ampera:

Ampera Technologies, a purpose driven Digital IT Services with primary focus on supporting our client with their Data, AI / ML, Accessibility and other Digital IT needs. We also ensure that equal opportunities are provided to Persons with Disabilities Talent. Ampera Technologies has its Global Headquarters in Chicago, USA and its Global Delivery Center is based out of Chennai, India. We are actively expanding our Tech Delivery team in Chennai and across India. We offer exciting benefits for our teams, such as 1) Hybrid and Remote work options available, 2) Opportunity to work directly with our Global Enterprise Clients, 3) Opportunity to learn and implement evolving Technologies, 4) Comprehensive healthcare, and 5) Conducive workplace for Persons with Disability Talent meeting Physical and Digital Accessibility standards

Skills:- Machine Learning (ML), MLOps, Kubernetes, Large Language Models (LLM), openshift and CI/CD

📌 Senior AI Platform / MLOps Engineer (Chennai)
🏢 Ampera Technologies
📍 Chennai

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior ai platform / mlops engineer (chennai) / chennai

Subscribe to this job alert:

Get the latest job offers by email for: senior ai platform / mlops engineer (chennai) / chennai