Run and optimize LLM infrastructure, serving, and operational workflows.
Engineering
Full time
Location: Mumbai
Experience: 1-2 years
Key Requirements
- Experience operating LLM or ML inference systems
- Understanding of latency, cost, and quality trade-offs
- Familiarity with observability for AI systems
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.