Location:
Remote, Bangalore & Hyderabad area - to reach a nearby office occasionally)
Joining:
Immediate
About the engagement
You will be working as part of the Nestingale delivery team: : building and shipping
applied-AI systems — LLM/RAG and GenAI features — into production
. Hands-on engineering that takes models from prototype to reliable production service — not research, not advisory.
What you'll do
- Build LLM/RAG and GenAI features
end to end
and ship them (APIs, pipelines, monitoring)
- Work real data: ingestion, cleaning, embeddings, evaluation
- Improve what's live — latency, cost, accuracy, reliability
- Partner with product/eng to define and measure impact
Who you are
- 1–3 years hands-on in AI/ML (substantive internships count)
- Real
Python
+ a real ML framework (PyTorch / TensorFlow / scikit-learn) — not just calling APIs
- Hands-on with
LLMs / RAG
(LangChain/LangGraph, vector DBs, retrieval/prompt or fine-tuning)
- Think in data flows, latency, and evaluation; enjoy debugging why a model underperforms
Must have
Demonstrated
production
experience — you've shipped an AI/ML feature real users hit and can walk us through it in detail — with solid Python and hands-on LLM/RAG work.