We are looking for a hands-on AI Engineer experienced in building and deploying production-grade Generative AI, LLM, RAG and Agentic AI solutions.
Key Responsibilities:
Build and deploy LLM/GenAI applications and AI agents.
Develop RAG pipelines including chunking, embeddings, vector search, hybrid retrieval and reranking.
Design multi-agent workflows using LangChain/LangGraph or similar frameworks.
Develop scalable Python/FastAPI backend services and REST APIs.
Work with Azure/AWS/GCP and Docker/Kubernetes for deployment.
Implement LLM evaluation, monitoring, guardrails and hallucination reduction.
Optimize AI applications for accuracy, latency and cost.
Take end-to-end ownership from development to production deployment.