Youll work at the bleeding edge of generative AI - building RAG pipelines, agentic workflows, and LLM-powered products that are already shaping industries. This is a high-impact role with direct ownership over production AI systems.
Responsibilities
- Build and maintain RAG (Retrieval-Augmented Generation) pipelines for enterprise clients
- Design multi-agent orchestration systems using frameworks like LangChain or CrewAI
- Integrate LLM APIs (OpenAI, Anthropic, Mistral) into scalable backend services
- Implement prompt engineering strategies and evaluation frameworks
- Work closely with clients to translate business requirements into AI solutions
- Instrument and monitor LLM systems for quality, latency, and cost
Requirements
- 3+ years of software engineering experience with 1+ years in LLM/GenAI
- Hands-on experience with LangChain,
LlamaIndex, or similar frameworks
- Robust Python and API integration skills
- Understanding of vector databases (Pinecone, Weaviate, pgvector)
- Comfortable working across the full stack from model to UI
Nice to Have
- Experience with open-source model deployment (Ollama, vLLM, TGI)
- Knowledge of fine-tuning workflows (LoRA, QLoRA, PEFT)
- Prior startup or consulting experience
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.