Job Summary: Design and deploy production-grade AI/ML and Generative AI solutions, with focus on scalable model deployment and enterprise AI pipelines.
Key Responsibilities:
●Design robust enterprise Generative AI (GenAI) pipelines.
●Deploy and optimize machine-learning models in high-throughput production settings.
●Develop Large Language Model (LLM) fine-tuning and Retrieval-Augmented Generation (RAG) architectures.
●Optimize model inference latency and compute efficiency.
●Design and deploy scalable vector-database infrastructure.
●Build, standardize, and improve automated MLOps pipelines.