Experience: 3-6 years
Job Location: Noida, Gurugram, Bengaluru
Notice Period: 30 Days or Immediate Joiner
We are seeking an experienced ML Engineer with robust expertise in Machine Learning, Large Language Models (LLMs), Generative AI, MLOps, and Cloud Technologies. The ideal candidate will be responsible for designing, developing, deploying, and optimizing scalable AI/ML solutions, including LLM-powered applications, RAG systems, and enterprise-grade data pipelines.
Key Responsibilities
- Design, develop, and deploy ML and LLM-based applications in production environments.
- Build and optimize RAG pipelines, semantic search solutions, and AI-powered chatbots.
- Develop scalable microservices and APIs using Python, FastAPI, REST, and gRPC.
- Create and maintain robust CI/CD pipelines and MLOps workflows.
- Design enterprise-grade data pipelines using Spark and cloud platforms.
- Implement model serving, monitoring, optimization, quantization, and retraining strategies.
- Collaborate with data scientists, product teams, and stakeholders to deliver AI-driven solutions.
- Create LLDs and contribute to architecture design for scalable AI systems.
Mandatory Skills
- Python
- Machine Learning & • Deep Learning
- Generative AI & • Large Language Models (LLMs)
- Hugging Face Transformers
- LangChain
- RAG (Retrieval Augmented Generation)
- Vector Databases (FAISS, Pinecone, ChromaDB)
- FastAPI
- PyTorch / TensorFlow
- Apache Spark
- SQL
- Git & • MLflow
- Microservices Architecture
- Kubernetes
- CI/CD (GitHub Actions, Jenkins)
- AWS / Azure / GCP
Preferred Skills
- Go or Rust
- vLLM
- Terraform / CloudFormation / Pulumi
- Helm & • Service Mesh
- TorchServe / TensorFlow Serving
- Model Quantization & • Distillation
📌 Machine Learning Engineer (Noida)
🏢 EXL
📍 Noida