Experience: 3 -6 years
Job Location: Noida, Gurugram, Bengaluru
Notice Period: 30 Days or Immediate Joiner
We are seeking an experienced ML Engineer with robust expertise in Machine Learning, Large Language Models (LLMs), Generative AI, MLOps, and Cloud Technologies . The ideal candidate will be responsible for designing, developing, deploying, and optimizing scalable AI/ML solutions, including LLM-powered applications, RAG systems, and enterprise-grade data pipelines.
Key Responsibilities
- Design, develop, and deploy ML and LLM-based applications in production environments.
- Build and optimize RAG pipelines, semantic search solutions, and AI-powered chatbots.
- Develop scalable microservices and APIs using Python, FastAPI, REST, and gRPC.
- Create and maintain robust CI/CD pipelines and MLOps workflows.
- Design enterprise-grade data pipelines using Spark and cloud platforms.
- Implement model serving, monitoring, optimization, quantization,
and retraining strategies.
- Collaborate with data scientists, product teams, and stakeholders to deliver AI-driven solutions.
- Create LLDs and contribute to architecture design for scalable AI systems.
Mandatory Skills
- Python
- Machine Learning & Deep Learning
- Generative AI & Large Language Models (LLMs)
- Hugging Face Transformers
- LangChain
- RAG (Retrieval Augmented Generation)
- Vector Databases (FAISS, Pinecone, ChromaDB)
- FastAPI
- PyTorch / TensorFlow
- Apache Spark
- SQL
- Git & MLflow
- Microservices Architecture
- Kubernetes
- CI/CD (GitHub Actions, Jenkins)
- AWS / Azure / GCP
Preferred Skills
- Go or Rust
- vLLM
- Terraform / CloudFormation / Pulumi
- Helm & Service Mesh
- TorchServe / TensorFlow Serving
- Model Quantization & Distillation
📌 Machine Learning Engineer (Noida)
🏢 EXL
📍 Noida