03 Oct
|
Beas%20consultancy%20and%20services%20pvt%20ltd
|
India
03 Oct
Beas%20consultancy%20and%20services%20pvt%20ltd
India
Years Of Experience: 5+ Years
Location: Kolkata
Level: Senior
Duration: 4 Months
Vacancy: 1
Type: Full Time
Key Responsibilities
Fine-tune and adapt multimodal LLMs (e.g., Qwen-VL, LLaVA, or similar) for domain-specific document understanding.
Design prompt templates and instruction sets to improve JSON-structured output quality.
Perform different fine-tuning and incremental learning across multiple datasets to ensure generalization.
Implement evaluation metrics and validation datasets to track model accuracy and performance.
Optimize inference performance using quantization, LoRA adapters, and Ray Serve / vLLM / Unsloth frameworks.
Collaborate with backend teams to integrate the model into production pipelines.
Build monitoring tools to log model confidence, token usage, and inference latency
Required Skills
Robust Python programming skills with experience in PyTorch and Transformers (Hugging Face).
Experience working with OCR tools such as PaddleOCR,
Tesseract, or EasyOCR.
Hands-on experience fine-tuning or serving LLMs / VLMs (e.g., Qwen, LLaVA, Mistral, or Vicuna).
Knowledge of LoRA / QLoRA / PEFT adapters for productive model training.
Familiarity with JSON schema generation, prompt engineering, and structured data extraction.
Experience using Unsloth, Ray Serve, or vLLM for large-scale inference.
Comfortable working with Docker, CUDA, and NVIDIA GPUs for model deployment.
Robust understanding of tokenization, attention mechanisms, and quantization (4-bit / 8-bit )
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Ai Engineer Kolkata (India)
🏢 Beas%20consultancy%20and%20services%20pvt%20ltd
📍 India