16 Sep
|
Technozis
|
Karnataka
16 Sep
Technozis
Karnataka
Job DescriptionSenior VLM Developer – Vision Language Models
NLocation: Bangalore, India
NExperience: 7.5+ Years
NJoining: Immediate – September 2026
NWe are looking for an experienced Senior VLM Developer with strong foundations in Machine Learning and Deep Learning to build, fine-tune, and evaluate Vision-Language Models (VLMs) for multimodal perception and understanding applications.
NThe role involves developing large-scale training pipelines, curating multimodal datasets, experimenting with model architectures, and optimizing models for downstream applications.
NKey Responsibilities
n
N
- Design, implement, and fine-tune VLM architectures such as LLaVA, Qwen-VL, or similar models.
N
- Develop data preprocessing and training pipelines for large-scale multimodal datasets.
N
- Optimize model performance across visual-language benchmarks and internal use cases.
N
- Collaborate with research and synthetic data teams to integrate generated data into model training.
N
- Conduct ablation studies and experiments to evaluate model performance.
N
- Maintain documentation of experiments, model behavior, and evaluation results.
N
- Develop and optimize ML/DL solutions for complex business problems.
N
- Work independently on statistical, machine learning, and research-oriented projects.
N
- Collaborate with cross-functional teams to translate research and experimentation into practical AI solutions.
N
nRequired Skills
N
n
- 7.5+ years of experience in Machine Learning / Deep Learning / Data Science.
N
- Strong hands-on experience with Vision-Language Models (VLMs) and multimodal AI.
N
- Experience designing and fine-tuning VLM architectures such as LLaVA, Qwen-VL, or equivalent models.
N
- Strong hands-on expertise in PyTorch.
N
- Experience with PyTorchLightning.
N
- Strong experience with the Hugging Face ecosystem.
N
- Strong understanding ofTransformers and vision-language architectures.
N
- Experience with multimodal fusion techniques.
N
- Experience with distributed training frameworks.
N
- Hands-on experience with large-scale model training, fine-tuning, and evaluation.
N
- Experience with PEFT / parameter-effective fine-tuning and/or reinforcement learning techniques.
N
- Strong experience in data preprocessing and multimodal dataset preparation.
N
- Experience working withsynthetic data is preferred.
N
nMandatory Domain Experience
NCandidates must have experience in at least one of the following domains:
NCPG | Retail | Pharma
nCandidates without relevant CPG, Retail, or Pharma domain experience should not be considered.
NEducation
nB.Tech / M.Tech / Ph.D. in:
N
n
- Computer Science
N
- Artificial Intelligence
N
- Machine Learning
n
- Data Science
N
- Or a related technical discipline
N
nJoining Requirement: Immediate joiners only.
NCandidates must be able to join on or before 30 September 2026.
NCandidates with a joiningdate after 30 September 2026 will not be considered.
📌 Senior Vlm Developer - Vision Language Models (Karnataka)
🏢 Technozis
📍 Karnataka