13 Sep
|
Technozis
|
Karnataka
13 Sep
Technozis
Karnataka
Job DescriptionSenior VLM Developer – Vision Language Models
NLocation: Bangalore, India
NExperience: 7.5+ Years
NJoining: Immediate – September 2026
NWe are looking for an experienced Senior VLM Developer with solid foundations in Machine Learning and Deep Learning to build, fine-tune, and evaluate Vision-Language Models (VLMs) for multimodal perception and understanding applications.
NThe role involves developing large-scale training pipelines, curating multimodal datasets, experimenting with model architectures, and optimizing models for downstream applications.
NKey Responsibilities
n
N
Design, implement, and fine-tune VLM architectures such as LLaVA, Qwen-VL, or similar models. N
Develop data preprocessing and training pipelines for large-scale multimodal datasets. N
Optimize model performance across visual-language benchmarks and internal use cases. N
Collaborate with research and synthetic data teams to integrate generated data into model training. N
Conduct ablation studies and experiments to evaluate model performance. N
Maintain documentation of experiments, model behavior, and evaluation results. N
Develop and optimize ML/DL solutions for complex business problems. N
Work independently on statistical, machine learning, and research-oriented projects. N
Collaborate with cross-functional teams to translate research and experimentation into practical AI solutions. N
nRequired Skills
N
n
7.5+ years of experience in Machine Learning / Deep Learning / Data Science. N
Robust hands-on experience with Vision-Language Models (VLMs) and multimodal AI. N
Experience designing and fine-tuning VLM architectures such as LLaVA, Qwen-VL, or equivalent models. N
Strong hands-on expertise in PyTorch. N
Experience with PyTorchLightning. N
Strong experience with the Hugging Face ecosystem. N
Solid understanding ofTransformers and vision-language architectures. N
Experience with multimodal fusion techni
📌 Hiring: Senior Vlm Developer Karnataka
🏢 Technozis
📍 Karnataka