10 Aug
|
HERE Technologies
|
Mumbai
10 Aug
HERE Technologies
Mumbai
What's the role?
We are seeking an experienced Lead Data Scientist to drive the development of advanced Vision Foundation Models, Vision-Language Models, and multimodal AI systems for large-scale image and video understanding.You will lead the design, development, and deployment of scalable AI solutions, translating cutting-edge research into real-world applications. Key Responsibilities
Design, train, and optimize large-scale vision foundation models across image and video modalities
Develop multimodal AI systems using architectures such as Vision Transformers (ViT), SAM, DINOv3, CLIP, and VLMs
Apply self-supervised learning, transfer learning, and fine-tuning approaches for downstream tasks
Build and enhance Vision-Language Models for visual reasoning and multimodal understanding
Develop Retrieval-Augmented Generation (RAG) pipelines and multimodal knowledge retrieval systems
Work with embeddings, vector databases, and semantic search frameworks
Build scalable pipelines for training, evaluation, and deployment
Manage large-scale image, video, and multimodal datasets
Optimize distributed training workflows and model performance
Translate research into production-ready solutions and explore emerging approaches in multimodal AI and generative AI
Evaluate model quality, robustness, and retrieval effectiveness
Who are you?
You bring solid expertise in computer vision, foundation models, and multimodal AI systems, along with the ability to deliver scalable solutions from research to production.
Master’s or PhD in Computer Science, Artificial Intelligence, Machine Learning, or a related field
Extensive experience in deep learning, computer vision, or multimodal AI
Strong programming skills in Python and experience with PyTorch
Deep understanding of computer vision, Vision Transformers, self-supervised learning, Vision-Language Models, and multimodal systems
Hands-on experience with foundation models such as SAM, DINOv3, CLIP, BLIP/BLIP-2, LLaVA, or diffusion-based visio
📌 Lead Data Scientist-(VLM-Multimodal AI (Mumbai)
🏢 HERE Technologies
📍 Mumbai