13 Aug
|
Unicorn workforce
|
India
13 Aug
Unicorn workforce
India
Position: Senior AI Engineer – Vision Language Models (VLM)
Experience: 4+ Years
Relevant Experience: 2+ Years in Multimodal AI / VLM
Employment Type: Contractual
Work Mode: Remote
Notice Period:
Immediate Joiners / Short Notice Preferred
About the Role
We are looking for a
Senior AI Engineer – Vision Language Models (VLM)
with strong hands-on experience in
Multimodal AI, Computer Vision, Generative AI, and Video Analytics
.
The selected candidate will design, develop, optimize, and deploy AI solutions focused on
video understanding, scene analysis, activity recognition, and real-time incident detection
.
The role requires experience taking AI solutions from
research and prototyping through production deployment
, with strong expertise in VLMs, deep learning, Python, PyTorch, and GPU-based inference.
Key Responsibilities
VLM & Multimodal AI
- Design and develop
Vision Language Model (VLM)
solutions for video understanding and incident detection.
- Build multimodal AI systems combining visual, textual, and contextual information.
- Work with up-to-date multimodal foundation models such as
GPT-4o Vision, Gemini, Qwen-VL, LLaVA, InternVL, Florence
, or equivalent.
- Develop prompting, reasoning, and evaluation strategies for complex visual scenarios.
- Improve model performance through experimentation, prompt engineering, feedback loops, and iterative optimization.
Video Understanding & Incident Detection
- Develop AI pipelines for
real-time or near-real-time video analysis
.
- Build solutions for detecting incidents such as:
- Falls
- Physical restraints
- Altercations
- Unsafe or unusual activities
- Other safety-related incidents
- Work on
scene understanding, activity recognition, action recognition, temporal reasoning, and video analytics
.
- Process and analyze large-scale image and video datasets.
Model Evaluation & Optimization
- Benchmark VLMs against traditional
Computer Vision and Action Recognition models
.
- Define appropriate evaluation methodolo
📌 Senior AI Engineer – Vision Language Models (India)
🏢 Unicorn workforce
📍 India