31 Jul
|
Video SDK
|
Surat
Job Description
Position : AI Research Engineer – Small Language Models & On-Device AI
Work Type : Full Time/Hybrid (India)
Company Location : Surat, Gujarat
At VideoSDK, we’re building the real-time intelligence layer for the next generation of applications — rapid, private, multimodal, and on-device.
We're hiring an AI Research Engineer to lead infrastructure development for small language models (SLMs) and speech systems, optimized for on-device multimodal AI. Your work will directly power products like real-time voice agents, live translation, intelligent video systems, and low-latency assistants — all running beyond the cloud.
What You'll Work On
Build scalable infrastructure to train, evaluate, and deploy small language models efficiently.
Design and implement RL-based training loops (e.g., PPO, DPO, RLAIF) for tuning small models in constrained environments.
Work with speech systems, including speech-to-text, text-to-speech, and voice activity detection.
Optimize models for on-device inference – targeting mobile, browser, and edge hardware (CPU, GPU).
Contribute to building real-time multimodal AI pipelines combining text, audio, and video.
Translate cutting-edge research into clean, production-ready code.
Drive experiments on quantization, distillation, and architecture search for effective deployment.
What We're Looking For
Robust understanding of training, testing, and fine-tuning small models (≤2B parameters).
Experience with reinforcement learning for language models (GRPO, PPO).
Proven experience working with speech models (Whisper, xTTS, Silero etc).
Ability to read and implement research papers quickly and effectively.
Solid Python and PyTorch skills with an eye for clean, modular code.
Experience with building and managing custom datasets, training pipelines, and evaluation suites.
Familiarity with multimodal AI, particularly combining audio, text, or video.
Bonus Points
Experience optimizing inference for mobile or
📌 Ai Research Engineer Surat
🏢 Video SDK
📍 Surat