Location: Hyderabad (primary hub — deepest voice-AI talent pool in India outside Bangalore)
Experience: 7–10 years in ML/speech engineering, with 2+ years architecting production voice pipelines
Responsibilities
Own end-to-end architecture: ASR → NLU/LLM → dialogue manager → TTS → telephony
Make build-vs-buy calls on ASR/TTS vendors (Deepgram, Azure Speech, ElevenLabs, or open-source Whisper/Coqui)
Own latency budget across the pipeline (target sub-1.5s round-trip for live calls)
Mentor Conversational AI and ML/Voice engineers; own technical hiring bar
Qualifications
Must-Have Technical Qualifications
Voice AI & Speech Technologies
Production experience with at least one ASR engine (Whisper, Deepgram, Google STT, Azure Speech) at scale
Comfortable working across Hindi/Hinglish and English accent handling for Indian/UAE markets
Experience handling dialect variation and code-switching (e.g. Hinglish, Gulf Arabic variants)
in production ASR/NLU
LLM & Conversational AI
Hands-on experience with LLM-based dialogue systems (function calling, RAG, or fine-tuning)
Telephony & Infrastructure
Robust grasp of telephony integration: SIP, WebRTC, Twilio/Exotel/Ozonetel or similar
Performance & Optimization
Experience optimizing for latency and cost simultaneously in a live voice pipeline
Track record of hitting production accuracy benchmarks: 95%+ transcription (WER-based) accuracy and 90%+ intent classification accuracy at scale — not just in a demo
Required Skills
Positive-to-Have
Prior experience at Amazon Alexa, Microsoft Cortana/Speech, Google Assistant, or a voice-AI startup
Exposure to Arabic ASR/TTS for UAE market
2+ years hands-on with LLMs/Generative AI in a production (not POC) setting
📌 Tech Lead Hyderabad
🏢 Hireologist
📍 Hyderabad
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.