Voice AI Engineer | Foundation Models, ASR & TTS | 3–5 Years (Bengaluru)

Voice AI Engineer | Foundation Models, ASR & TTS | 3–5 Years (Bengaluru)

17 Sep
|
TalixoHR
|
Bengaluru

17 Sep

TalixoHR

Bengaluru

Train the models powering the next generation of Voice AI.

A fast-growing AI technology team is looking for a hands-on Voice AI Engineer who has already worked on training or pre-training foundation models .

This is not a conventional ML application role. You’ll work at the model layer across speech, foundation models, large-scale datasets and production optimization .

ABOUT THE OPPORTUNITY

You’ll own the journey from speech/text data → model training → evaluation → optimization → production deployment , solving challenging problems across accuracy, latency and inference efficiency.

The role is ideal for an engineer with strong Speech AI / ML / NLP fundamentals who has actually trained models rather than only consuming existing AI APIs.

WHAT YOU'LL OWN

- Develop, train, fine-tune and evaluate foundation models for Voice AI .
- Build speech solutions across ASR, TTS, speaker identification and voice activity detection .
- Work on conversational AI and speech-language models.
- Prepare and optimize large-scale speech and text datasets .
- Design training pipelines and run model experiments.
- Work with GPU-based and distributed training workloads.
- Improve model accuracy, latency, robustness and inference efficiency.
- Apply model optimization techniques including quantization .
- Debug model, data and training-pipeline issues.
- Support production deployment and performance optimization.

THE IDEAL CANDIDATE We're looking for someone who has actually trained or pre-trained foundation models ,



not someone whose experience is limited to prompting, API integration or using pre-trained models.

MUST HAVE

- 3–4 years of hands-on experience in ML, Speech AI, NLP or closely related areas.
- Demonstrated experience training/pre-training foundation models .
- Strong understanding of Transformers and deep-learning architectures .
- Hands-on experience with ASR and/or TTS models .
- Strong Python skills.
- Hands-on experience with PyTorch and/or TensorFlow .
- Experience with distributed training and GPU workloads.
- Experience handling large-scale speech/text data pipelines.
- Model evaluation, optimization and production deployment exposure.
- Strong analytical and debugging skills.

VALUABLE TO HAVE

- Multilingual or Indian-language speech models .
- Whisper, wav2vec 2.0, HuBERT, NeMo, SpeechBrain or Hugging Face.
- Audio preprocessing, augmentation or annotation.
- Dataset quality improvement experience.

WHY THIS ROLE

- Work directly on Voice AI foundation-model development .
- Go beyond application-layer GenAI into actual model training .
- Work with large-scale speech datasets and GPU infrastructure.
- Solve challenging problems around accuracy, latency and inference .
- Build expertise across ASR, TTS and conversational AI .

R OLE DETAILS Role: Voice AI Engineer / Speech ML Engineer

Experience: 3–5 Years

Domain: Voice AI / Speech ML / NLP

Employment: Full-Time

Location: HSR Layout Bangalore

📌 Voice AI Engineer | Foundation Models, ASR & TTS | 3–5 Years (Bengaluru)
🏢 TalixoHR
📍 Bengaluru

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: voice ai engineer | foundation models, asr & tts | 3–5 years (bengaluru) / bengaluru

Subscribe to this job alert:

Get the latest job offers by email for: voice ai engineer | foundation models, asr & tts | 3–5 years (bengaluru) / bengaluru