24 Sep
|
TalixoHR
|
Bangalore Metropolitan Area
24 Sep
TalixoHR
Bangalore Metropolitan Area
We are hiring a hands-on Machine Learning Engineer to work on Voice AI, Speech AI and Foundation Model initiatives.
The ideal candidate should have 3–4 years of hands-on experience in Machine Learning, Speech AI, NLP or related fields , with proven experience in training or pre-training foundation models .
This is a hands-on role involving ASR, TTS, Speech Recognition, Speech Synthesis, Transformer models, Large Language Models, Deep Learning, GPU-based training and large-scale ML pipelines .
Key Responsibilities
- Develop, train, pre-train, fine-tune and evaluate foundation models for Voice AI applications
- Build and improve Automatic Speech Recognition (ASR) and Text-to-Speech (TTS) models
- Work on Speech Recognition, Speech Synthesis, Speaker Identification, Voice Activity Detection (VAD) and Conversational AI
- Design and implement machine learning and deep learning training pipelines
- Prepare, process and optimize large-scale speech, audio and text datasets
- Conduct model experiments and improve model accuracy, robustness, latency and inference performance
- Work with Transformer architectures and deep learning models
- Handle GPU-based workloads and distributed model training
- Optimize models using techniques such as quantization and inference optimization
- Support production deployment, model evaluation and performance debugging
Required Qualifications
- 3–4 years of hands-on experienc e in Machine Learning, Speech AI, Speech Processing, NLP, Deep Learning or related area
- sProven hands-on experience with training / pre-training foundation model
- sStrong understanding of Deep Learning and Transformer architecture
- sHands-on experience with ASR / Automatic Speech Recognitio
- nHands-on experience with TTS / Text-to-Speech / Speech Synthesi
- sStrong programming skills in Pytho
- nStrong experience with PyTorch and/or TensorFlo
- wExperience with GPU computing, distributed training and large-scale model trainin
- gExperience working with large-scale speech, audio or text dataset
- sUnderstanding of model evaluation, optimization, quantization and inferenc
- eStrong analytical, debugging and problem-solving skill
s Preferred Qualification
- s Experience buildin g multilingual speech mode
- lsExperience wit h Indian language / Indic language speech mode
- lsHands-on experience wit h Whisper, wav2vec 2.0, HuBERT, NVIDIA NeMo, SpeechBrain or Hugging Fa
- ceExperience wit h audio preprocessing, speech data augmentation, audio annotation and dataset quality improveme
- ntExperience deployin g ML / Deep Learning / Speech AI models in producti
- onExperience wit h Conversational AI, Voice AI or Generative < /li>A **I What We Are Looking F** or
We are specifically looking for candidates who have work ed at the model-training le vel.
Candidates with experience only in using OpenAI APIs, building chatbots, prompt engineering, RAG, API integration or consuming pre-trained Voice AI mod els may not be relevant unless they also have hands-on foundation-model training/pre-training experienc **e.
Experien** **ce
3–4 yea** **rs
App** ly
If your experience includ es Foundation Model Training, Speech AI, ASR, TTS, Transformers, PyTorch/TensorFlow and large-scale GPU/distributed train ing, apply with your updated C **V.
Relevant profiles will be prioritized for screeni** ng.
📌 Machine Learning Engineer - Voice AI / Speech AI / Foundation Models (Bangalore Metropolitan Area)
🏢 TalixoHR
📍 Bangalore Metropolitan Area