24 Sep
|
Digivance Solutions
|
India
24 Sep
Digivance Solutions
India
– AI Speech & Audio Specialist (ASR | TTS | Voice AI)
Role: AI Speech & Audio Specialist – ASR | TTS | Voice AI | Audio Data
Location: India (Remote)
Interview Mode: Virtual
We are looking for an AI Speech & Audio Specialist to join our AI Community and help evaluate and improve next-generation speech AI technologies, including Automatic Speech Recognition (ASR), Text-to-Speech (TTS), Voice AI, and audio machine learning systems.
Key Responsibilities
- Evaluate AI-generated speech (TTS) for naturalness, pronunciation, rhythm, tone, emotion, and overall voice quality.
- Review and improve ASR outputs by identifying transcription errors, accent issues, and linguistic nuances.
- Perform qualified audio editing: segmentation, trimming, noise reduction, silence adjustment, and quality enhancement.
- Analyze audio using waveform and spectrogram tools to detect quality issues.
- Support creation and validation of high-quality speech datasets for AI model training and evaluation.
- Conduct audio quality checks (loudness, background noise, clipping, distortion, export specs).
- Contribute to voice AI projects: TTS, ASR, voice generation, voice similarity, mimicry, and conversational AI evaluation.
Mandatory Qualifications Education / Background (any one):
- Bachelor’s degree or equivalent in: Audio Engineering,
Sound Engineering, Music Technology, Linguistics, Phonetics, Speech-Language Pathology, Computational Linguistics, Computer Science (Speech/AI focus), or related fields.
OR
- Professional experience in: Audio production, sound design, dialogue editing, voice recording, podcast/audio post-production, speech data creation, or AI speech projects.
Technical Experience:
- Hands-on with at least one professional audio tool: Adobe Audition, iZotope RX, Pro Tools, Reaper, Audacity, Logic Pro, Cubase, or other DAWs.
- Knowledge of: audio editing & restoration, noise reduction, speech segmentation, spectrogram analysis, audio quality assessment, and voice/audio processing workflows.
Language Requirements:
- Native-level proficiency (C1/C2+) in the target language preferred.
- Ability to identify pronunciation, accent, and linguistic nuances.
- Good English skills to understand technical instructions and guidelines.
Preferred Qualifications Experience with any of the following is a strong plus:
- ASR, TTS, Voice AI, speech datasets, audio annotation
- Voice cloning / voice mimicry projects
- ADR / dubbing, localization audio
- Broadcast or media production
- AI model evaluation for speech/audio
📌 AI Speech & Audio Specialist – ASR | TTS | Voice AI | Audio Data (India)
🏢 Digivance Solutions
📍 India