23 Sep
|
Innodata India
|
Noida
23 Sep
Innodata India
Noida
We are looking for exceptional Speech & Audio AI Evaluation Specialists with genuine in-house Global Capability Centre (GCC) / Captive international
voice experience to evaluate state-of-the-art Speech-to-Speech (S2S), Text-to-Speech (TTS), and real-time conversational voice agents. Unlike traditional
transcription or BPO roles, this specialist position demands a highly trained auditory ear to evaluate prosody, cadence, phonetic accuracy, emotional
steering, and paralinguistic nuance across global English dialects. Candidates must possess C2level near-native fluency to execute rigorous human
preference and synchronization benchmarks.
KEY RESPONSIBILITIES & CORE WORKFLOWS
- S2S & TTS Naturalness Scoring: Evaluate live Speech-to-Speech and neural Text-to-Speech outputs across intelligibility, rhythm, and conversational
cadence.
- Paralinguistic & Emotion Steering Evaluation: Benchmark how effectively models express nuanced vocal attributes including empathy,
hesitation,
tone inflection, irony, and situational urgency.
- Pairwise Audio Preference Ratings: Conduct blinded, head-to-head auditory preference evaluations between candidate audio completions, providing
detailed perceptual rationales.
- Multi-Modal data validations: Verify multi-modal synchronization, lip-sync alignment, and acoustic scene consistency for video dubbing and avatar
driven speech models.
- Accent & Dialect Calibration: Apply standardized rubric metrics uniformly across North American, British, Australian, and international English
accents without regional bias.
- Defensible Auditory Documentation: Document timestamped acoustic anomalies, unnatural vocal artifacts, metallic distortion, and hallucinated
phonetic segments
📌 Speech & Audio AI Evaluation Specialist (International Voice) (Noida)
🏢 Innodata India
📍 Noida