ROLE OVERVIEW & OBJECTIVE
We are looking for exceptional Speech & Audio AI Evaluation Specialists with genuine in-house Global Capability Centre (GCC) / Captive international voice experience to evaluate state-of-the-art Speech-to-Speech (S2S), Text-to-Speech (TTS), and real-time conversational voice agents. Unlike traditional transcription or BPO roles, this specialist position demands a highly trained auditory ear to evaluate prosody, cadence, phonetic accuracy, emotional steering, and paralinguistic nuance across global English dialects. Candidates must possess C2 level near-native fluency to execute rigorous human-preference and synchronization benchmarks.
KEY RESPONSIBILITIES & CORE WORKFLOWS
• S2S & TTS Naturalness Scoring: Evaluate live Speech-to-Speech and neural Text-to-Speech outputs across intelligibility, rhythm, and conversational cadence.
CANDIDATE PROFILE & QUALIFICATIONS
Mandatory Requirements
Preferred Qualifications
• Linguistics & Phonetics: Academic coursework or practical background in phonetics, phonology, auditory acoustics, or speech-language pathology.
• Speech QA Background: Prior professional experience in voice quality analytics, acoustic data annotation, or speech synthesis evaluation.
• Acoustic Acuity: Formally trained auditory ear for subtle vocal inflections, pitch modulation, cadence shifts, and articulation artifacts.
• Multilingual Ability: Additional fluency in major European, Asian, or Latin American languages to support cross-lingual speech benchmarks.
This listing is sourced from a third-party job board. Applying will redirect you to the original posting.
UDIT - University of Design, Innovation and Technology
UDIT - University of Design, Innovation and Technology