الباحثون

Soham Ray

المنشورات 2

نسخة أولية وصول مفتوح

mu-bench: A Multilingual Utterance Transcription Benchmark

Andrea Li, Soham Ray · 2026

Voice agents depend on accurate automatic speech recognition (ASR) to act on what callers say, yet ASR is evaluated on read, English-centric speech with word error rate (WER), which penalizes surface rather than semantic differences. We introduce mu-bench, a dataset of 4,270 caller utterances from 250 phone calls to an …

نسخة أولية وصول مفتوح

$τ$-Multilingual: Benchmarking Voice Agents Across Languages

English-only benchmarks expose only a narrow slice of voice-agent behavior. We introduce $τ$-Multilingual, extending $τ$-Voice to Spanish, Brazilian Portuguese, Hindi, Korean, and Mandarin with native-speaker review and evaluation of generated language and spoken output. Across 4,500 full-duplex calls and five voice co …

المؤلفون المشاركون