الباحثون

Helen M. Meng

المنشورات 1

نسخة أولية وصول مفتوح

DuraS2ST: Chain-of-Thought and Reinforcement Learning for Duration-Aligned Speech-to-Speech Translation

Yayue Deng, Dingdong Wang, Yuxuan Hu وآخرون · 2026

Speech-to-speech translation (S2ST) in time-sensitive applications such as video dubbing requires not only semantic fidelity and speaker preservation, but also strict duration consistency to avoid audio-visual misalignment. However, existing S2ST systems largely generate target speech without explicit temporal planning …

المؤلفون المشاركون