الباحثون

Wenhao Guan

المنشورات 1

نسخة أولية وصول مفتوح

JoyAI-Voice 2.0: A Full-Continuous Autoregressive Speech Generation Model with Semantic-Acoustic Joint Representation

Yafeng Chen, Boya Dong, Yankun Huang وآخرون · 2026

We present JoyAI-Voice~2.0, an end-to-end anthropomorphic speech generation model built upon a fully continuous, dual-encoder architecture. Raw speech is encoded into continuous latents and partitioned into patches. Each patch is decomposed by a semantic-acoustic dual encoder into a semantically purified representation …

المؤلفون المشاركون