الباحثون

Ziyang Ma

المنشورات 3

نسخة أولية وصول مفتوح

WorldSonus: Bringing Sound to Worlds

Pengjun Fang, Jingyi Fa, Kam Man Wu وآخرون · 2026

Recent advances in world models have enabled increasingly realistic visual synthesis. However, these generated environments remain largely silent. Bringing sound to world models poses three core challenges: real-time generation to keep pace with interactive video streams, interactive control to respond to mid-stream so …

نسخة أولية وصول مفتوح

Omni Demand Understanding: A Benchmark for Contextual User-Intent Inference in Multimodal Interaction

Qi Chen, Yunfei Chu, Haolin He وآخرون · 2026

Natural audio-visual interaction is emerging as an important interface for AI assistants, allowing users to communicate through speech and vision rather than carefully composed text prompts. However, existing benchmarks of interactive capabilities still focus primarily on response quality, leaving a more fundamental qu …

نسخة أولية وصول مفتوح

ECHO: Early-layer Collaborative Hierarchical Orchestration with Bonus Logits in Speculative Decoding

Ziyang Ma, Zihong Zhang, Zuchao Li وآخرون · 2026

While draft-model-free speculative decoding offers a promising path to efficient LLM inference, it is frequently constrained by stale draft candidates and the high computational cost of the verification. To address these challenges, we propose ECHO, a hierarchical dual-loop framework that exploits the functional asymme …

المؤلفون المشاركون