الباحثون

Yuki Mitsufuji

المنشورات 6

نسخة أولية وصول مفتوح

Syn-Omni: Structured Specialization and Progressive Collaboration for Omnimodal Embeddings

Youngtaek Oh, Qiyu Wu, Hiromi Wakaki وآخرون · 2026

Omnimodal embeddings naturally involve both shared representations and modality-specific features across heterogeneous inputs. However, existing omnimodal embedding methods often rely on a single shared parameter space over mixed-modality data, limiting structural separation between universal and modality-specific repr …

نسخة أولية وصول مفتوح

End-to-End Historical Music Restoration in Latent Space

Steven Cho, Junghyun Koo, Raphael Lafargue وآخرون · 2026

Historical music restoration (HMR) has almost exclusively focused on constrained problems such as Super-Resolution or the restoration of solo pieces, under-exploring the general task of restoring orchestral historical music, which has multiple instruments. This under-exploration is largely because the HMR domain, early …

نسخة أولية وصول مفتوح

Distilling Diffusion Score Discrepancy for Efficient Training Data Attribution

Shixuan Liu, Joan Serrà, Kin Wai Cheuk وآخرون · 2026

Training data attribution for diffusion models aims to identify the training samples that influence a generated instance, but existing methods either require costly per-sample gradient computation or query-specific model optimization. Moreover, most methods attribute changes in a proxy loss rather than changes in the a …

نسخة أولية وصول مفتوح

Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models

Beomsu Kim, Chieh-Hsin Lai, Bac Nguyen وآخرون · 2026

World models predict future observations from current experience and actions, yet prediction can depend on observations seen far in the past. Episodic memory preserves past observations for later recall; however, as memory accumulates, it raises a fundamental question: which memories are useful for the current predicti …

نسخة أولية وصول مفتوح

Does Uniform Discrete Diffusion Need Time?

Uniform discrete diffusion models (UDMs) commonly use explicit time conditioning, but we find that it can often be unnecessary in practice. In this paper, we first show that the population-optimal UDM predictor generally depends on time: time controls how much the model should trust the observed context. We then show t …

نسخة أولية وصول مفتوح

LYRIC: Language-Driven Physics-Based Character Control for Contact-Rich Whole-Body Object Interaction

Zeyu Han, Zichong Meng, Julian Tanke وآخرون · 2026

We present LYRIC, a generative flow-matching controller for language-driven physics-based contact-rich interaction control, that enables simulated characters to perform contact-rich whole-body object interactions from a free-form language instruction and a sparse terminal object goal. To obtain reliable expert trajecto …

المؤلفون المشاركون