الباحثون

Zhixuan Liang

المنشورات 4

نسخة أولية وصول مفتوح

Immiscible Diffusion Policy: Preserving Multimodal Robot Actions through Label-Free Noise Assignment

Xiao Zhang, Yuxin Chen, Zhixuan Liang وآخرون · 2026

When diffusion policies were first introduced, they were expected to recover multi-modal action distributions. However, we find this expectation does not always hold, as diffusion policies often collapse to a single modality even when we guarantee the balance of dataset modalities and exact within-batch symmetry. Our a …

نسخة أولية وصول مفتوح

EmbodiedRSI: Active Continual Robot Learning Through Hypothesis-Guided Co-Evolution

Python Song, Zhixuan Liang, Kelsey Fu وآخرون · 2026

Robot foundation models provide strong visuomotor control, yet their performance can degrade when object positions or task instructions change. Further improvements often require post-training on substantial robot data, which can be costly to collect through methods such as teleoperation. Agentic harnesses can adapt ar …

نسخة أولية وصول مفتوح

Beyond Future Prediction: Denoising as Generative Adaptation for Robot Control

Zanyi Wang, Yuheng Lei, Dengyang Jiang وآخرون · 2026

Pretrained generative Diffusion Transformers (DiTs) capture rich pixel-level visual and language-conditioned structure through large-scale image and video generation training. A growing line of robot policies builds on this generative prior, but how it should be transferred to control remains unclear, and existing appr …

نسخة أولية وصول مفتوح

M2Tok: Multi-head Multi-codebook Discrete Action Tokenization for Vision-Language-Action Models

Chunpu Xu, Zhixuan Liang, Yuhao Zhang وآخرون · 2026

Recent advancements have successfully adapted autoregressive language models to process multimodal signals, such as images and actions. Since raw action signals are continuous, effective tokenization is essential to map high-dimensional inputs into compact discrete tokens for autoregressive processing. However, existin …

المؤلفون المشاركون