الباحثون

Xinyu Wang

المنشورات 10

نسخة أولية وصول مفتوح

Memory Forcing: Attendable Mid-Horizon History for Streaming Video Generation

Jiaming Zhang, Xinyu Wang, Huafeng Shi وآخرون · 2026

Autoregressive video diffusion enables causal video streaming without a bidirectional pass over the full clip, but existing few-step systems usually retain only the opening and most recent frames in a fixed-size KV cache. Once an event leaves this window, later frames can no longer attend to it, a failure we term mid-h …

نسخة أولية وصول مفتوح

When Is Deletion Ordering Tractable? From Update Dynamics to Permutation Structure

Xinyu Wang, Ziyu Zhao, Yixuan He وآخرون · 2026

Given a fixed set of pending deletion requests, retraining from scratch after each request is prohibitive, so a prescribed request-wise policy processes them sequentially. The resulting terminal model can depend on their order. Rather than prescribing an ordering rule, we study the permutation objective induced by the …

نسخة أولية وصول مفتوح

Sequential Functional Structured Tucker Compression for Large Language Model Attentions

Jiangfeng Chen, Xinyu Wang, Tianshuo Yan وآخرون · 2026

Post-training compression of LLM attention is often formulated as independent matrix approximation, ignoring both the shared structure among attention projections and the representation shift introduced by earlier compression. We propose FTC, a sequential structured compression framework that adapts the approximation t …

نسخة أولية وصول مفتوح

JARQ: Joint Alternating Refinement for Quantization

Group-wise post-training quantizers for large language models round weights onto a grid that is not refit to the resulting integer codes. We show that this leaves accuracy on the table: the best grid depends on the codes, input correlations couple the errors of different groups, and useful code changes often involve ma …

نسخة أولية وصول مفتوح

Zero2Repo: Can Coding Agents Build Repositories from Scratch?

Pei Yang, Tianyu Shi, Yuhang Yao وآخرون · 2026

Coding agents are increasingly asked to build software rather than patch it, yet benchmarks for from-scratch repository construction are mostly limited to a single language and depend on manually curated tasks. We introduce Zero2Repo, a benchmark in which an agent receives a product requirements document, an interface …

نسخة أولية وصول مفتوح

What Paired Evaluations Reveal under Visual Perturbations

Yongda Wei, Chen Zhang, Yifei Wang وآخرون · 2026

Robustness evaluation must examine diverse visual perturbations, while benchmarks cover only some real-world conditions and physical testing is costly. Paired evaluations link clean and perturbed predictions for the same image, capturing changes in correctness, confidence, and acceptance beyond aggregate accuracy. We i …

نسخة أولية وصول مفتوح

Dependency-Aware Trajectory Refinement for Efficient Multi-Turn Agent Fine-Tuning

Zhuo Chen, Zhen Zhang, Xinyu Wang وآخرون · 2026

Multi-turn agent trajectories often contain redundant rounds (failed tool calls, parallel sub-queries, verification-only steps) that inflate both training and inference cost. We propose viewing each trajectory as a \emph{round-level dependency DAG} that exposes which rounds are globally load-bearing for the final answe …

المؤلفون المشاركون