الباحثون

Xiaoyu Li

المنشورات 6

نسخة أولية وصول مفتوح

The Ball and the Box: Two Geometries of Computation in Superposition

Xiaoyu Li, Lequan Lin, Dai Shi وآخرون · 2026

Neural representations can encode more features than they have dimensions, a phenomenon known as superposition. We study the dimension needed to compute Boolean gates from such representations. For a single threshold layer with a Gaussian random dictionary and uniformly random sparse Boolean inputs, we derive sharp dim …

نسخة أولية وصول مفتوح

PulseBound: Future-Beat State Forecasting Under an Explicit Information Boundary

Chenyang Xu, Donglin Xie, Xi Xiang وآخرون · 2026

Predictive representation learning from photoplethysmography (PPG) can violate causal information access even with causal attention, as normalization, nonlocal transforms, or companion views may depend on withheld samples. We introduce PulseBound, a PPG representation learner combining physiologically structured future …

نسخة أولية وصول مفتوح

Rethinking Visual Provenance: Detection and Watermarking Across Direct Visual Generation and LLM-Driven Code Rendering

Zheng Gao, Xiaoyu Li, Zhicheng Bao وآخرون · 2026

AI systems create images and videos with image/video generation models or by writing code and graphics descriptions that are then rendered. These routes can produce similar visible artifacts but expose different representations, intervention points, and provenance evidence. We develop a production-centered framework th …

نسخة أولية وصول مفتوح

The Birkhoff Geometry of Manifold-Constrained Hyper-Connections: Two Channels, Vertex Viscosity, and Sinkhorn as a Retraction

Hyper-connections widen the residual stream of a Transformer to $n$ parallel streams. Their manifold-constrained version (mHC) mixes the streams at each layer with a doubly stochastic matrix, which it computes by Sinkhorn normalization of exponentiated logits. We give a geometric theory of this design on the Birkhoff p …

نسخة أولية وصول مفتوح

What Pretraining and Midtraining Make Learnable from Rewards?

A reward can identify a correct answer while leaving the computation needed for new inputs undetermined. We study how pretraining and midtraining supply the information and computation that make reward adaptation effective. In sequential state computation and contextual memory, we characterize mechanisms that agree on …

نسخة أولية وصول مفتوح

MatchFusion: Explicit-Implicit Instance Matching for Spatio-Temporal Multimodal Autonomous Driving

Xiaoyu Li, Jiajia Fu, Long Shi وآخرون · 2026

Sparse instance representations provide a compact interface for spatial LiDAR-camera and temporal past-current interaction in multimodal perception and E2EAD. Effective interaction requires reliable instance correspondences despite geometric discrepancies and heterogeneous semantic representations. Attention-based meth …

المؤلفون المشاركون