الباحثون

Longbo Huang

المنشورات 5

نسخة أولية وصول مفتوح

FlexLoop: Depth-Elastic Looped Policies for Adaptive Test-Time Computation in Deep RL

Xun Wang, Ruishuo Chen, Yu Chen وآخرون · 2026

Looped architectures scale computation by reusing the same parameters across recurrent steps, and recent work shows that they substantially improve deep reinforcement learning policies on long-horizon tasks. Since recurrent depth directly controls computation, one may expect looped policies to naturally support elastic …

نسخة أولية وصول مفتوح

One Proposal for Every Margin: Zero-Shot Amortized Sequential Importance Sampling for Binary Matrices

Ruishuo Chen, Weijia Li, Xun Wang وآخرون · 2026

In ecology, psychometrics, and the analysis of social and financial networks, binary matrices are often analyzed conditional on their observed row and column sums, which restricts the problem to a finite sample space of matrices with the same margins. Two fundamental problems are to count this space and to sample unifo …

نسخة أولية وصول مفتوح

The Router Within: Eliciting Native Skill Routing from a Frozen LLM

Ruishuo Chen, Xun Wang, Yu Chen وآخرون · 2026

Skills extend an LLM agent beyond its parametric knowledge, and the gain they promise rests on picking the right one. Deployed harnesses route by preloading every skill's metadata into the context, which disperses the agent's attention and caps the library size. Retrieval pipelines move the selection out of the context …

المؤلفون المشاركون