الباحثون

Yiqiao Jin

المنشورات 3

نسخة أولية وصول مفتوح

Can Language Models Learn to Reject Their Own Bad Reasoning Steps?

Siheng Xiong, Xiaoze Liu, Yiqiao Jin وآخرون · 2026

Verifier-guided decoding can prevent harmful reasoning steps from contaminating subsequent generation, but typically relies on an external learned verifier. We ask whether a language model can instead reject its own bad reasoning steps. We define a prefix's recoverability as the probability that the frozen generator ca …

نسخة أولية وصول مفتوح

From Knowledge Access to Source Learning: Developing Source-Specific Competence

Lucheng Fu, Kejing Xia, Yiyang Wang وآخرون · 2026

Large language model (LLM) agents increasingly rely on persistent external sources to solve sequences of knowledge-intensive tasks. Existing methods improve how source content is accessed and organized, while agent-memory systems preserve reusable knowledge from prior interactions, but repeated use of the same source i …

المؤلفون المشاركون