الباحثون

Han Zheng

المنشورات 3

نسخة أولية وصول مفتوح

Evolving in Thought Space: Training a Small Model at Test Time Unlocks Better Discoveries

Chonghe Jiang, Ao Qu, Siyuan Liu وآخرون · 2026

Open-ended scientific discovery often requires repeatedly proposing and evaluating candidate solutions. LLM-based systems can support this process by generating and refining executable solutions from verifier feedback. Methods such as TTT-Discover use test-time training (TTT) to update the solution-generating LLM from …

نسخة أولية وصول مفتوح

CoEM: Empowering Long-Context Reasoning with Commit-on-Evidence Memory

Jingguang Li, Yebo Wu, Zuyi Guo وآخرون · 2026

Long-context reasoning is essential for complex and long-horizon tasks, yet the performance of large language models (LLMs) degrades as context length increases. Recent approaches address this by processing input chunk by chunk while maintaining a bounded textual memory in model context. However, premature information …

المؤلفون المشاركون