الباحثون

Jinpeng Li

المنشورات 4

نسخة أولية وصول مفتوح

MedZERO: Self-Evolving Agents for Open-Ended Medical Reasoning Through Controlled Knowledge Accumulation

Xilin Dang, Weilin Ruan, Xue Yang وآخرون · 2026

Large language models (LLMs) have shown promise in medical question answering and clinical reasoning, yet their improvement remains constrained by static parametric knowledge and costly expert supervision. Self-evolving agents offer a promising alternative by enabling models to improve through iterative task generation …

نسخة أولية وصول مفتوح

HARPO: Hallucination-Aware Reinforcement Learning for Faithful and Creative Language Generation

Tiezheng Yu, Yuxin Jiang, Jinpeng Li وآخرون · 2026

Large Language Models (LLMs) are prone to generating hallucinated content, which compromises their reliability in knowledge-intensive tasks. To address this challenge without sacrificing creativity, we propose HARPO, a reinforcement learning framework designed to jointly optimize faithfulness and creativity. HARPO inco …

نسخة أولية وصول مفتوح

Code Owns the Simulation, Jev Owns the Evaluation

Yaodong Yang, Hongyao Tang, Yi Ma وآخرون · 2026

Judgment models such as \jev{} return, in a single call and without reasoning text, a probability for each described option. This makes them attractive as an agent's action-selection layer, but it is unclear which decisions they can be trusted with. We test \jev{} on reflection tests, one-shot matrix games, the text ga …

المؤلفون المشاركون