الباحثون

Jingwei Sun

المنشورات 3

نسخة أولية وصول مفتوح

Test-time Calibration Learning for Large Language Model Reasoning

Zizhuo Zhang, Xiong Peng, Jingwei Sun وآخرون · 2026

Reliable large language models (LLMs) must not only produce accurate answers but also express confidence that faithfully reflects their probability of being correct. Such calibration is essential for identifying uncertain predictions and supporting reliable decision-making in real-world deployment. Recent studies incor …

نسخة أولية وصول مفتوح

LatentSift: Policy-State Filtering for Token-Efficient Verification of Software Engineering Agents

Yuning Han, Yangchenchen Jin, Tyler Jandreau وآخرون · 2026

Test-time scaling improves software engineering agents by generating multiple candidate trajectories and selecting the best one. Verifying and selecting among these long interactions can consume as many tokens as generation itself. Existing hybrid workflows first apply an LLM-based execution-free (EF) verifier to filte …

المؤلفون المشاركون