الباحثون

Jie Gao

المنشورات 4

نسخة أولية وصول مفتوح

FastJEV: Understanding Redundancy for Compact JEV Inference

Jie Ma, Jie Gao, Yihang Liu وآخرون · 2026

JEV models make multimodal decisions by directly scoring candidates. Although the common context is encoded once, candidate evaluation can still repeat matching token histories, duplicate inference states, and execute the full backbone. In this paper, we study these sources of redundancy and present FastJEV for compact …

نسخة أولية وصول مفتوح

Accurate but Not Humble: Evaluating Epistemic Humility in LLM Agents under Knowledge Conflict

When retrieved evidence contradicts an agent's prior beliefs, does it revise its answer, acknowledge uncertainty, or persist with an incorrect conclusion? Existing evaluations of agentic systems focus primarily on task success, offering limited insight into how agents handle such conflicts. We propose to evaluate agent …

نسخة أولية وصول مفتوح

On the Behavioral Traits of LLM Agents

Haokai Zhao, Jie Gao, Yunze Xiao وآخرون · 2026

Users increasingly describe different AI agents as distinct colleagues to work with. AI personality research aims to quantify such impressions by attributing human-like "traits" to agents. However, existing measures fall short: models' self-reports (S-data) diverge from their actual behavior, while informant ratings fr …

نسخة أولية وصول مفتوح

Where Does Retrieval-Based Open-Ended Evaluation Fail? Automatic Taxonomy Induction from Long-Form Medical Answer Factuality Verification

Heyuan Huang, Jirui Dai, Alexandra DeLucia وآخرون · 2026

Retrieval-based factuality evaluation, where LLM-generated claims are verified against evidence from authoritative medical corpora, has become the dominant paradigm for scalable hallucination detection in high-stakes clinical settings. Despite the urgency of reliable and transparent medical fact verification, most syst …

المؤلفون المشاركون