الباحثون

Chao Shen

المنشورات 3

نسخة أولية وصول مفتوح

DecepEval: A Benchmark for Evaluating Deception in LLM Agents

Yiming Xu, Hongyue Yu, Beihua Yang وآخرون · 2026

As large language model (LLM) agents become increasingly autonomous, they may pursue task performance through deception, raising concerns about their reliable deployment. Existing evaluations show that LLM agents can deceive, but often examine isolated scenarios or narrowly defined conditions, limiting systematic under …

نسخة أولية وصول مفتوح

MTOR: Generalizable AI-Generated Video Detection with Multimodal Semantics and Temporal Over-Regularity

Hang Wang, Chao Shen, Lei Zhang وآخرون · 2026

The rapid evolution of video generation has narrowed the perceptual gap between authentic and synthetic videos, making generalizable AI-generated video detection increasingly challenging. Existing detectors predominantly rely on visual representations, leaving caption-derived textual semantics underexplored. Meanwhile, …

نسخة أولية وصول مفتوح

When Confidence Rises Too Early: Detecting Shortcut Reasoning via Premature Answer Commitment

Zhaohan Zhang, Junjie Liu, Chengzhengxu Li وآخرون · 2026

The reasoning trajectory of a Large Language Model (LLM) is often treated as a verbalized description of its internal reasoning. However, such trajectories can be unfaithful: a model may rely on shortcuts to reach an answer and then post-rationalize the decision with a seemingly coherent chain of thought. Detecting thi …

المؤلفون المشاركون