الباحثون

Xiuying Chen

المنشورات 4

نسخة أولية وصول مفتوح

Do MLLM Judges Judge the Edit? Auditing Bias in Image Editing Evaluation with Verified Quality Preservation

Multimodal large language models (MLLMs) are increasingly used as automated judges for instruction-based image editing and as reward signals for model training. However, systematically auditing whether these judges are influenced by cues irrelevant to editing quality is challenging because visual interventions may them …

نسخة أولية وصول مفتوح

UnlearningSoup: Is Repeated Tuning Necessary for Large Language Model Unlearning?

Puning Yang, Qizhou Wang, Junchi Yu وآخرون · 2026

Large language models trained on vast corpora inherently risk memorizing harmful content that may later re-emerge in their outputs. To mitigate this issue, existing unlearning methods typically rely on training-based parameter updates, such as gradient ascent and its variants, to delete targeted content while preservin …

نسخة أولية وصول مفتوح

JevOut: Natural Context Can Flip Decision Models

Zixiang Xu, Zirui Song, Chiyu Zhang وآخرون · 2026

An ordinary-looking background detail can turn a correct model decision into a confident mistake. We demonstrate this fragility in four decision systems, including Jev, across seven datasets covering knowledge, reasoning, and tool routing. Within 64 accepted target evaluations per decision, we uncover short context add …

المؤلفون المشاركون