الباحثون

Zihan Dong

المنشورات 2

نسخة أولية وصول مفتوح

When Verifiable Counts Depend on Wording: Auditing Wording Robustness in Instruction Following

Qishi Zhan, Seoyeon Jang, Zihan Dong وآخرون · 2026

Verifiable instruction-following benchmarks often express each constraint through one fixed template. We test whether scores remain stable when the operational requirement is unchanged but its wording varies. We introduce WISE, a matched evaluation suite and reporting protocol instantiated on exact word count, keyword …

نسخة أولية وصول مفتوح

Count Evidence, Not Sentences: Tempered Evidence Fusion of LLM Judgments for Long-Text Value Measurement

Yuhe Wu, Rui Qian, Guangyu Wang وآخرون · 2026

Large language models (LLMs) are increasingly used to measure public value orientations from long social media posts, yet such posts often mix background, quotations, concessions, and only a few stance-bearing sentences. Existing approaches either ask the model to predict a document-level label directly, which can be o …

المؤلفون المشاركون