الباحثون

Xinyi Liu

المنشورات 4

نسخة أولية وصول مفتوح

LongSocialBench: Do Long-Context LLMs Understand Online Discussion Threads?

Xinyi Liu, Rinat Khaziev, Dilek Hakkani-Tür وآخرون · 2026

Long-context LLMs can now ingest entire online discussion threads, but understanding their social discourse requires more than reading a long document: models must track parent-reply relations, turning points, scoped subtrees, cross-branch contrasts, and participant trajectories. To test this structure-aware social rea …

نسخة أولية وصول مفتوح

Maintaining Benchmarks Against Increasingly Capable Agents: Detection and Remediation of Unearned Passes

Weijun Luo, Kelvin Luu, Xinyi Liu وآخرون · 2026

Agentic benchmarks guide model selection and training. Yet an agent can pass a task without demonstrating the intended capability. Such outcomes constitute unearned passes; their proportion among all passes defines the integrity gap. As agents improve, benchmark surfaces that once seemed harmless can become exploitable …

نسخة أولية وصول مفتوح

SV2V-RSim: A Comprehensive Benchmark for Self-Selective V2V Cooperative Perception with Near-Realistic Data

Yulu Wu, Chao Wei, Jujun Cheng وآخرون · 2026

Vehicle-to-Vehicle (V2V) cooperative perception enhances autonomous driving by enabling vehicles to share information beyond their direct line of sight. However, existing V2V datasets are limited by a small number of participating agents, static collaborator selection strategies, and a significant domain gap between si …

المؤلفون المشاركون