الباحثون

Yue Liu

المنشورات 3

نسخة أولية وصول مفتوح

Collective Bias Mitigation via Model Routing and Collaboration

Mingzhe Du, Luu Anh Tuan, Xiaobao Wu وآخرون · 2026

Large language models (LLMs) are increasingly deployed in public health, finance, and governance, requiring both accuracy and societal value alignment. Despite recent advances, LLMs often perpetuate or amplify bias embedded in their training data, posing challenges to fairness. While self-debiasing encourages an LLM to …

نسخة أولية وصول مفتوح

Aletheia: Permission-Minimality Testing for Coding-Agent Rules

Jieke Shi, Yuchen Chen, Junda He وآخرون · 2026

Repository instruction files guide coding agents, but also expose them to prompt injection. Malicious rules can request credential access or data transfer while the agent produces a correct patch. We present Aletheia, a framework for permission-minimality testing. Aletheia translates requested authority into a typed la …

المؤلفون المشاركون