Authors

Bin Hu

Publications 4

Preprint Open access

DNAlign: Dynamic Null-Space Safe Alignment for LLMs

Ensuring the safe and reliable deployment of large language models (LLMs) remains a fundamental challenge. Existing safety alignment approaches either incur high computational cost or unintentionally disrupt the model's core knowledge, leading to degraded fluency and factual accuracy on benign tasks. This reveals a per …

Preprint Open access

Empty Commitments: When Agents Promise What They Cannot Deliver

Jiaqi Tang, Bingyu Shen, Lan Wei et al. · 2026

A chatbot that says "I will remind you tomorrow" will not run again until the user writes. We call such a promise an empty commitment: a promise of action after the current turn that nothing in the agent's tools or runtime can carry out. Unlike a broken promise, its emptiness is decided by the agent's configuration at …

Co-authors