الباحثون

Qiang Huang

المنشورات 6

نسخة أولية وصول مفتوح

Can Agent Harnesses and Inference Engines Hear Each Other? The HEAR Protocol for Agentic LLM Serving

Jiaqi Zhao, Haodong Chen, Jitai Hao وآخرون · 2026

LLM agents increasingly execute complex workflows involving multi-turn reasoning, tool use, and parallel agents. Efficient serving requires decisions that span two layers with complementary information: the agent harness understands workflow dependencies, context lifecycles, and execution objectives, whereas the infere …

نسخة أولية وصول مفتوح

Beyond Refusal Patterns: Safe-Role Internalization for Robust and Generalizable LLM Safety Alignment

Jinghao Pang, Jitai Hao, Qiang Huang وآخرون · 2026

Large Language Models (LLMs) have achieved remarkable capabilities but remain vulnerable to jailbreak attacks that elicit harmful or unsafe outputs. Existing safety alignment approaches, including Supervised Fine-Tuning (SFT) and Reinforcement Learning from Human Feedback (RLHF), often require substantial attack-specif …

نسخة أولية وصول مفتوح

SparseEngine: Sparse-First Inference Engine

Jitai Hao, Quansheng Gu, Qiang Huang وآخرون · 2026

Long-context LLM agents accumulate interaction histories that strain KV-cache memory and attention computation. Although sparse attention reduces these costs, heterogeneous cache representations and workflows hinder integration with existing inference engines, while prior sparse-serving abstractions support only specif …

نسخة أولية وصول مفتوح

EvoSteer: Online Self-Evolving Graph Orchestration via Reference-Anchored Credit Assignment

Mingda Zhang, Hanwen Zhang, Qiang Huang وآخرون · 2026

In recent years, LLM-based multi-agent systems have been widely applied to orchestrate tool-using agents into executable communication graphs. However, existing self-evolving orchestration still faces key challenges, including post-hoc evolution that revises the team only after the trajectory ends, credit diffusion tha …

نسخة أولية وصول مفتوح

CollabFlow: Recursive Self-Improvement of Agent Collaboration

Xiao Huang, Mingda Zhang, Junming Zhang وآخرون · 2026

Recursive self-improvement (RSI) lets a system improve from its own outcomes; in LLM-based multi-agent systems, Agents refine one another within a task, and outcomes improve how they collaborate across tasks. However, existing multi-agent collaboration leaves this loop open: collaboration is pre-defined at the operator …

المؤلفون المشاركون