الباحثون

An Zhang

المنشورات 4

نسخة أولية وصول مفتوح

Selective Transfer of RL Updates for Visual Reasoning

Suxin Ji, Hungtao Wan, Mingjun Liu وآخرون · 2026

Model merging provides a training-free way to transfer reasoning capabilities from language models to vision-language models (VLMs), but endpoint-based transfer can conflate pre-existing model differences with changes acquired during reasoning post-training. We instead formulate capability transfer around the training- …

نسخة أولية وصول مفتوح

Semantic Behavioral Watermarking: Paraphrase-Robust and Forgery-Resistant Provenance for LLM Agents

Suxin Ji, Hungtao Wan, Shaoxuan Chen وآخرون · 2026

Behavioral watermarking embeds an owner identifier in an LLM agent's high-level action choices, giving provenance without touching output tokens. Prior agent watermarks break in two ways. First, all three prior schemes bind the watermark to the exact action symbol, so renaming a tool desynchronizes decoding even when t …

المؤلفون المشاركون