الباحثون

Chenxi Wang

المنشورات 3

نسخة أولية وصول مفتوح

How to Tame a Multi-Headed Hydra? Adaptive Multi-Category Safety Steering for Large Language Models

Chenxi Wang, Ruiyang Huang, Li Huang وآخرون · 2026

As large language models (LLMs) become increasingly widespread, preventing unsafe responses to harmful prompts is essential for their safe deployment. Activation steering offers an approach to improving LLM safety by modifying internal activations during inference without updating model parameters. However, a single pr …

نسخة أولية وصول مفتوح

Rethinking Visual Embodiment Dependence in Visuomotor Policies

Hongjie Fang, Yuxuan Lu, Chenxi Wang وآخرون · 2026

Visuomotor policies observe both the task scene and the acting embodiment, allowing embodiment-specific visual cues to influence action prediction. We study this phenomenon as visual embodiment dependence (VED) and show, through cue-conflict interventions across representative policies, that visible robot configuration …

المؤلفون المشاركون