Authors

Dehao Zhang

Publications 2

Preprint Open access

Spike-driven Vision-Language-Action Model

Vision-language-action (VLA) models bridge multimodal understanding and robotic control, advancing the dominant paradigm for embodied intelligence. However, most existing models rely on large Transformers, whose latency and energy costs hinder deployment on resource-constrained platforms. Through sparse event-driven co …

Co-authors