الملخص
Although the advancement of vision-language models (VLMs) has endowed robots with enhanced environmental understanding and task reasoning, a comprehensive evaluation methodology is important to advance the integration of VLMs in robotic navigation and manipulation. However, current benchmarks lack a comprehensive method to evaluate diverse robotic tasks, and evaluation metrics remain relatively constrained, making it difficult to assess the embodied capabilities of VLMs in a thorough and fine-grained manner. To address this issue, we propose RMMBench, an evaluation benchmark that requires robots to understand language instructions and perform long-horizon tasks in continuous spaces. RMMBench seamlessly integrates high- and low-level embodied tasks into a unified framework, constructing a "navigation-manipulation" task suite comprising 70 canonical task scenarios that range from localized manipulation to long-horizon composite navigation. The results reveal that leading VLMs still face major challenges in spatial localization when performing mobile manipulation tasks, and also highlight the necessity of enhancing the spatial perception capability of robots during long-horizon interactions. RMMBench can be accessed at https://mxxq-stack.github.io/rmmbench-project/
الكلمات المفتاحية
الموضوع
بيانات النشر
- المجلة
- غير متاح
- وصول مفتوح
- وصول مفتوح أخضر
اقتبس هذه المقالة
APA 7
Li, H., Feng, F., Fan, J., Yang, S., Chen, F., Cao, X., Song, R., & Zhang, W. (2026). RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation. https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation
MLA 9
Li, Huapeng, et al. "RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation." https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation.
شيكاغو (المؤلف–التاريخ)
Li, Huapeng, Fuxiang Feng, Jinqiu Fan, Shuo Yang, Fengjiao Chen, Xuezhi Cao, Ran Song, and Wei Zhang. 2026. "RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation." https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation.
هارفارد
Li, H., Feng, F., Fan, J., Yang, S., Chen, F., Cao, X., Song, R. and Zhang, W. (2026) 'RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation', Available at: https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation.
فانكوفر
Li H, Feng F, Fan J, Yang S, Chen F, Cao X, et al. RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation. https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation
IEEE
H. Li, F. Feng, J. Fan, S. Yang, F. Chen, X. Cao, R. Song, and W. Zhang, "RMMBench: A Comprehensive Benchmark for Robotic Mobile Manipulation," https://omanscience.com/ar/articles/rmmbench-a-comprehensive-benchmark-for-robotic-mobile-manipulation.