Moon, S., Jeon, S., Kim, S., Joo, H., & Shin, J. (2026). RLHND: Video Foundation Models as Physically Grounded Hand Trackers for Robot Learning. https://omanscience.com/en/articles/rlhnd-video-foundation-models-as-physically-grounded-hand-trackers-for-robot-learning