الباحثون

Daniel Seita

المنشورات 3

نسخة أولية وصول مفتوح

ExStereo: Lifting 2D Vision-Language-Action Models to 3D with Explicit Stereo Representations

Three-dimensional perception is critical for robotic manipulation, particularly for high-precision tasks, as recovering metric depth and precise 3D object positions from monocular RGB observations is inherently ill-posed. However, many Vision-Language-Action (VLA) models rely solely on RGB observations for perception. …

نسخة أولية وصول مفتوح

When Does Touch Matter? Charting the Vision-Interaction Gap in Cluttered Dexterous Grasping

Dexterous grasping in clutter poses a basic sensing question: when do tactile measurements and external wrench estimates improve on visual geometry? Occlusion and contact can obscure grasp quality, motivating a controlled evaluation of these interaction signals. We present a controlled real-world study over five tablet …

المؤلفون المشاركون