نسخة أولية وصول مفتوح
VGGT-Bridge: Beyond Sequential Pose Graphs via Coarse-Stride Skip Edges
Feed-forward visual geometry transformers such as VGGT reconstruct dense 3D structure from images in a single forward pass, simplifying multi-view 3D reconstruction. However, their quadratic attention complexity makes them difficult to scale to long sequences with thousands of frames. Chunk-and-align frameworks address …