نسخة أولية وصول مفتوح
Less Context, Better Geometry: Masked Geometric Encoder for Robust 3D Foundation Models
Recent progress in 3D foundation models has enabled rapid 3D reconstruction and camera calibration by leveraging learned 3D priors from vast amount of spatial data. However, the all-to-all global attention design leads to quadratic complexity and limits long-sequence inference; unconstrained cross-view interactions als …