Abstract

A camera and an IMU are the minimal sensor setup for metric localization and dense mapping, yet classical visual--inertial filters must wait for parallax before they start and then retain only sparse landmarks. Feed-forward geometry models, in contrast, predict dense structure from a few images but provide neither metric scale nor gravity. We present DAVIO, which uses a single multi-view depth model, Depth Anything~3, for both start-up and mapping. At start-up, a five-image window and preintegrated IMU measurements form a feature-free linear system. Its robust, conditioning-checked solution bootstraps a VIO filter through buffered replay. During tracking, the filter's metric poses condition the depth model. Residual scale is corrected only along viewing rays, which preserves the metric camera baselines, and a gravity-preserving submap graph with drift-gated revisits refines the map. On EuRoC, DAVIO starts markedly earlier, reduces the localization error, and maps more accurately than SOTA feed-forward mappers given identical poses. On building-scale ORI sequences, DAVIO is on bar or better than SOTA mappers on the same odometry, and degrades far less when GT poses are replaced by real odometry. We release the code of DAVIO, a real-time dense metric SLAM system, to the community.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Mahmoud, J., Movsesyan, A., Iumanov, M., & Kolyubin, S. (2026). DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping. https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping

MLA 9

Mahmoud, Jaafar, et al. "DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping." https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping.

Chicago (author–date)

Mahmoud, Jaafar, Arthur Movsesyan, Mikhail Iumanov, and Sergey Kolyubin. 2026. "DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping." https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping.

Harvard

Mahmoud, J., Movsesyan, A., Iumanov, M. and Kolyubin, S. (2026) 'DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping', Available at: https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping.

Vancouver

Mahmoud J, Movsesyan A, Iumanov M, Kolyubin S. DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping. https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping

IEEE

J. Mahmoud, A. Movsesyan, M. Iumanov, and S. Kolyubin, "DAVIO: Dense Monocular-Inertial SLAM with Feed-Forward Initialization and Pose-Conditioned Mapping," https://omanscience.com/en/articles/davio-dense-monocular-inertial-slam-with-feed-forward-initialization-and-pose-conditioned-mapping.