Building temporally coherent 3D maps with VGGT for memory-efficient Semantic SLAM
Fuente:
arXiv
Saved in:
| Main Authors: | Dinya, Gergely, Halász, Péter, Lőrincz, András, Karacs, Kristóf, Gelencsér-Horváth, Anna |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAMannot: A Memory-Efficient, Local, Open-source Framework for Interactive Video Instance Segmentation based on SAM2
by: Dinya, Gergely, et al.
Published: (2026)
by: Dinya, Gergely, et al.
Published: (2026)
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
Automatic camera orientation estimation for a partially calibrated camera above a plane with a line at known planar distance
by: Dinya, Gergely, et al.
Published: (2025)
by: Dinya, Gergely, et al.
Published: (2025)
VGGT-SLAM++
by: Mandal, Avilasha, et al.
Published: (2026)
by: Mandal, Avilasha, et al.
Published: (2026)
Mitigating Pretraining-Induced Attention Asymmetry in 2D+ Electron Microscopy Image Segmentation
by: Molnár, Zsófia, et al.
Published: (2025)
by: Molnár, Zsófia, et al.
Published: (2025)
VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold
by: Maggio, Dominic, et al.
Published: (2025)
by: Maggio, Dominic, et al.
Published: (2025)
Post-Hoc MOTS: Exploring the Capabilities of Time-Symmetric Multi-Object Tracking
by: Szabó, Gergely, et al.
Published: (2024)
by: Szabó, Gergely, et al.
Published: (2024)
Stable Diffusion with Continuous-time Neural Network
by: Horvath, Andras
Published: (2024)
by: Horvath, Andras
Published: (2024)
Dense Semantic Matching with VGGT Prior
by: Yang, Songlin, et al.
Published: (2025)
by: Yang, Songlin, et al.
Published: (2025)
A Self-Supervised Method for Body Part Segmentation and Keypoint Detection of Rat Images
by: Kopácsi, László, et al.
Published: (2024)
by: Kopácsi, László, et al.
Published: (2024)
VGGT-Motion: Motion-Aware Calibration-Free Monocular SLAM for Long-Range Consistency
by: Xiong, Zhuang, et al.
Published: (2026)
by: Xiong, Zhuang, et al.
Published: (2026)
VGGT-SLAM 2.0: Real-time Dense Feed-forward Scene Reconstruction
by: Maggio, Dominic, et al.
Published: (2026)
by: Maggio, Dominic, et al.
Published: (2026)
NIS-SLAM: Neural Implicit Semantic RGB-D SLAM for 3D Consistent Scene Understanding
by: Zhai, Hongjia, et al.
Published: (2024)
by: Zhai, Hongjia, et al.
Published: (2024)
VGGT-Det: Mining VGGT Internal Priors for Sensor-Geometry-Free Multi-View Indoor 3D Object Detection
by: Cao, Yang, et al.
Published: (2026)
by: Cao, Yang, et al.
Published: (2026)
VGGT-$Ω$
by: Wang, Jianyuan, et al.
Published: (2026)
by: Wang, Jianyuan, et al.
Published: (2026)
Enhancing Apparent Personality Trait Analysis with Cross-Modal Embeddings
by: Fodor, Ádám, et al.
Published: (2024)
by: Fodor, Ádám, et al.
Published: (2024)
VGGT-World: Transforming VGGT into an Autoregressive Geometry World Model
by: Sun, Xiangyu, et al.
Published: (2026)
by: Sun, Xiangyu, et al.
Published: (2026)
VGGT-X: When VGGT Meets Dense Novel View Synthesis
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
VGGT-MPR: VGGT-Enhanced Multimodal Place Recognition in Autonomous Driving Environments
by: Xu, Jingyi, et al.
Published: (2026)
by: Xu, Jingyi, et al.
Published: (2026)
FrameVGGT: Geometry-Aligned Frame-Level Memory for Bounded Streaming VGGT
by: Xu, Zhisong, et al.
Published: (2026)
by: Xu, Zhisong, et al.
Published: (2026)
RGBDS-SLAM: A RGB-D Semantic Dense SLAM Based on 3D Multi Level Pyramid Gaussian Splatting
by: Cao, Zhenzhong, et al.
Published: (2024)
by: Cao, Zhenzhong, et al.
Published: (2024)
S-VGGT: Structure-Aware Subscene Decomposition for Scalable 3D Foundation Models
by: Li, Xinze, et al.
Published: (2026)
by: Li, Xinze, et al.
Published: (2026)
PanoVGGT: Feed-Forward 3D Reconstruction from Panoramic Imagery
by: Guo, Yijing, et al.
Published: (2026)
by: Guo, Yijing, et al.
Published: (2026)
LiteVGGT: Boosting Vanilla VGGT via Geometry-aware Cached Token Merging
by: Shu, Zhijian, et al.
Published: (2025)
by: Shu, Zhijian, et al.
Published: (2025)
NEDS-SLAM: A Neural Explicit Dense Semantic SLAM Framework using 3D Gaussian Splatting
by: Ji, Yiming, et al.
Published: (2024)
by: Ji, Yiming, et al.
Published: (2024)
VGGT-Long: Chunk it, Loop it, Align it -- Pushing VGGT's Limits on Kilometer-scale Long RGB Sequences
by: Deng, Kai, et al.
Published: (2025)
by: Deng, Kai, et al.
Published: (2025)
V3D-SLAM: Robust RGB-D SLAM in Dynamic Environments with 3D Semantic Geometry Voting
by: Dang, Tuan, et al.
Published: (2024)
by: Dang, Tuan, et al.
Published: (2024)
GS3LAM: Gaussian Semantic Splatting SLAM
by: Li, Linfei, et al.
Published: (2026)
by: Li, Linfei, et al.
Published: (2026)
VGGT: Visual Geometry Grounded Transformer
by: Wang, Jianyuan, et al.
Published: (2025)
by: Wang, Jianyuan, et al.
Published: (2025)
SegVGGT: Joint 3D Reconstruction and Instance Segmentation from Multi-View Images
by: Qu, Jinyuan, et al.
Published: (2026)
by: Qu, Jinyuan, et al.
Published: (2026)
VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction
by: Chen, Xun, et al.
Published: (2026)
by: Chen, Xun, et al.
Published: (2026)
An Evaluation of DUSt3R/MASt3R/VGGT 3D Reconstruction on Photogrammetric Aerial Blocks
by: Wu, Xinyi, et al.
Published: (2025)
by: Wu, Xinyi, et al.
Published: (2025)
GS-SLAM: Dense Visual SLAM with 3D Gaussian Splatting
by: Yan, Chi, et al.
Published: (2023)
by: Yan, Chi, et al.
Published: (2023)
Taming the Light: Illumination-Invariant Semantic 3DGS-SLAM
by: Zhang, Shouhe, et al.
Published: (2025)
by: Zhang, Shouhe, et al.
Published: (2025)
4DLangVGGT: 4D Language-Visual Geometry Grounded Transformer
by: Wu, Xianfeng, et al.
Published: (2025)
by: Wu, Xianfeng, et al.
Published: (2025)
SemGauss-SLAM: Dense Semantic Gaussian Splatting SLAM
by: Zhu, Siting, et al.
Published: (2024)
by: Zhu, Siting, et al.
Published: (2024)
Review of Feed-forward 3D Reconstruction: From DUSt3R to VGGT
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
VGGT-CD: Training-Free Robust Registration for 3D Change Detection
by: Zhang, Wei, et al.
Published: (2026)
by: Zhang, Wei, et al.
Published: (2026)
EndoVGGT: GNN-Enhanced Depth Estimation for Surgical 3D Reconstruction
by: Fan, Falong, et al.
Published: (2026)
by: Fan, Falong, et al.
Published: (2026)
AVGGT: Rethinking Global Attention for Accelerating VGGT
by: Sun, Xianbing, et al.
Published: (2025)
by: Sun, Xianbing, et al.
Published: (2025)
Similar Items
-
SAMannot: A Memory-Efficient, Local, Open-source Framework for Interactive Video Instance Segmentation based on SAM2
by: Dinya, Gergely, et al.
Published: (2026) -
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
by: Gelencsér-Horváth, Anna, et al.
Published: (2026) -
Automatic camera orientation estimation for a partially calibrated camera above a plane with a line at known planar distance
by: Dinya, Gergely, et al.
Published: (2025) -
VGGT-SLAM++
by: Mandal, Avilasha, et al.
Published: (2026) -
Mitigating Pretraining-Induced Attention Asymmetry in 2D+ Electron Microscopy Image Segmentation
by: Molnár, Zsófia, et al.
Published: (2025)