SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Gelencsér-Horváth, Anna, Dinya, Gergely, Erős, Dorka Boglárka, Halász, Péter, Muqsit, Islam Muhammad, Karacs, Kristóf |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Building temporally coherent 3D maps with VGGT for memory-efficient Semantic SLAM
di: Dinya, Gergely, et al.
Pubblicazione: (2025)
di: Dinya, Gergely, et al.
Pubblicazione: (2025)
SAMannot: A Memory-Efficient, Local, Open-source Framework for Interactive Video Instance Segmentation based on SAM2
di: Dinya, Gergely, et al.
Pubblicazione: (2026)
di: Dinya, Gergely, et al.
Pubblicazione: (2026)
Automatic camera orientation estimation for a partially calibrated camera above a plane with a line at known planar distance
di: Dinya, Gergely, et al.
Pubblicazione: (2025)
di: Dinya, Gergely, et al.
Pubblicazione: (2025)
VGGT-SLAM++
di: Mandal, Avilasha, et al.
Pubblicazione: (2026)
di: Mandal, Avilasha, et al.
Pubblicazione: (2026)
VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold
di: Maggio, Dominic, et al.
Pubblicazione: (2025)
di: Maggio, Dominic, et al.
Pubblicazione: (2025)
VGGT-SLAM 2.0: Real-time Dense Feed-forward Scene Reconstruction
di: Maggio, Dominic, et al.
Pubblicazione: (2026)
di: Maggio, Dominic, et al.
Pubblicazione: (2026)
VGGT-$Ω$
di: Wang, Jianyuan, et al.
Pubblicazione: (2026)
di: Wang, Jianyuan, et al.
Pubblicazione: (2026)
VGGT-World: Transforming VGGT into an Autoregressive Geometry World Model
di: Sun, Xiangyu, et al.
Pubblicazione: (2026)
di: Sun, Xiangyu, et al.
Pubblicazione: (2026)
VGGT-X: When VGGT Meets Dense Novel View Synthesis
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
VGGT-MPR: VGGT-Enhanced Multimodal Place Recognition in Autonomous Driving Environments
di: Xu, Jingyi, et al.
Pubblicazione: (2026)
di: Xu, Jingyi, et al.
Pubblicazione: (2026)
FrameVGGT: Geometry-Aligned Frame-Level Memory for Bounded Streaming VGGT
di: Xu, Zhisong, et al.
Pubblicazione: (2026)
di: Xu, Zhisong, et al.
Pubblicazione: (2026)
Assessing the generalization performance of SAM for ureteroscopy scene understanding
di: Villagrana, Martin, et al.
Pubblicazione: (2025)
di: Villagrana, Martin, et al.
Pubblicazione: (2025)
LiteVGGT: Boosting Vanilla VGGT via Geometry-aware Cached Token Merging
di: Shu, Zhijian, et al.
Pubblicazione: (2025)
di: Shu, Zhijian, et al.
Pubblicazione: (2025)
Enhancing Cell Tracking with a Time-Symmetric Deep Learning Approach
di: Szabó, Gergely, et al.
Pubblicazione: (2023)
di: Szabó, Gergely, et al.
Pubblicazione: (2023)
VGGT-Motion: Motion-Aware Calibration-Free Monocular SLAM for Long-Range Consistency
di: Xiong, Zhuang, et al.
Pubblicazione: (2026)
di: Xiong, Zhuang, et al.
Pubblicazione: (2026)
ACE-SLAM: Scene Coordinate Regression for Neural Implicit Real-Time SLAM
di: Alzugaray, Ignacio, et al.
Pubblicazione: (2025)
di: Alzugaray, Ignacio, et al.
Pubblicazione: (2025)
VGGT-Long: Chunk it, Loop it, Align it -- Pushing VGGT's Limits on Kilometer-scale Long RGB Sequences
di: Deng, Kai, et al.
Pubblicazione: (2025)
di: Deng, Kai, et al.
Pubblicazione: (2025)
VGGT: Visual Geometry Grounded Transformer
di: Wang, Jianyuan, et al.
Pubblicazione: (2025)
di: Wang, Jianyuan, et al.
Pubblicazione: (2025)
Dense Semantic Matching with VGGT Prior
di: Yang, Songlin, et al.
Pubblicazione: (2025)
di: Yang, Songlin, et al.
Pubblicazione: (2025)
GPA-VGGT:Adapting VGGT to Large Scale Localization by Self-Supervised Learning with Geometry and Physics Aware Loss
di: Xu, Yangfan, et al.
Pubblicazione: (2026)
di: Xu, Yangfan, et al.
Pubblicazione: (2026)
AVGGT: Rethinking Global Attention for Accelerating VGGT
di: Sun, Xianbing, et al.
Pubblicazione: (2025)
di: Sun, Xianbing, et al.
Pubblicazione: (2025)
VGGT-Det: Mining VGGT Internal Priors for Sensor-Geometry-Free Multi-View Indoor 3D Object Detection
di: Cao, Yang, et al.
Pubblicazione: (2026)
di: Cao, Yang, et al.
Pubblicazione: (2026)
Supertoroid fitting of objects with holes for robotic grasping and scene generation
di: Torres, Joan Badia, et al.
Pubblicazione: (2024)
di: Torres, Joan Badia, et al.
Pubblicazione: (2024)
SwiftVGGT: A Scalable Visual Geometry Grounded Transformer for Large-Scale Scenes
di: Lee, Jungho, et al.
Pubblicazione: (2025)
di: Lee, Jungho, et al.
Pubblicazione: (2025)
Hilti SLAM Challenge 2023: Benchmarking Single + Multi-session SLAM across Sensor Constellations in Construction
di: Nair, Ashish Devadas, et al.
Pubblicazione: (2024)
di: Nair, Ashish Devadas, et al.
Pubblicazione: (2024)
Unblur-SLAM: Dense Neural SLAM for Blurry Inputs
di: Zhang, Qi, et al.
Pubblicazione: (2026)
di: Zhang, Qi, et al.
Pubblicazione: (2026)
VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation
di: Gao, Yulu, et al.
Pubblicazione: (2026)
di: Gao, Yulu, et al.
Pubblicazione: (2026)
HD-VGGT: High-Resolution Visual Geometry Transformer
di: Chen, Tianrun, et al.
Pubblicazione: (2026)
di: Chen, Tianrun, et al.
Pubblicazione: (2026)
DynamicVGGT: Learning Dynamic Point Maps for 4D Scene Reconstruction in Autonomous Driving
di: He, Zhuolin, et al.
Pubblicazione: (2026)
di: He, Zhuolin, et al.
Pubblicazione: (2026)
VGGT-Edit: Feed-forward Native 3D Scene Editing with Residual Field Prediction
di: Zhu, Kaixin, et al.
Pubblicazione: (2026)
di: Zhu, Kaixin, et al.
Pubblicazione: (2026)
Language-EXtended Indoor SLAM (LEXIS): A Versatile System for Real-time Visual Scene Understanding
di: Kassab, Christina, et al.
Pubblicazione: (2023)
di: Kassab, Christina, et al.
Pubblicazione: (2023)
An event-based implementation of saliency-based visual attention for rapid scene analysis
di: Chane, Camille Simon, et al.
Pubblicazione: (2024)
di: Chane, Camille Simon, et al.
Pubblicazione: (2024)
Compressed learning based onboard semantic compression for remote sensing platforms
di: Bhattacharjee, Protim, et al.
Pubblicazione: (2024)
di: Bhattacharjee, Protim, et al.
Pubblicazione: (2024)
Reconfigurable, large-format D-ToF/photon-counting SPAD image sensors with embedded FPGA for scene adaptability
di: Milanese, Tommaso, et al.
Pubblicazione: (2025)
di: Milanese, Tommaso, et al.
Pubblicazione: (2025)
VGGT4D: Mining Motion Cues in Visual Geometry Transformers for 4D Scene Reconstruction
di: Hu, Yu, et al.
Pubblicazione: (2025)
di: Hu, Yu, et al.
Pubblicazione: (2025)
InfiniteVGGT: Visual Geometry Grounded Transformer for Endless Streams
di: Yuan, Shuai, et al.
Pubblicazione: (2026)
di: Yuan, Shuai, et al.
Pubblicazione: (2026)
HTTM: Head-wise Temporal Token Merging for Faster VGGT
di: Wang, Weitian, et al.
Pubblicazione: (2025)
di: Wang, Weitian, et al.
Pubblicazione: (2025)
FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
di: Shen, You, et al.
Pubblicazione: (2025)
di: Shen, You, et al.
Pubblicazione: (2025)
HeSS: Head Sensitivity Score for Sparsity Redistribution in VGGT
di: Kim, Yongsung, et al.
Pubblicazione: (2026)
di: Kim, Yongsung, et al.
Pubblicazione: (2026)
Reloc-VGGT: Visual Re-localization with Geometry Grounded Transformer
di: Deng, Tianchen, et al.
Pubblicazione: (2025)
di: Deng, Tianchen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Building temporally coherent 3D maps with VGGT for memory-efficient Semantic SLAM
di: Dinya, Gergely, et al.
Pubblicazione: (2025) -
SAMannot: A Memory-Efficient, Local, Open-source Framework for Interactive Video Instance Segmentation based on SAM2
di: Dinya, Gergely, et al.
Pubblicazione: (2026) -
Automatic camera orientation estimation for a partially calibrated camera above a plane with a line at known planar distance
di: Dinya, Gergely, et al.
Pubblicazione: (2025) -
VGGT-SLAM++
di: Mandal, Avilasha, et al.
Pubblicazione: (2026) -
VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold
di: Maggio, Dominic, et al.
Pubblicazione: (2025)