DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Ziren, Li, Xiaohan, Tosi, Fabio, Zhang, Youmin, Mattoccia, Stefano, Wu, Jun, Poggi, Matteo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HS-SLAM: Hybrid Representation with Structural Supervision for Improved Dense SLAM
by: Gong, Ziren, et al.
Published: (2025)
by: Gong, Ziren, et al.
Published: (2025)
How NeRFs and 3D Gaussian Splatting are Reshaping SLAM: a Survey
by: Tosi, Fabio, et al.
Published: (2024)
by: Tosi, Fabio, et al.
Published: (2024)
Ov3R: Open-Vocabulary Semantic 3D Reconstruction from RGB Videos
by: Gong, Ziren, et al.
Published: (2025)
by: Gong, Ziren, et al.
Published: (2025)
Stereo 3D Gaussian Splatting SLAM for Outdoor Urban Scenes
by: Li, Xiaohan, et al.
Published: (2025)
by: Li, Xiaohan, et al.
Published: (2025)
Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail
by: Bartolomei, Luca, et al.
Published: (2024)
by: Bartolomei, Luca, et al.
Published: (2024)
FoundationSLAM: Unleashing the Power of Depth Foundation Models for End-to-End Dense Visual SLAM
by: Wu, Yuchen, et al.
Published: (2025)
by: Wu, Yuchen, et al.
Published: (2025)
D$^3$FlowSLAM: Self-Supervised Dynamic SLAM with Flow Motion Decomposition and DINO Guidance
by: Yu, Xingyuan, et al.
Published: (2022)
by: Yu, Xingyuan, et al.
Published: (2022)
GlORIE-SLAM: Globally Optimized RGB-only Implicit Encoding Point Cloud SLAM
by: Zhang, Ganlin, et al.
Published: (2024)
by: Zhang, Ganlin, et al.
Published: (2024)
ToF-Splatting: Dense SLAM using Sparse Time-of-Flight Depth and Multi-Frame Integration
by: Conti, Andrea, et al.
Published: (2025)
by: Conti, Andrea, et al.
Published: (2025)
Depth AnyEvent: A Cross-Modal Distillation Paradigm for Event-Based Monocular Depth Estimation
by: Bartolomei, Luca, et al.
Published: (2025)
by: Bartolomei, Luca, et al.
Published: (2025)
EventHub: Data Factory for Generalizable Event-Based Stereo Networks without Active Sensors
by: Bartolomei, Luca, et al.
Published: (2026)
by: Bartolomei, Luca, et al.
Published: (2026)
Active Stereo in the Wild through Virtual Pattern Projection
by: Bartolomei, Luca, et al.
Published: (2024)
by: Bartolomei, Luca, et al.
Published: (2024)
NIS-SLAM: Neural Implicit Semantic RGB-D SLAM for 3D Consistent Scene Understanding
by: Zhai, Hongjia, et al.
Published: (2024)
by: Zhai, Hongjia, et al.
Published: (2024)
Benchmarking Implicit Neural Representation and Geometric Rendering in Real-Time RGB-D SLAM
by: Hua, Tongyan, et al.
Published: (2024)
by: Hua, Tongyan, et al.
Published: (2024)
DINO-Foresight: Looking into the Future with DINO
by: Karypidis, Efstathios, et al.
Published: (2024)
by: Karypidis, Efstathios, et al.
Published: (2024)
S3-SLAM: Sparse Tri-plane Encoding for Neural Implicit SLAM
by: Zhang, Zhiyao, et al.
Published: (2024)
by: Zhang, Zhiyao, et al.
Published: (2024)
DINO-Tok: Adapting DINO for Visual Tokenizers
by: Jia, Mingkai, et al.
Published: (2025)
by: Jia, Mingkai, et al.
Published: (2025)
NeB-SLAM: Neural Blocks-based Salable RGB-D SLAM for Unknown Scenes
by: Bai, Lizhi, et al.
Published: (2024)
by: Bai, Lizhi, et al.
Published: (2024)
GauS-SLAM: Dense RGB-D SLAM with Gaussian Surfels
by: Su, Yongxin, et al.
Published: (2025)
by: Su, Yongxin, et al.
Published: (2025)
MCN-SLAM: Multi-Agent Collaborative Neural SLAM with Hybrid Implicit Neural Scene Representation
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
RGB Guided ToF Imaging System: A Survey of Deep Learning-based Methods
by: Qiao, Xin, et al.
Published: (2024)
by: Qiao, Xin, et al.
Published: (2024)
FlowSeek: Optical Flow Made Easier with Depth Foundation Models and Motion Bases
by: Poggi, Matteo, et al.
Published: (2025)
by: Poggi, Matteo, et al.
Published: (2025)
Federated Online Adaptation for Deep Stereo
by: Poggi, Matteo, et al.
Published: (2024)
by: Poggi, Matteo, et al.
Published: (2024)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
by: Liu, Shilong, et al.
Published: (2023)
by: Liu, Shilong, et al.
Published: (2023)
Evaluating Stenosis Detection with Grounding DINO, YOLO, and DINO-DETR
by: Ansari, Muhammad Musab
Published: (2025)
by: Ansari, Muhammad Musab
Published: (2025)
Booster: a Benchmark for Depth from Images of Specular and Transparent Surfaces
by: Ramirez, Pierluigi Zama, et al.
Published: (2023)
by: Ramirez, Pierluigi Zama, et al.
Published: (2023)
DF-SLAM: Dictionary Factors Representation for High-Fidelity Neural Implicit Dense Visual SLAM System
by: Wei, Weifeng, et al.
Published: (2024)
by: Wei, Weifeng, et al.
Published: (2024)
EvenNICER-SLAM: Event-based Neural Implicit Encoding SLAM
by: Chen, Shi, et al.
Published: (2024)
by: Chen, Shi, et al.
Published: (2024)
Range-Agnostic Multi-View Depth Estimation With Keyframe Selection
by: Conti, Andrea, et al.
Published: (2024)
by: Conti, Andrea, et al.
Published: (2024)
LiDAR-Event Stereo Fusion with Hallucinations
by: Bartolomei, Luca, et al.
Published: (2024)
by: Bartolomei, Luca, et al.
Published: (2024)
Depth on Demand: Streaming Dense Depth from a Low Frame Rate Active Sensor
by: Conti, Andrea, et al.
Published: (2024)
by: Conti, Andrea, et al.
Published: (2024)
Implicit Event-RGBD Neural SLAM
by: Qu, Delin, et al.
Published: (2023)
by: Qu, Delin, et al.
Published: (2023)
Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
by: Sandström, Erik, et al.
Published: (2024)
by: Sandström, Erik, et al.
Published: (2024)
DDN-SLAM: Real-time Dense Dynamic Neural Implicit SLAM
by: Li, Mingrui, et al.
Published: (2024)
by: Li, Mingrui, et al.
Published: (2024)
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
by: Guo, Hao, et al.
Published: (2024)
by: Guo, Hao, et al.
Published: (2024)
PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training
by: Fu, Weifu, et al.
Published: (2026)
by: Fu, Weifu, et al.
Published: (2026)
Bidirectional Cross-Modal Prompting for Event-Frame Asymmetric Stereo
by: Xu, Ninghui, et al.
Published: (2026)
by: Xu, Ninghui, et al.
Published: (2026)
Deploy DINO with Many-to-Many Association
by: Jiang, Haodong, et al.
Published: (2026)
by: Jiang, Haodong, et al.
Published: (2026)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
by: Tumanyan, Narek, et al.
Published: (2024)
by: Tumanyan, Narek, et al.
Published: (2024)
SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3
by: Yang, Sicheng, et al.
Published: (2025)
by: Yang, Sicheng, et al.
Published: (2025)
Similar Items
-
HS-SLAM: Hybrid Representation with Structural Supervision for Improved Dense SLAM
by: Gong, Ziren, et al.
Published: (2025) -
How NeRFs and 3D Gaussian Splatting are Reshaping SLAM: a Survey
by: Tosi, Fabio, et al.
Published: (2024) -
Ov3R: Open-Vocabulary Semantic 3D Reconstruction from RGB Videos
by: Gong, Ziren, et al.
Published: (2025) -
Stereo 3D Gaussian Splatting SLAM for Outdoor Urban Scenes
by: Li, Xiaohan, et al.
Published: (2025) -
Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail
by: Bartolomei, Luca, et al.
Published: (2024)