RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nasser, Zaid, Iumanov, Mikhail, Li, Tianhao, Popov, Maxim, Mahmoud, Jaafar, Kolyubin, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KM-ViPE: Online Tightly Coupled Vision-Language-Geometry Fusion for Open-Vocabulary Semantic SLAM
von: Nasser, Zaid, et al.
Veröffentlicht: (2025)
von: Nasser, Zaid, et al.
Veröffentlicht: (2025)
OSMa-Bench: Evaluating Open Semantic Mapping Under Varying Lighting Conditions
von: Popov, Maxim, et al.
Veröffentlicht: (2025)
von: Popov, Maxim, et al.
Veröffentlicht: (2025)
OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes
von: Kurkova, Regina, et al.
Veröffentlicht: (2026)
von: Kurkova, Regina, et al.
Veröffentlicht: (2026)
R5DGS: Semantic-Aware 4D Gaussian Splatting with Rigid Body Constraints for Efficient Dynamic Scene Reconstruction
von: Gridusov, Denis, et al.
Veröffentlicht: (2026)
von: Gridusov, Denis, et al.
Veröffentlicht: (2026)
ViPE: Video Pose Engine for 3D Geometric Perception
von: Huang, Jiahui, et al.
Veröffentlicht: (2025)
von: Huang, Jiahui, et al.
Veröffentlicht: (2025)
AgentGrounder: Zero-Shot 3D Visual Pointcloud Grounding using Multimodal Language Models
von: Huynh, Cuong, et al.
Veröffentlicht: (2026)
von: Huynh, Cuong, et al.
Veröffentlicht: (2026)
Open-Vocabulary Online Semantic Mapping for SLAM
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
Tightly Coupled SLAM with Imprecise Architectural Plans
von: Shaheer, Muhammad, et al.
Veröffentlicht: (2024)
von: Shaheer, Muhammad, et al.
Veröffentlicht: (2024)
LEG-SLAM: Real-Time Language-Enhanced Gaussian Splatting for SLAM
von: Titkov, Roman, et al.
Veröffentlicht: (2025)
von: Titkov, Roman, et al.
Veröffentlicht: (2025)
GeoFlow-SLAM: A Robust Tightly-Coupled RGBD-Inertial and Legged Odometry Fusion SLAM for Dynamic Legged Robotics
von: Xiao, Tingyang, et al.
Veröffentlicht: (2025)
von: Xiao, Tingyang, et al.
Veröffentlicht: (2025)
AQUA-SLAM: Tightly-Coupled Underwater Acoustic-Visual-Inertial SLAM with Sensor Calibration
von: Xu, Shida, et al.
Veröffentlicht: (2025)
von: Xu, Shida, et al.
Veröffentlicht: (2025)
ViMGuard: A Novel Multi-Modal System for Video Misinformation Guarding
von: Kan, Andrew, et al.
Veröffentlicht: (2024)
von: Kan, Andrew, et al.
Veröffentlicht: (2024)
Semantic Library Adaptation: LoRA Retrieval and Fusion for Open-Vocabulary Semantic Segmentation
von: Qorbani, Reza, et al.
Veröffentlicht: (2025)
von: Qorbani, Reza, et al.
Veröffentlicht: (2025)
Tight and Efficient Upper Bound on Spectral Norm of Convolutional Layers
von: Grishina, Ekaterina, et al.
Veröffentlicht: (2024)
von: Grishina, Ekaterina, et al.
Veröffentlicht: (2024)
Synergizing Morphological Computation and Generative Design: Automatic Synthesis of Tendon-Driven Grippers
von: Zharkov, Kirill, et al.
Veröffentlicht: (2024)
von: Zharkov, Kirill, et al.
Veröffentlicht: (2024)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
von: Reichard, Klara, et al.
Veröffentlicht: (2025)
von: Reichard, Klara, et al.
Veröffentlicht: (2025)
OpenMonoGS-SLAM: Monocular Gaussian Splatting SLAM with Open-set Semantics
von: Yoo, Jisang, et al.
Veröffentlicht: (2025)
von: Yoo, Jisang, et al.
Veröffentlicht: (2025)
VISTA: Open-Vocabulary, Task-Relevant Robot Exploration with Online Semantic Gaussian Splatting
von: Nagami, Keiko, et al.
Veröffentlicht: (2025)
von: Nagami, Keiko, et al.
Veröffentlicht: (2025)
Category-Adaptive Cross-Modal Semantic Refinement and Transfer for Open-Vocabulary Multi-Label Recognition
von: Liu, Haijing, et al.
Veröffentlicht: (2024)
von: Liu, Haijing, et al.
Veröffentlicht: (2024)
ViSTA-SLAM: Visual SLAM with Symmetric Two-view Association
von: Zhang, Ganlin, et al.
Veröffentlicht: (2025)
von: Zhang, Ganlin, et al.
Veröffentlicht: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
Tightly-Coupled LiDAR-Visual-Inertial SLAM and Large-Scale Volumetric Occupancy Mapping
von: Boche, Simon, et al.
Veröffentlicht: (2024)
von: Boche, Simon, et al.
Veröffentlicht: (2024)
RSV-SLAM: Toward Real-Time Semantic Visual SLAM in Indoor Dynamic Environments
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
Beyond Bare Queries: Open-Vocabulary Object Grounding with 3D Scene Graph
von: Linok, Sergey, et al.
Veröffentlicht: (2024)
von: Linok, Sergey, et al.
Veröffentlicht: (2024)
Open-Vocabulary Semantic Segmentation with Uncertainty Alignment for Robotic Scene Understanding in Indoor Building Environments
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
von: Xu, Yifan, et al.
Veröffentlicht: (2025)
OVMR: Open-Vocabulary Recognition with Multi-Modal References
von: Ma, Zehong, et al.
Veröffentlicht: (2024)
von: Ma, Zehong, et al.
Veröffentlicht: (2024)
Exploring Simple Open-Vocabulary Semantic Segmentation
von: Lai, Zihang
Veröffentlicht: (2024)
von: Lai, Zihang
Veröffentlicht: (2024)
Open Vocabulary Semantic Scene Sketch Understanding
von: Bourouis, Ahmed, et al.
Veröffentlicht: (2023)
von: Bourouis, Ahmed, et al.
Veröffentlicht: (2023)
Open-Vocabulary Segmentation with Semantic-Assisted Calibration
von: Liu, Yong, et al.
Veröffentlicht: (2023)
von: Liu, Yong, et al.
Veröffentlicht: (2023)
Open-Vocabulary Audio-Visual Semantic Segmentation
von: Guo, Ruohao, et al.
Veröffentlicht: (2024)
von: Guo, Ruohao, et al.
Veröffentlicht: (2024)
Towards Open-Vocabulary Video Semantic Segmentation
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
von: Li, Xinhao, et al.
Veröffentlicht: (2024)
Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation
von: Truong, Hoang M., et al.
Veröffentlicht: (2026)
von: Truong, Hoang M., et al.
Veröffentlicht: (2026)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
von: Shin, Heeseong, et al.
Veröffentlicht: (2024)
von: Shin, Heeseong, et al.
Veröffentlicht: (2024)
Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
GSplatLoc: Grounding Keypoint Descriptors into 3D Gaussian Splatting for Improved Visual Localization
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
von: Sidorov, Gennady, et al.
Veröffentlicht: (2024)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
Opti-Acoustic Semantic SLAM with Unknown Objects in Underwater Environments
von: Singh, Kurran, et al.
Veröffentlicht: (2024)
von: Singh, Kurran, et al.
Veröffentlicht: (2024)
OpenIN: Open-Vocabulary Instance-Oriented Navigation in Dynamic Domestic Environments
von: Tang, Yujie, et al.
Veröffentlicht: (2025)
von: Tang, Yujie, et al.
Veröffentlicht: (2025)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KM-ViPE: Online Tightly Coupled Vision-Language-Geometry Fusion for Open-Vocabulary Semantic SLAM
von: Nasser, Zaid, et al.
Veröffentlicht: (2025) -
OSMa-Bench: Evaluating Open Semantic Mapping Under Varying Lighting Conditions
von: Popov, Maxim, et al.
Veröffentlicht: (2025) -
OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes
von: Kurkova, Regina, et al.
Veröffentlicht: (2026) -
R5DGS: Semantic-Aware 4D Gaussian Splatting with Rigid Body Constraints for Efficient Dynamic Scene Reconstruction
von: Gridusov, Denis, et al.
Veröffentlicht: (2026) -
ViPE: Video Pose Engine for 3D Geometric Perception
von: Huang, Jiahui, et al.
Veröffentlicht: (2025)