Online Embedding Multi-Scale CLIP Features into 3D Maps
Fuente:
arXiv
Guardado en:
| Autores principales: | Taguchi, Shun, Deguchi, Hideki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Efficient and Distributed Large-Scale 3D Map Registration using Tomographic Features
por: Unlu, Halil Utku, et al.
Publicado: (2024)
por: Unlu, Halil Utku, et al.
Publicado: (2024)
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
por: Taguchi, Shun, et al.
Publicado: (2025)
por: Taguchi, Shun, et al.
Publicado: (2025)
CLIP-Loc: Multi-modal Landmark Association for Global Localization in Object-based Maps
por: Matsuzaki, Shigemichi, et al.
Publicado: (2024)
por: Matsuzaki, Shigemichi, et al.
Publicado: (2024)
CLIP feature-based randomized control using images and text for multiple tasks and robots
por: Shibata, Kazuki, et al.
Publicado: (2024)
por: Shibata, Kazuki, et al.
Publicado: (2024)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
por: Wang, Jiasen, et al.
Publicado: (2024)
por: Wang, Jiasen, et al.
Publicado: (2024)
LatentAM: Real-Time, Large-Scale Latent Gaussian Attention Mapping via Online Dictionary Learning
por: Lee, Junwoon, et al.
Publicado: (2026)
por: Lee, Junwoon, et al.
Publicado: (2026)
LIVE-GS: Online LiDAR-Inertial-Visual State Estimation and Globally Consistent Mapping with 3D Gaussian Splatting
por: Park, Jaeseok, et al.
Publicado: (2025)
por: Park, Jaeseok, et al.
Publicado: (2025)
Robotic-CLIP: Fine-tuning CLIP on Action Data for Robotic Applications
por: Nguyen, Nghia, et al.
Publicado: (2024)
por: Nguyen, Nghia, et al.
Publicado: (2024)
Online Knowledge Integration for 3D Semantic Mapping: A Survey
por: Igelbrink, Felix, et al.
Publicado: (2024)
por: Igelbrink, Felix, et al.
Publicado: (2024)
Multi-Agent 3D Map Reconstruction and Change Detection in Microgravity with Free-Flying Robots
por: Dinkel, Holly, et al.
Publicado: (2023)
por: Dinkel, Holly, et al.
Publicado: (2023)
MLFM: Multi-Layered Feature Maps for Richer Language Understanding in Zero-Shot Semantic Navigation
por: Raychaudhuri, Sonia, et al.
Publicado: (2025)
por: Raychaudhuri, Sonia, et al.
Publicado: (2025)
Open-Vocabulary Online Semantic Mapping for SLAM
por: Martins, Tomas Berriel, et al.
Publicado: (2024)
por: Martins, Tomas Berriel, et al.
Publicado: (2024)
Active Neural Mapping at Scale
por: Kuang, Zijia, et al.
Publicado: (2024)
por: Kuang, Zijia, et al.
Publicado: (2024)
SDTagNet: Leveraging Text-Annotated Navigation Maps for Online HD Map Construction
por: Immel, Fabian, et al.
Publicado: (2025)
por: Immel, Fabian, et al.
Publicado: (2025)
MapGCLR: Geospatial Contrastive Learning of Representations for Online Vectorized HD Map Construction
por: Merkert, Jonas, et al.
Publicado: (2026)
por: Merkert, Jonas, et al.
Publicado: (2026)
MapTRv2: An End-to-End Framework for Online Vectorized HD Map Construction
por: Liao, Bencheng, et al.
Publicado: (2023)
por: Liao, Bencheng, et al.
Publicado: (2023)
Language to Map: Topological map generation from natural language path instructions
por: Deguchi, Hideki, et al.
Publicado: (2024)
por: Deguchi, Hideki, et al.
Publicado: (2024)
RecNet: An Invertible Point Cloud Encoding through Range Image Embeddings for Multi-Robot Map Sharing and Reconstruction
por: Stathoulopoulos, Nikolaos, et al.
Publicado: (2024)
por: Stathoulopoulos, Nikolaos, et al.
Publicado: (2024)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
Accelerating Online Mapping and Behavior Prediction via Direct BEV Feature Attention
por: Gu, Xunjiang, et al.
Publicado: (2024)
por: Gu, Xunjiang, et al.
Publicado: (2024)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
por: Mahdavian, Mohammad, et al.
Publicado: (2026)
por: Mahdavian, Mohammad, et al.
Publicado: (2026)
Online Mapping for Autonomous Driving: Addressing Sensor Generalization and Dynamic Map Updates in Campus Environments
por: Zhang, Zihan, et al.
Publicado: (2025)
por: Zhang, Zihan, et al.
Publicado: (2025)
MapGS: Generalizable Pretraining and Data Augmentation for Online Mapping via Novel View Synthesis
por: Zhang, Hengyuan, et al.
Publicado: (2025)
por: Zhang, Hengyuan, et al.
Publicado: (2025)
SMR-Net:Robot Snap Detection Based on Multi-Scale Features and Self-Attention Network
por: Hou, Kuanxu
Publicado: (2026)
por: Hou, Kuanxu
Publicado: (2026)
Volumetric Semantically Consistent 3D Panoptic Mapping
por: Miao, Yang, et al.
Publicado: (2023)
por: Miao, Yang, et al.
Publicado: (2023)
3D Feature Distillation with Object-Centric Priors
por: Tziafas, Georgios, et al.
Publicado: (2024)
por: Tziafas, Georgios, et al.
Publicado: (2024)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
por: Jiang, Jiajun, et al.
Publicado: (2025)
por: Jiang, Jiajun, et al.
Publicado: (2025)
Mono-Hydra++: Real-Time Monocular Scene Graph Construction with Multi-Task Learning for 3D Indoor Mapping
por: Udugama, U. V. B. L., et al.
Publicado: (2026)
por: Udugama, U. V. B. L., et al.
Publicado: (2026)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
por: Sun, Haowen, et al.
Publicado: (2026)
por: Sun, Haowen, et al.
Publicado: (2026)
Lifelong 3D Mapping Framework for Hand-held & Robot-mounted LiDAR Mapping Systems
por: Yang, Liudi, et al.
Publicado: (2025)
por: Yang, Liudi, et al.
Publicado: (2025)
Enhancing Online Road Network Perception and Reasoning with Standard Definition Maps
por: Zhang, Hengyuan, et al.
Publicado: (2024)
por: Zhang, Hengyuan, et al.
Publicado: (2024)
Impact of Localization Errors on Label Quality for Online HD Map Construction
por: Blumberg, Alexander, et al.
Publicado: (2026)
por: Blumberg, Alexander, et al.
Publicado: (2026)
MoD-SLAM: Monocular Dense Mapping for Unbounded 3D Scene Reconstruction
por: Zhou, Heng, et al.
Publicado: (2024)
por: Zhou, Heng, et al.
Publicado: (2024)
SegVec3D: A Method for Vector Embedding of 3D Objects Oriented Towards Robot manipulation
por: Kang, Zhihan, et al.
Publicado: (2025)
por: Kang, Zhihan, et al.
Publicado: (2025)
VMGNet: A Low Computational Complexity Robotic Grasping Network Based on VMamba with Multi-Scale Feature Fusion
por: Jin, Yuhao, et al.
Publicado: (2024)
por: Jin, Yuhao, et al.
Publicado: (2024)
Online 3D Scene Reconstruction Using Neural Object Priors
por: Chabal, Thomas, et al.
Publicado: (2025)
por: Chabal, Thomas, et al.
Publicado: (2025)
123D: Unifying Multi-Modal Autonomous Driving Data at Scale
por: Dauner, Daniel, et al.
Publicado: (2026)
por: Dauner, Daniel, et al.
Publicado: (2026)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
por: Nanwani, Laksh, et al.
Publicado: (2024)
por: Nanwani, Laksh, et al.
Publicado: (2024)
OpenNavMap: Structure-Free Topometric Mapping via Large-Scale Collaborative Localization
por: Jiao, Jianhao, et al.
Publicado: (2026)
por: Jiao, Jianhao, et al.
Publicado: (2026)
NextBestPath: Efficient 3D Mapping of Unseen Environments
por: Li, Shiyao, et al.
Publicado: (2025)
por: Li, Shiyao, et al.
Publicado: (2025)
Ejemplares similares
-
Efficient and Distributed Large-Scale 3D Map Registration using Tomographic Features
por: Unlu, Halil Utku, et al.
Publicado: (2024) -
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
por: Taguchi, Shun, et al.
Publicado: (2025) -
CLIP-Loc: Multi-modal Landmark Association for Global Localization in Object-based Maps
por: Matsuzaki, Shigemichi, et al.
Publicado: (2024) -
CLIP feature-based randomized control using images and text for multiple tasks and robots
por: Shibata, Kazuki, et al.
Publicado: (2024) -
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
por: Wang, Jiasen, et al.
Publicado: (2024)