Real-Time 3D Vision-Language Embedding Mapping
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rauch, Christian, Ellensohn, Björn, Nwankwo, Linus, Dave, Vedant, Rueckert, Elmar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SIL: Symbiotic Interactive Learning for Language-Conditioned Human-Agent Co-Adaptation
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025)
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025)
ReLI: A Language-Agnostic Approach to Human-Robot Interaction
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025)
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025)
The Conversation is the Command: Interacting with Real-World Autonomous Robot Through Natural Language
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
Multimodal Human-Autonomous Agents Interaction Using Pre-Trained Language and Visual Foundation Models
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
EnvoDat: A Large-Scale Multisensory Dataset for Robotic Spatial Awareness and Semantic Reasoning in Heterogeneous Environments
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)
Understanding why SLAM algorithms fail in modern indoor environments
von: Linus, Nwankwo, et al.
Veröffentlicht: (2023)
von: Linus, Nwankwo, et al.
Veröffentlicht: (2023)
ROMR: A ROS-based Open-source Mobile Robot
von: Linus, Nwankwo, et al.
Veröffentlicht: (2022)
von: Linus, Nwankwo, et al.
Veröffentlicht: (2022)
Multimodal Visual-Tactile Representation Learning through Self-Supervised Contrastive Pre-Training
von: Dave, Vedant, et al.
Veröffentlicht: (2024)
von: Dave, Vedant, et al.
Veröffentlicht: (2024)
M2CURL: Sample-Efficient Multimodal Reinforcement Learning via Self-Supervised Representation Learning for Robotic Manipulation
von: Lygerakis, Fotios, et al.
Veröffentlicht: (2024)
von: Lygerakis, Fotios, et al.
Veröffentlicht: (2024)
SteelDS: A High-Resolution Video Dataset of E40 Steel Scrap for Object Detection and Instance Segmentation
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
PASTA: Vision Transformer Patch Aggregation for Weakly Supervised Target and Anomaly Segmentation
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
von: Feith, Nikolaus, et al.
Veröffentlicht: (2024)
von: Feith, Nikolaus, et al.
Veröffentlicht: (2024)
MAP-ADAPT: Real-Time Quality-Adaptive Semantic 3D Maps
von: Zheng, Jianhao, et al.
Veröffentlicht: (2024)
von: Zheng, Jianhao, et al.
Veröffentlicht: (2024)
Gleanmer: A 6 mW SoC for Real-Time 3D Gaussian Occupancy Mapping
von: Fu, Zih-Sing, et al.
Veröffentlicht: (2026)
von: Fu, Zih-Sing, et al.
Veröffentlicht: (2026)
Vision-Language Models on the Edge for Real-Time Robotic Perception
von: Ahmad, Sarat, et al.
Veröffentlicht: (2026)
von: Ahmad, Sarat, et al.
Veröffentlicht: (2026)
Probabilistic Height Grid Terrain Mapping for Mining Shovels using LiDAR
von: Bhandari, Vedant, et al.
Veröffentlicht: (2024)
von: Bhandari, Vedant, et al.
Veröffentlicht: (2024)
SLIM-VDB: A Real-Time 3D Probabilistic Semantic Mapping Framework
von: Sheppard, Anja, et al.
Veröffentlicht: (2025)
von: Sheppard, Anja, et al.
Veröffentlicht: (2025)
BlabberSeg: Real-Time Embedded Open-Vocabulary Aerial Segmentation
von: Bong, Haechan Mark, et al.
Veröffentlicht: (2024)
von: Bong, Haechan Mark, et al.
Veröffentlicht: (2024)
Real Time Collision Avoidance with GPU-Computed Distance Maps
von: Bedada, Wendwosen Bellete, et al.
Veröffentlicht: (2024)
von: Bedada, Wendwosen Bellete, et al.
Veröffentlicht: (2024)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
LIV-GaussMap: LiDAR-Inertial-Visual Fusion for Real-time 3D Radiance Field Map Rendering
von: Hong, Sheng, et al.
Veröffentlicht: (2024)
von: Hong, Sheng, et al.
Veröffentlicht: (2024)
Multi-Domain Motion Embedding: Expressive Real-Time Mimicry for Legged Robots
von: Heyrman, Matthias, et al.
Veröffentlicht: (2025)
von: Heyrman, Matthias, et al.
Veröffentlicht: (2025)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
Real-Time Planning Under Uncertainty for AUVs Using Virtual Maps
von: Collado-Gonzalez, Ivana, et al.
Veröffentlicht: (2024)
von: Collado-Gonzalez, Ivana, et al.
Veröffentlicht: (2024)
Can Pretrained Vision-Language Embeddings Alone Guide Robot Navigation?
von: Subedi, Nitesh, et al.
Veröffentlicht: (2025)
von: Subedi, Nitesh, et al.
Veröffentlicht: (2025)
Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Real-time 3D Semantic Scene Perception for Egocentric Robots with Binocular Vision
von: Nguyen, K., et al.
Veröffentlicht: (2024)
von: Nguyen, K., et al.
Veröffentlicht: (2024)
Online Embedding Multi-Scale CLIP Features into 3D Maps
von: Taguchi, Shun, et al.
Veröffentlicht: (2024)
von: Taguchi, Shun, et al.
Veröffentlicht: (2024)
Embedded Flexible Circumferential Sensing for Real-Time Intraoperative Environmental Perception in Continuum Robots
von: Luo, Peiyu, et al.
Veröffentlicht: (2025)
von: Luo, Peiyu, et al.
Veröffentlicht: (2025)
Splat-Nav: Safe Real-Time Robot Navigation in Gaussian Splatting Maps
von: Chen, Timothy, et al.
Veröffentlicht: (2024)
von: Chen, Timothy, et al.
Veröffentlicht: (2024)
MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2025)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2025)
Highly Deformable Proprioceptive Membrane for Real-Time 3D Shape Reconstruction
von: Xu, Guanyu, et al.
Veröffentlicht: (2026)
von: Xu, Guanyu, et al.
Veröffentlicht: (2026)
Xiaomi-Robotics-0: An Open-Sourced Vision-Language-Action Model with Real-Time Execution
von: Cai, Rui, et al.
Veröffentlicht: (2026)
von: Cai, Rui, et al.
Veröffentlicht: (2026)
Semantic Segmentation and Depth Estimation for Real-Time Lunar Surface Mapping Using 3D Gaussian Splatting
von: Vila, Guillem Casadesus, et al.
Veröffentlicht: (2026)
von: Vila, Guillem Casadesus, et al.
Veröffentlicht: (2026)
Real-Time Thermal-Inertial Odometry on Embedded Hardware for High-Speed GPS-Denied Flight
von: Stone, Austin, et al.
Veröffentlicht: (2026)
von: Stone, Austin, et al.
Veröffentlicht: (2026)
Towards Safe Autonomous Driving: A Real-Time Motion Planning Algorithm on Embedded Hardware
von: Moller, Korbinian, et al.
Veröffentlicht: (2026)
von: Moller, Korbinian, et al.
Veröffentlicht: (2026)
Characterizing and Optimizing Real-Time Optimal Control for Embedded SoCs
von: Dong, Kris Shengjun, et al.
Veröffentlicht: (2024)
von: Dong, Kris Shengjun, et al.
Veröffentlicht: (2024)
Range-SLAM: Ultra-Wideband-Based Smoke-Resistant Real-Time Localization and Mapping
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
Mesh-based Photorealistic and Real-time 3D Mapping for Robust Visual Perception of Autonomous Underwater Vehicle
von: Lee, Jungwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jungwoo, et al.
Veröffentlicht: (2024)
No Need for Real 3D: Fusing 2D Vision with Pseudo 3D Representations for Robotic Manipulation Learning
von: Yu, Run, et al.
Veröffentlicht: (2025)
von: Yu, Run, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SIL: Symbiotic Interactive Learning for Language-Conditioned Human-Agent Co-Adaptation
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025) -
ReLI: A Language-Agnostic Approach to Human-Robot Interaction
von: Nwankwo, Linus, et al.
Veröffentlicht: (2025) -
The Conversation is the Command: Interacting with Real-World Autonomous Robot Through Natural Language
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024) -
Multimodal Human-Autonomous Agents Interaction Using Pre-Trained Language and Visual Foundation Models
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024) -
EnvoDat: A Large-Scale Multisensory Dataset for Robotic Spatial Awareness and Semantic Reasoning in Heterogeneous Environments
von: Nwankwo, Linus, et al.
Veröffentlicht: (2024)