Compositional Semantics for Open Vocabulary Spatio-semantic Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Karlsson, Robin, Lepe-Salazar, Francisco, Takeda, Kazuya |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
by: Lemke, Oliver, et al.
Published: (2024)
by: Lemke, Oliver, et al.
Published: (2024)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
by: Hu, X., et al.
Published: (2025)
by: Hu, X., et al.
Published: (2025)
PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution
by: Deichler, Anna
Published: (2026)
by: Deichler, Anna
Published: (2026)
A Survey of Spatial Memory Representations for Efficient Robot Navigation
by: Pangaliman, Ma. Madecheen S., et al.
Published: (2026)
by: Pangaliman, Ma. Madecheen S., et al.
Published: (2026)
RAE-NWM: Navigation World Model in Dense Visual Representation Space
by: Zhang, Mingkun, et al.
Published: (2026)
by: Zhang, Mingkun, et al.
Published: (2026)
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment
by: Kang, Li, et al.
Published: (2026)
by: Kang, Li, et al.
Published: (2026)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
by: Tourani, Ali, et al.
Published: (2024)
by: Tourani, Ali, et al.
Published: (2024)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
by: Mahdian, Navid, et al.
Published: (2024)
by: Mahdian, Navid, et al.
Published: (2024)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
by: Li, Xinting, et al.
Published: (2023)
by: Li, Xinting, et al.
Published: (2023)
VIN-NBV: A View Introspection Network for Next-Best-View Selection
by: Frahm, Noah, et al.
Published: (2025)
by: Frahm, Noah, et al.
Published: (2025)
PRISM-TopoMap: Online Topological Mapping with Place Recognition and Scan Matching
by: Muravyev, Kirill, et al.
Published: (2024)
by: Muravyev, Kirill, et al.
Published: (2024)
M3CAD: Towards Generic Cooperative Autonomous Driving Benchmark
by: Zhu, Morui, et al.
Published: (2025)
by: Zhu, Morui, et al.
Published: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
by: Li, Danyang, et al.
Published: (2025)
by: Li, Danyang, et al.
Published: (2025)
Transformers for Image-Goal Navigation
by: Pelluri, Nikhilanj
Published: (2024)
by: Pelluri, Nikhilanj
Published: (2024)
Single-Shot Metric Depth from Focused Plenoptic Cameras
by: Lasheras-Hernandez, Blanca, et al.
Published: (2024)
by: Lasheras-Hernandez, Blanca, et al.
Published: (2024)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
by: Özeren, Enes, et al.
Published: (2025)
by: Özeren, Enes, et al.
Published: (2025)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
by: Xiao, Jiasong, et al.
Published: (2026)
by: Xiao, Jiasong, et al.
Published: (2026)
Multi-modal Loop Closure Detection with Foundation Models in Severely Unstructured Environments
by: Gonzalez, Laura Alejandra Encinar, et al.
Published: (2025)
by: Gonzalez, Laura Alejandra Encinar, et al.
Published: (2025)
SelfOccFlow: Towards end-to-end self-supervised 3D Occupancy Flow prediction
by: Timoneda, Xavier, et al.
Published: (2026)
by: Timoneda, Xavier, et al.
Published: (2026)
Unsupervised Decomposition and Recombination with Discriminator-Driven Diffusion Models
by: Wang, Archer, et al.
Published: (2026)
by: Wang, Archer, et al.
Published: (2026)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
by: Mansour, Jad, et al.
Published: (2025)
by: Mansour, Jad, et al.
Published: (2025)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
by: Mansour, Jad, et al.
Published: (2024)
by: Mansour, Jad, et al.
Published: (2024)
Leveraging GNSS and Onboard Visual Data from Consumer Vehicles for Robust Road Network Estimation
by: Opra, Balázs, et al.
Published: (2024)
by: Opra, Balázs, et al.
Published: (2024)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
by: Tourani, Ali, et al.
Published: (2025)
by: Tourani, Ali, et al.
Published: (2025)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
by: Syed, Shahram Najam, et al.
Published: (2025)
by: Syed, Shahram Najam, et al.
Published: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
by: Mehta, Vinit, et al.
Published: (2025)
by: Mehta, Vinit, et al.
Published: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
by: Tourani, Ali, et al.
Published: (2025)
by: Tourani, Ali, et al.
Published: (2025)
The Impact of 2D Segmentation Backbones on Point Cloud Predictions Using 4D Radar
by: Muckelroy III, William, et al.
Published: (2025)
by: Muckelroy III, William, et al.
Published: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
by: Zorina, Kateryna, et al.
Published: (2026)
by: Zorina, Kateryna, et al.
Published: (2026)
Bayesian Data Augmentation and Training for Perception DNN in Autonomous Aerial Vehicles
by: Rasul, Ashik E, et al.
Published: (2024)
by: Rasul, Ashik E, et al.
Published: (2024)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
by: Riva, Paolo, et al.
Published: (2026)
by: Riva, Paolo, et al.
Published: (2026)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
by: He, Yuankai, et al.
Published: (2025)
by: He, Yuankai, et al.
Published: (2025)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
by: Käs, Stephanie, et al.
Published: (2025)
by: Käs, Stephanie, et al.
Published: (2025)
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
by: Ziakas, Christos, et al.
Published: (2025)
by: Ziakas, Christos, et al.
Published: (2025)
CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models
by: Liu, Zhi
Published: (2026)
by: Liu, Zhi
Published: (2026)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
by: Tourani, Ali, et al.
Published: (2023)
by: Tourani, Ali, et al.
Published: (2023)
Towards Global Localization using Multi-Modal Object-Instance Re-Identification
by: Chavan, Aneesh, et al.
Published: (2024)
by: Chavan, Aneesh, et al.
Published: (2024)
NavTopo: Leveraging Topological Maps For Autonomous Navigation Of a Mobile Robot
by: Muravyev, Kirill, et al.
Published: (2024)
by: Muravyev, Kirill, et al.
Published: (2024)
Real-Time Localization Framework for Autonomous Basketball Robots
by: Medarametla, Naren, et al.
Published: (2026)
by: Medarametla, Naren, et al.
Published: (2026)
MAP: End-to-End Autonomous Driving with Map-Assisted Planning
by: Yin, Huilin, et al.
Published: (2025)
by: Yin, Huilin, et al.
Published: (2025)
Similar Items
-
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
by: Lemke, Oliver, et al.
Published: (2024) -
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
by: Hu, X., et al.
Published: (2025) -
PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution
by: Deichler, Anna
Published: (2026) -
A Survey of Spatial Memory Representations for Efficient Robot Navigation
by: Pangaliman, Ma. Madecheen S., et al.
Published: (2026) -
RAE-NWM: Navigation World Model in Dense Visual Representation Space
by: Zhang, Mingkun, et al.
Published: (2026)