Leveraging Large Language Model-based Room-Object Relationships Knowledge for Enhancing Multimodal-Input Object Goal Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Leyuan, Kanezaki, Asako, Caron, Guillaume, Yoshiyasu, Yusuke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embodied Navigation with Auxiliary Task of Action Description Prediction
by: Kondoh, Haru, et al.
Published: (2025)
by: Kondoh, Haru, et al.
Published: (2025)
DiffSurf: A Transformer-based Diffusion Model for Generating and Reconstructing 3D Surfaces in Pose
by: Yoshiyasu, Yusuke, et al.
Published: (2024)
by: Yoshiyasu, Yusuke, et al.
Published: (2024)
TransFusionOdom: Interpretable Transformer-based LiDAR-Inertial Fusion Odometry Estimation
by: Sun, Leyuan, et al.
Published: (2023)
by: Sun, Leyuan, et al.
Published: (2023)
Zero-Shot Peg Insertion: Identifying Mating Holes and Estimating SE(2) Poses with Vision-Language Models
by: Yajima, Masaru, et al.
Published: (2025)
by: Yajima, Masaru, et al.
Published: (2025)
OP-Align: Object-level and Part-level Alignment for Self-supervised Category-level Articulated Object Pose Estimation
by: Che, Yuchen, et al.
Published: (2024)
by: Che, Yuchen, et al.
Published: (2024)
WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation
by: Nie, Dujun, et al.
Published: (2025)
by: Nie, Dujun, et al.
Published: (2025)
FOM-Nav: Frontier-Object Maps for Object Goal Navigation
by: Chabal, Thomas, et al.
Published: (2025)
by: Chabal, Thomas, et al.
Published: (2025)
Language-Based Augmentation to Address Shortcut Learning in Object Goal Navigation
by: Hoftijzer, Dennis, et al.
Published: (2024)
by: Hoftijzer, Dennis, et al.
Published: (2024)
Leveraging Unknown Objects to Construct Labeled-Unlabeled Meta-Relationships for Zero-Shot Object Navigation
by: Zheng, Yanwei, et al.
Published: (2024)
by: Zheng, Yanwei, et al.
Published: (2024)
MOPA: Modular Object Navigation with PointGoal Agents
by: Raychaudhuri, Sonia, et al.
Published: (2023)
by: Raychaudhuri, Sonia, et al.
Published: (2023)
Cognitive Planning for Object Goal Navigation using Generative AI Models
by: S, Arjun P, et al.
Published: (2024)
by: S, Arjun P, et al.
Published: (2024)
SR-Nav: Spatial Relationships Matter for Zero-shot Object Goal Navigation
by: Fang, Leyuan, et al.
Published: (2026)
by: Fang, Leyuan, et al.
Published: (2026)
MeshMamba: State Space Models for Articulated 3D Mesh Generation and Reconstruction
by: Yoshiyasu, Yusuke, et al.
Published: (2025)
by: Yoshiyasu, Yusuke, et al.
Published: (2025)
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
by: Wu, Pengying, et al.
Published: (2024)
by: Wu, Pengying, et al.
Published: (2024)
NeuralMeshing: Complete Object Mesh Extraction from Casual Captures
by: Erich, Floris, et al.
Published: (2025)
by: Erich, Floris, et al.
Published: (2025)
UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents
by: Xiao, Jianqiang, et al.
Published: (2025)
by: Xiao, Jianqiang, et al.
Published: (2025)
APEX: A Decoupled Memory-based Explorer for Asynchronous Aerial Object Goal Navigation
by: Zhang, Daoxuan, et al.
Published: (2026)
by: Zhang, Daoxuan, et al.
Published: (2026)
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)
by: Li, Badi, et al.
Published: (2025)
Zero-shot Degree of Ill-posedness Estimation for Active Small Object Change Detection
by: Takeda, Koji, et al.
Published: (2024)
by: Takeda, Koji, et al.
Published: (2024)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
by: Li, Xinting, et al.
Published: (2023)
by: Li, Xinting, et al.
Published: (2023)
COG: Confidence-aware Optimal Geometric Correspondence for Unsupervised Single-reference Novel Object Pose Estimation
by: Che, Yuchen, et al.
Published: (2026)
by: Che, Yuchen, et al.
Published: (2026)
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
by: Xu, Nuo, et al.
Published: (2024)
by: Xu, Nuo, et al.
Published: (2024)
CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs
by: Cao, Yihan, et al.
Published: (2024)
by: Cao, Yihan, et al.
Published: (2024)
TDANet: Target-Directed Attention Network For Object-Goal Visual Navigation With Zero-Shot Ability
by: Lian, Shiwei, et al.
Published: (2024)
by: Lian, Shiwei, et al.
Published: (2024)
Collision-Aware Object-Goal Visual Navigation via Two-Stage Deep Reinforcement Learning
by: Wang, Hongwu, et al.
Published: (2025)
by: Wang, Hongwu, et al.
Published: (2025)
LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation
by: Guan, Tianrui, et al.
Published: (2024)
by: Guan, Tianrui, et al.
Published: (2024)
RSRNav: Reasoning Spatial Relationship for Image-Goal Navigation
by: Qin, Zheng, et al.
Published: (2025)
by: Qin, Zheng, et al.
Published: (2025)
HOH: Markerless Multimodal Human-Object-Human Handover Dataset with Large Object Count
by: Wiederhold, Noah, et al.
Published: (2023)
by: Wiederhold, Noah, et al.
Published: (2023)
REST: Receding Horizon Explorative Steiner Tree for Zero-Shot Object-Goal Navigation
by: Xiao, Shuqi, et al.
Published: (2026)
by: Xiao, Shuqi, et al.
Published: (2026)
GMT: Goal-Conditioned Multimodal Transformer for 6-DOF Object Trajectory Synthesis in 3D Scenes
by: Zeng, Huajian, et al.
Published: (2026)
by: Zeng, Huajian, et al.
Published: (2026)
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation
by: Bao, Muyi, et al.
Published: (2026)
by: Bao, Muyi, et al.
Published: (2026)
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation
by: Yu, Yinfeng, et al.
Published: (2025)
by: Yu, Yinfeng, et al.
Published: (2025)
FlowLoss: Dynamic Flow-Conditioned Loss Strategy for Video Diffusion Models
by: Wu, Kuanting, et al.
Published: (2025)
by: Wu, Kuanting, et al.
Published: (2025)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
Enhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour Videos
by: Xu, Haoxuan, et al.
Published: (2026)
by: Xu, Haoxuan, et al.
Published: (2026)
Personalized Embodied Navigation for Portable Object Finding
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
LOVON: Legged Open-Vocabulary Object Navigator
by: Peng, Daojie, et al.
Published: (2025)
by: Peng, Daojie, et al.
Published: (2025)
Enhancing Autonomous Navigation by Imaging Hidden Objects using Single-Photon LiDAR
by: Young, Aaron, et al.
Published: (2024)
by: Young, Aaron, et al.
Published: (2024)
Personalized Instance-based Navigation Toward User-Specific Objects in Realistic Environments
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
by: He, Yu, et al.
Published: (2025)
by: He, Yu, et al.
Published: (2025)
Similar Items
-
Embodied Navigation with Auxiliary Task of Action Description Prediction
by: Kondoh, Haru, et al.
Published: (2025) -
DiffSurf: A Transformer-based Diffusion Model for Generating and Reconstructing 3D Surfaces in Pose
by: Yoshiyasu, Yusuke, et al.
Published: (2024) -
TransFusionOdom: Interpretable Transformer-based LiDAR-Inertial Fusion Odometry Estimation
by: Sun, Leyuan, et al.
Published: (2023) -
Zero-Shot Peg Insertion: Identifying Mating Holes and Estimating SE(2) Poses with Vision-Language Models
by: Yajima, Masaru, et al.
Published: (2025) -
OP-Align: Object-level and Part-level Alignment for Self-supervised Category-level Articulated Object Pose Estimation
by: Che, Yuchen, et al.
Published: (2024)