HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yokoyama, Naoki, Ramrakhya, Ram, Das, Abhishek, Batra, Dhruv, Ha, Sehoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
von: Khanna, Mukul, et al.
Veröffentlicht: (2024)
von: Khanna, Mukul, et al.
Veröffentlicht: (2024)
FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning
von: Yokoyama, Naoki, et al.
Veröffentlicht: (2025)
von: Yokoyama, Naoki, et al.
Veröffentlicht: (2025)
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
von: Ortega-Peimbert, Jesús, et al.
Veröffentlicht: (2025)
von: Ortega-Peimbert, Jesús, et al.
Veröffentlicht: (2025)
OpenObject-NAV: Open-Vocabulary Object-Oriented Navigation Based on Dynamic Carrier-Relationship Scene Graph
von: Tang, Yujie, et al.
Veröffentlicht: (2024)
von: Tang, Yujie, et al.
Veröffentlicht: (2024)
A Design Co-Pilot for Task-Tailored Manipulators
von: Külz, Jonathan, et al.
Veröffentlicht: (2025)
von: Külz, Jonathan, et al.
Veröffentlicht: (2025)
Semantic Environment Atlas for Object-Goal Navigation
von: Kim, Nuri, et al.
Veröffentlicht: (2024)
von: Kim, Nuri, et al.
Veröffentlicht: (2024)
RobotDesignGPT: Automated Robot Design Synthesis using Vision Language Models
von: Sontakke, Nitish, et al.
Veröffentlicht: (2026)
von: Sontakke, Nitish, et al.
Veröffentlicht: (2026)
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
von: Busch, Finn Lukas, et al.
Veröffentlicht: (2024)
von: Busch, Finn Lukas, et al.
Veröffentlicht: (2024)
LGR: LLM-Guided Ranking of Frontiers for Object Goal Navigation
von: Uno, Mitsuaki, et al.
Veröffentlicht: (2025)
von: Uno, Mitsuaki, et al.
Veröffentlicht: (2025)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
von: Yang, Zebin, et al.
Veröffentlicht: (2025)
von: Yang, Zebin, et al.
Veröffentlicht: (2025)
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
Reliable Semantic Understanding for Real World Zero-shot Object Goal Navigation
von: Unlu, Halil Utku, et al.
Veröffentlicht: (2024)
von: Unlu, Halil Utku, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Multi-Agent Coordination
von: Aina, Kehinde O., et al.
Veröffentlicht: (2025)
von: Aina, Kehinde O., et al.
Veröffentlicht: (2025)
HomeRobot: Open-Vocabulary Mobile Manipulation
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
von: Rho, Seungeun, et al.
Veröffentlicht: (2025)
von: Rho, Seungeun, et al.
Veröffentlicht: (2025)
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
von: Debnath, Arnab, et al.
Veröffentlicht: (2025)
von: Debnath, Arnab, et al.
Veröffentlicht: (2025)
LOST-3DSG: Lightweight Open-Vocabulary 3D Scene Graphs with Semantic Tracking in Dynamic Environments
von: Ferraina, Sara Micol, et al.
Veröffentlicht: (2026)
von: Ferraina, Sara Micol, et al.
Veröffentlicht: (2026)
Open-Vocabulary Action Localization with Iterative Visual Prompting
von: Wake, Naoki, et al.
Veröffentlicht: (2024)
von: Wake, Naoki, et al.
Veröffentlicht: (2024)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
LineRides: Line-Guided Reinforcement Learning for Bicycle Robot Stunts
von: Rho, Seungeun, et al.
Veröffentlicht: (2026)
von: Rho, Seungeun, et al.
Veröffentlicht: (2026)
FUS3DMaps: Scalable and Accurate Open-Vocabulary Semantic Mapping by 3D Fusion of Voxel- and Instance-Level Layers
von: Homberger, Timon, et al.
Veröffentlicht: (2026)
von: Homberger, Timon, et al.
Veröffentlicht: (2026)
Transforming a Quadruped into a Guide Robot for the Visually Impaired: Formalizing Wayfinding, Interaction Modeling, and Safety Mechanism
von: Kim, J. Taery, et al.
Veröffentlicht: (2023)
von: Kim, J. Taery, et al.
Veröffentlicht: (2023)
OVSegDT: Segmenting Transformer for Open-Vocabulary Object Goal Navigation
von: Zemskova, Tatiana, et al.
Veröffentlicht: (2025)
von: Zemskova, Tatiana, et al.
Veröffentlicht: (2025)
Uncertainty-Informed Active Perception for Open Vocabulary Object Goal Navigation
von: Bajpai, Utkarsh, et al.
Veröffentlicht: (2025)
von: Bajpai, Utkarsh, et al.
Veröffentlicht: (2025)
Point2Graph: An End-to-end Point Cloud-based 3D Open-Vocabulary Scene Graph for Robot Navigation
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
Dynamic Manipulation of Deformable Objects in 3D: Simulation, Benchmark and Learning Strategy
von: Lan, Guanzhou, et al.
Veröffentlicht: (2025)
von: Lan, Guanzhou, et al.
Veröffentlicht: (2025)
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2024)
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2024)
BINDER: Instantly Adaptive Mobile Manipulation with Open-Vocabulary Commands
von: Cho, Seongwon, et al.
Veröffentlicht: (2025)
von: Cho, Seongwon, et al.
Veröffentlicht: (2025)
PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks
von: Chang, Matthew, et al.
Veröffentlicht: (2024)
von: Chang, Matthew, et al.
Veröffentlicht: (2024)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
3DGSNav: Enhancing Vision-Language Model Reasoning for Object Navigation via Active 3D Gaussian Splatting
von: Zheng, Wancai, et al.
Veröffentlicht: (2026)
von: Zheng, Wancai, et al.
Veröffentlicht: (2026)
OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation
von: Pei, Jiahua, et al.
Veröffentlicht: (2026)
von: Pei, Jiahua, et al.
Veröffentlicht: (2026)
GoalSwarm: Multi-UAV Semantic Coordination for Open-Vocabulary Object Navigation
von: James, MoniJesu Wonders, et al.
Veröffentlicht: (2026)
von: James, MoniJesu Wonders, et al.
Veröffentlicht: (2026)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
E2E Parking Dataset: An Open Benchmark for End-to-End Autonomous Parking
von: Gao, Kejia, et al.
Veröffentlicht: (2025)
von: Gao, Kejia, et al.
Veröffentlicht: (2025)
Hierarchical Reinforcement Learning in Multi-Goal Spatial Navigation with Autonomous Mobile Robots
von: Johnson, Brendon, et al.
Veröffentlicht: (2025)
von: Johnson, Brendon, et al.
Veröffentlicht: (2025)
Open-World Drone Active Tracking with Goal-Centered Rewards
von: Sun, Haowei, et al.
Veröffentlicht: (2024)
von: Sun, Haowei, et al.
Veröffentlicht: (2024)
Reflex-Based Open-Vocabulary Navigation without Prior Knowledge Using Omnidirectional Camera and Multiple Vision-Language Models
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025) -
GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
von: Khanna, Mukul, et al.
Veröffentlicht: (2024) -
FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning
von: Yokoyama, Naoki, et al.
Veröffentlicht: (2025) -
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025) -
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
von: Ortega-Peimbert, Jesús, et al.
Veröffentlicht: (2025)