Affordance-Guided Coarse-to-Fine Exploration for Base Placement in Open-Vocabulary Mobile Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Tzu-Jung, Yeh, Jia-Fong, Su, Hung-Ting, Lin, Chung-Yi, Chen, Yi-Ting, Hsu, Winston H. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints
by: Chen, Pei-An, et al.
Published: (2026)
by: Chen, Pei-An, et al.
Published: (2026)
VICtoR: Learning Hierarchical Vision-Instruction Correlation Rewards for Long-horizon Manipulation
by: Hung, Kuo-Han, et al.
Published: (2024)
by: Hung, Kuo-Han, et al.
Published: (2024)
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
by: Su, Hung-Ting, et al.
Published: (2026)
by: Su, Hung-Ting, et al.
Published: (2026)
AED: Adaptable Error Detection for Few-shot Imitation Policy
by: Yeh, Jia-Fong, et al.
Published: (2024)
by: Yeh, Jia-Fong, et al.
Published: (2024)
Context-Aware Replanning with Pre-explored Semantic Map for Object Navigation
by: Ko, Po-Chen, et al.
Published: (2024)
by: Ko, Po-Chen, et al.
Published: (2024)
OVAL-Prompt: Open-Vocabulary Affordance Localization for Robot Manipulation through LLM Affordance-Grounding
by: Tong, Edmond, et al.
Published: (2024)
by: Tong, Edmond, et al.
Published: (2024)
Mitigating Cross-Modal Distraction and Ensuring Geometric Feasibility via Affordance-Guided and Self-Consistent MLLMs for Task Planning in Instruction-Following Manipulation
by: Shen, Yu-Hong, et al.
Published: (2025)
by: Shen, Yu-Hong, et al.
Published: (2025)
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding
by: Faure, Gueter Josmy, et al.
Published: (2026)
by: Faure, Gueter Josmy, et al.
Published: (2026)
Improving Generalization Ability for 3D Object Detection by Learning Sparsity-invariant Features
by: Lu, Hsin-Cheng, et al.
Published: (2025)
by: Lu, Hsin-Cheng, et al.
Published: (2025)
Language-Conditioned Open-Vocabulary Mobile Manipulation with Pretrained Models
by: Tan, Shen, et al.
Published: (2025)
by: Tan, Shen, et al.
Published: (2025)
Enhancing Sustainable Urban Mobility Prediction with Telecom Data: A Spatio-Temporal Framework Approach
by: Lin, ChungYi, et al.
Published: (2024)
by: Lin, ChungYi, et al.
Published: (2024)
Dynamic Open-Vocabulary 3D Scene Graphs for Long-term Language-Guided Mobile Manipulation
by: Yan, Zhijie, et al.
Published: (2024)
by: Yan, Zhijie, et al.
Published: (2024)
SAGA: Open-World Mobile Manipulation via Structured Affordance Grounding
by: Fang, Kuan, et al.
Published: (2025)
by: Fang, Kuan, et al.
Published: (2025)
HomeRobot: Open-Vocabulary Mobile Manipulation
by: Yenamandra, Sriram, et al.
Published: (2023)
by: Yenamandra, Sriram, et al.
Published: (2023)
Tracking-Assisted Object Detection with Event Cameras
by: Yen, Ting-Kang, et al.
Published: (2024)
by: Yen, Ting-Kang, et al.
Published: (2024)
BINDER: Instantly Adaptive Mobile Manipulation with Open-Vocabulary Commands
by: Cho, Seongwon, et al.
Published: (2025)
by: Cho, Seongwon, et al.
Published: (2025)
Shared-unique Features and Task-aware Prioritized Sampling on Multi-task Reinforcement Learning
by: Lin, Po-Shao, et al.
Published: (2024)
by: Lin, Po-Shao, et al.
Published: (2024)
OVAL-Grasp: Open-Vocabulary Affordance Localization for Task Oriented Grasping
by: Tong, Edmond, et al.
Published: (2025)
by: Tong, Edmond, et al.
Published: (2025)
AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance
by: Wei, Yi-Lin, et al.
Published: (2025)
by: Wei, Yi-Lin, et al.
Published: (2025)
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
Articulated Object Manipulation with Coarse-to-fine Affordance for Mitigating the Effect of Point Cloud Noise
by: Ling, Suhan, et al.
Published: (2024)
by: Ling, Suhan, et al.
Published: (2024)
Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V
by: Zhi, Peiyuan, et al.
Published: (2024)
by: Zhi, Peiyuan, et al.
Published: (2024)
AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter
by: Tang, Yingbo, et al.
Published: (2025)
by: Tang, Yingbo, et al.
Published: (2025)
Subgoal Diffuser: Coarse-to-fine Subgoal Generation to Guide Model Predictive Control for Robot Manipulation
by: Huang, Zixuan, et al.
Published: (2024)
by: Huang, Zixuan, et al.
Published: (2024)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
by: Yenamandra, Sriram, et al.
Published: (2024)
by: Yenamandra, Sriram, et al.
Published: (2024)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
by: Jiang, Zebin, et al.
Published: (2025)
by: Jiang, Zebin, et al.
Published: (2025)
Reflective VLM Planning for Dual-Arm Desktop Cleaning: Bridging Open-Vocabulary Perception and Precise Manipulation
by: Liu, Yufan, et al.
Published: (2025)
by: Liu, Yufan, et al.
Published: (2025)
Open-Vocabulary Mobile Manipulation Based on Double Relaxed Contrastive Learning with Dense Labeling
by: Yashima, Daichi, et al.
Published: (2024)
by: Yashima, Daichi, et al.
Published: (2024)
SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization
by: Chen, Posheng, et al.
Published: (2026)
by: Chen, Posheng, et al.
Published: (2026)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
by: Ma, Teli, et al.
Published: (2024)
by: Ma, Teli, et al.
Published: (2024)
DORA: Object Affordance-Guided Reinforcement Learning for Dexterous Robotic Manipulation
by: Zhang, Lei, et al.
Published: (2025)
by: Zhang, Lei, et al.
Published: (2025)
UMIRobot: An Open-{Software, Hardware} Low-Cost Robotic Manipulator for Education
by: Marinho, Murilo M., et al.
Published: (2023)
by: Marinho, Murilo M., et al.
Published: (2023)
OVExp: Open Vocabulary Exploration for Object-Oriented Navigation
by: Wei, Meng, et al.
Published: (2024)
by: Wei, Meng, et al.
Published: (2024)
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
by: Korekata, Ryosuke, et al.
Published: (2025)
by: Korekata, Ryosuke, et al.
Published: (2025)
Learning Affordances from Interactive Exploration using an Object-level Map
by: Wulkop, Paula, et al.
Published: (2025)
by: Wulkop, Paula, et al.
Published: (2025)
Tel2Veh: Fusion of Telecom Data and Vehicle Flow to Predict Camera-Free Traffic via a Spatio-Temporal Framework
by: Lin, ChungYi, et al.
Published: (2024)
by: Lin, ChungYi, et al.
Published: (2024)
TelTrans: Applying Multi-Type Telecom Data to Transportation Evaluation and Prediction via Multifaceted Graph Modeling
by: Lin, ChungYi, et al.
Published: (2024)
by: Lin, ChungYi, et al.
Published: (2024)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
by: Chu, Hengshuo, et al.
Published: (2025)
by: Chu, Hengshuo, et al.
Published: (2025)
WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection
by: Tsou, Tsung-Lin, et al.
Published: (2023)
by: Tsou, Tsung-Lin, et al.
Published: (2023)
Similar Items
-
ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints
by: Chen, Pei-An, et al.
Published: (2026) -
VICtoR: Learning Hierarchical Vision-Instruction Correlation Rewards for Long-horizon Manipulation
by: Hung, Kuo-Han, et al.
Published: (2024) -
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
by: Su, Hung-Ting, et al.
Published: (2026) -
AED: Adaptable Error Detection for Few-shot Imitation Policy
by: Yeh, Jia-Fong, et al.
Published: (2024) -
Context-Aware Replanning with Pre-explored Semantic Map for Object Navigation
by: Ko, Po-Chen, et al.
Published: (2024)