Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions
Fuente:
arXiv
Saved in:
| Main Authors: | Oyama, Akira, Hasegawa, Shoichi, Taniguchi, Akira, Hagiwara, Yoshinobu, Taniguchi, Tadahiro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Whose Is This?: Context-Aware Object Ownership Inference with Uncertainty-Guided Questioning
by: Hashimoto, Saki, et al.
Published: (2026)
by: Hashimoto, Saki, et al.
Published: (2026)
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
by: Hashimoto, Saki, et al.
Published: (2025)
by: Hashimoto, Saki, et al.
Published: (2025)
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
by: Murata, Kento, et al.
Published: (2025)
by: Murata, Kento, et al.
Published: (2025)
Hierarchical Path-planning from Speech Instructions with Spatial Concept-based Topometric Semantic Mapping
by: Taniguchi, Akira, et al.
Published: (2022)
by: Taniguchi, Akira, et al.
Published: (2022)
Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
by: Sakaguchi, Taichi, et al.
Published: (2024)
by: Sakaguchi, Taichi, et al.
Published: (2024)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
by: Okumura, Ryota, et al.
Published: (2025)
by: Okumura, Ryota, et al.
Published: (2025)
Real-world Instance-specific Image Goal Navigation: Bridging Domain Gaps via Contrastive Learning
by: Sakaguchi, Taichi, et al.
Published: (2024)
by: Sakaguchi, Taichi, et al.
Published: (2024)
Reducing Mental Workload through On-Demand Human Assistance for Physical Action Failures in LLM-based Multi-Robot Coordination
by: Hasegawa, Shoichi, et al.
Published: (2026)
by: Hasegawa, Shoichi, et al.
Published: (2026)
DEQ-MCL: Discrete-Event Queue-based Monte-Carlo Localization
by: Taniguchi, Akira, et al.
Published: (2024)
by: Taniguchi, Akira, et al.
Published: (2024)
Generative Emergent Communication: Large Language Model is a Collective World Model
by: Taniguchi, Tadahiro, et al.
Published: (2024)
by: Taniguchi, Tadahiro, et al.
Published: (2024)
System 0/1/2/3: Quad-process theory for multi-timescale embodied collective cognitive systems
by: Taniguchi, Tadahiro, et al.
Published: (2025)
by: Taniguchi, Tadahiro, et al.
Published: (2025)
Beyond Individuals: Collective Predictive Coding for Memory, Attention, and the Emergence of Language
by: Taniguchi, Tadahiro
Published: (2025)
by: Taniguchi, Tadahiro
Published: (2025)
Public Evaluation on Potential Social Impacts of Fully Autonomous Cybernetic Avatars for Physical Support in Daily-Life Environments: Large-Scale Demonstration and Survey at Avatar Land
by: Hafi, Lotfi El, et al.
Published: (2025)
by: Hafi, Lotfi El, et al.
Published: (2025)
Reward-Independent Messaging for Decentralized Multi-Agent Reinforcement Learning
by: Yoshida, Naoto, et al.
Published: (2025)
by: Yoshida, Naoto, et al.
Published: (2025)
ProCompNav: Proactive Instance Navigation with Comparative Judgment for Ambiguous User Queries
by: Kwon, Junhyuk, et al.
Published: (2026)
by: Kwon, Junhyuk, et al.
Published: (2026)
Multi-Level Compositional Reasoning for Interactive Instruction Following
by: Bhambri, Suvaansh, et al.
Published: (2023)
by: Bhambri, Suvaansh, et al.
Published: (2023)
Stable Object Placing using Curl and Diff Features of Vision-based Tactile Sensors
by: Takahashi, Kuniyuki, et al.
Published: (2024)
by: Takahashi, Kuniyuki, et al.
Published: (2024)
A Contact Model based on Denoising Diffusion to Learn Variable Impedance Control for Contact-rich Manipulation
by: Okada, Masashi, et al.
Published: (2024)
by: Okada, Masashi, et al.
Published: (2024)
World-Model-Based Control for Industrial box-packing of Multiple Objects using NewtonianVAE
by: Kato, Yusuke, et al.
Published: (2023)
by: Kato, Yusuke, et al.
Published: (2023)
Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs
by: Chiang, Hao-Tien Lewis, et al.
Published: (2024)
by: Chiang, Hao-Tien Lewis, et al.
Published: (2024)
A Paragraph is All It Takes: Rich Robot Behaviors from Interacting, Trusted LLMs
by: OpenMind, et al.
Published: (2024)
by: OpenMind, et al.
Published: (2024)
Emergent Communication between Heterogeneous Visual Agents through Decentralized Learning
by: Ochiai, Mikako, et al.
Published: (2026)
by: Ochiai, Mikako, et al.
Published: (2026)
Diffusing in Someone Else's Shoes: Robotic Perspective Taking with Diffusion
by: Spisak, Josua, et al.
Published: (2024)
by: Spisak, Josua, et al.
Published: (2024)
NIMS-OS: An automation software to implement a closed loop between artificial intelligence and robotic experiments in materials science
by: Tamura, Ryo, et al.
Published: (2023)
by: Tamura, Ryo, et al.
Published: (2023)
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
by: Ivanova, Anastasiia, et al.
Published: (2025)
by: Ivanova, Anastasiia, et al.
Published: (2025)
Online Continual Learning For Interactive Instruction Following Agents
by: Kim, Byeonghwi, et al.
Published: (2024)
by: Kim, Byeonghwi, et al.
Published: (2024)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
I2EDL: Interactive Instruction Error Detection and Localization
by: Taioli, Francesco, et al.
Published: (2024)
by: Taioli, Francesco, et al.
Published: (2024)
Using Machine Learning to Take Stay-or-Go Decisions in Data-driven Drone Missions
by: Polychronis, Giorgos, et al.
Published: (2025)
by: Polychronis, Giorgos, et al.
Published: (2025)
Towards Affect-Adaptive Human-Robot Interaction: A Protocol for Multimodal Dataset Collection on Social Anxiety
by: Poprcova, Vesna, et al.
Published: (2025)
by: Poprcova, Vesna, et al.
Published: (2025)
Ablation Study of Multimodal Perception, Language Grounding, and Control for Human-Robot Interaction in an Object Detection and Grasping Task
by: Tian, Zi, et al.
Published: (2026)
by: Tian, Zi, et al.
Published: (2026)
TakeAD: Preference-based Post-optimization for End-to-end Autonomous Driving with Expert Takeover Data
by: Liu, Deqing, et al.
Published: (2025)
by: Liu, Deqing, et al.
Published: (2025)
Decentralized Collective World Model for Emergent Communication and Coordination
by: Nomura, Kentaro, et al.
Published: (2025)
by: Nomura, Kentaro, et al.
Published: (2025)
Can LLMs Generate Human-Like Wayfinding Instructions? Towards Platform-Agnostic Embodied Instruction Synthesis
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
Embodied Instruction Following in Unknown Environments
by: Wu, Zhenyu, et al.
Published: (2024)
by: Wu, Zhenyu, et al.
Published: (2024)
Learning to Build by Building Your Own Instructions
by: Walsman, Aaron, et al.
Published: (2024)
by: Walsman, Aaron, et al.
Published: (2024)
Real-Time Out-of-Distribution Failure Prevention via Multi-Modal Reasoning
by: Ganai, Milan, et al.
Published: (2025)
by: Ganai, Milan, et al.
Published: (2025)
Verifiably Following Complex Robot Instructions with Foundation Models
by: Quartey, Benedict, et al.
Published: (2024)
by: Quartey, Benedict, et al.
Published: (2024)
LaMOuR: Leveraging Language Models for Out-of-Distribution Recovery in Reinforcement Learning
by: Kim, Chan, et al.
Published: (2025)
by: Kim, Chan, et al.
Published: (2025)
Out-of-Distribution Recovery with Object-Centric Keypoint Inverse Policy for Visuomotor Imitation Learning
by: Gao, George Jiayuan, et al.
Published: (2024)
by: Gao, George Jiayuan, et al.
Published: (2024)
Similar Items
-
Whose Is This?: Context-Aware Object Ownership Inference with Uncertainty-Guided Questioning
by: Hashimoto, Saki, et al.
Published: (2026) -
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
by: Hashimoto, Saki, et al.
Published: (2025) -
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
by: Murata, Kento, et al.
Published: (2025) -
Hierarchical Path-planning from Speech Instructions with Spatial Concept-based Topometric Semantic Mapping
by: Taniguchi, Akira, et al.
Published: (2022) -
Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
by: Sakaguchi, Taichi, et al.
Published: (2024)