GIST: Multimodal Knowledge Extraction and Spatial Grounding via Intelligent Semantic Topology
Fuente:
arXiv
Saved in:
| Main Authors: | Agrawal, Shivendra, Hayes, Bradley |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
by: Prieto, S. A., et al.
Published: (2026)
by: Prieto, S. A., et al.
Published: (2026)
Exploring the Effect of Robotic Embodiment and Empathetic Tone of LLMs on Empathy Elicitation
by: Darwesh, Liza, et al.
Published: (2025)
by: Darwesh, Liza, et al.
Published: (2025)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024)
by: Bonial, Claire, et al.
Published: (2024)
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
by: Lukin, Stephanie M., et al.
Published: (2024)
by: Lukin, Stephanie M., et al.
Published: (2024)
Human Agency, Causality, and the Human Computer Interface in High-Stakes Artificial Intelligence
by: Hattab, Georges
Published: (2026)
by: Hattab, Georges
Published: (2026)
What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics
by: Wachowiak, Lennart, et al.
Published: (2025)
by: Wachowiak, Lennart, et al.
Published: (2025)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
by: Deichler, Anna, et al.
Published: (2025)
by: Deichler, Anna, et al.
Published: (2025)
Speech to Reality: On-Demand Production using Natural Language, 3D Generative AI, and Discrete Robotic Assembly
by: Kyaw, Alexander Htet, et al.
Published: (2024)
by: Kyaw, Alexander Htet, et al.
Published: (2024)
"The Data Says Otherwise"-Towards Automated Fact-checking and Communication of Data Claims
by: Fu, Yu, et al.
Published: (2024)
by: Fu, Yu, et al.
Published: (2024)
CommentScope: A Comment-Embedded Assisted Reading System for a Long Text
by: Chen, Shuai, et al.
Published: (2025)
by: Chen, Shuai, et al.
Published: (2025)
StudyAlign: A Software System for Conducting Web-Based User Studies with Functional Interactive Prototypes
by: Lehmann, Florian, et al.
Published: (2025)
by: Lehmann, Florian, et al.
Published: (2025)
Functional Flexibility in Generative AI Interfaces: Text Editing with LLMs through Conversations, Toolbars, and Prompts
by: Lehmann, Florian, et al.
Published: (2024)
by: Lehmann, Florian, et al.
Published: (2024)
Can LLMs and humans be friends? Uncovering factors affecting human-AI intimacy formation
by: Hong, Yeseon, et al.
Published: (2025)
by: Hong, Yeseon, et al.
Published: (2025)
Mitigating Response Delays in Free-Form Conversations with LLM-powered Intelligent Virtual Agents
by: Maslych, Mykola, et al.
Published: (2025)
by: Maslych, Mykola, et al.
Published: (2025)
Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains
by: Calboreanu, Elias
Published: (2026)
by: Calboreanu, Elias
Published: (2026)
OOPrompt: Reifying Intents into Structured Artifacts for Modular and Iterative Prompting
by: Xu, Tengyou, et al.
Published: (2026)
by: Xu, Tengyou, et al.
Published: (2026)
Unconscious and Intentional Human Motion Cues for Expressive Robot-Arm Motion Design
by: Tashiro, Taito, et al.
Published: (2025)
by: Tashiro, Taito, et al.
Published: (2025)
Otherness as a Quality in Designing Expressive Robotic Touch
by: Zhou, Ran, et al.
Published: (2026)
by: Zhou, Ran, et al.
Published: (2026)
HiSync: Spatio-Temporally Aligning Hand Motion from Wearable IMU and On-Robot Camera for Command Source Identification in Long-Range HRI
by: Zhang, Chengwen, et al.
Published: (2026)
by: Zhang, Chengwen, et al.
Published: (2026)
HoloSpot: Intuitive Object Manipulation via Mixed Reality Drag-and-Drop
by: Garcia, Pablo Soler, et al.
Published: (2024)
by: Garcia, Pablo Soler, et al.
Published: (2024)
IRL Dittos: Embodied Multimodal AI Agent Interactions in Open Spaces
by: Lee, Seonghee, et al.
Published: (2025)
by: Lee, Seonghee, et al.
Published: (2025)
MAESTRO: Adapting GUIs and Guiding Navigation with User Preferences in Conversational Agents with GUIs
by: Lee, Sangwook, et al.
Published: (2026)
by: Lee, Sangwook, et al.
Published: (2026)
Designing Transparent AI-Mediated Language Support for Intergenerational Family Communication
by: Kang, Sora, et al.
Published: (2026)
by: Kang, Sora, et al.
Published: (2026)
M2HRI: An LLM-Driven Multimodal Multi-Agent Framework for Personalized Human-Robot Interaction
by: Hasan, Shaid, et al.
Published: (2026)
by: Hasan, Shaid, et al.
Published: (2026)
A Service Robot's Guide to Interacting with Busy Customers
by: Nukala, Suraj, et al.
Published: (2025)
by: Nukala, Suraj, et al.
Published: (2025)
SAGE: Sensor-Augmented Grounding Engine for LLM-Powered Sleep Care Agent
by: Lee, Hansoo, et al.
Published: (2026)
by: Lee, Hansoo, et al.
Published: (2026)
WebNav: An Intelligent Agent for Voice-Controlled Web Navigation
by: Srinivasan, Trisanth, et al.
Published: (2025)
by: Srinivasan, Trisanth, et al.
Published: (2025)
Distorted Perspectives of LLM-Simulated Preferences: Can AI Mislead Design?
by: Kuric, Eduard, et al.
Published: (2026)
by: Kuric, Eduard, et al.
Published: (2026)
Inaccuracy of an E-Dictionary and Its Influence on Chinese Language Users
by: Zhang, Shiyang, et al.
Published: (2025)
by: Zhang, Shiyang, et al.
Published: (2025)
Know Your Author: Does the AI Penalty Hold in Short Fiction?
by: Todasco, Michael, et al.
Published: (2026)
by: Todasco, Michael, et al.
Published: (2026)
Exploring Mobile Touch Interaction with Large Language Models
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Content-Driven Local Response: Supporting Sentence-Level and Message-Level Mobile Email Replies With and Without AI
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Writer-Defined AI Personas for On-Demand Feedback Generation
by: Benharrak, Karim, et al.
Published: (2023)
by: Benharrak, Karim, et al.
Published: (2023)
Collage is the New Writing: Exploring the Fragmentation of Text and User Interfaces in AI Tools
by: Buschek, Daniel
Published: (2024)
by: Buschek, Daniel
Published: (2024)
LLMs and people both learn to form conventions -- just not with each other
by: Jones, Cameron R., et al.
Published: (2026)
by: Jones, Cameron R., et al.
Published: (2026)
The AI Memory Gap: Users Misremember What They Created With AI or Without
by: Zindulka, Tim, et al.
Published: (2025)
by: Zindulka, Tim, et al.
Published: (2025)
Real-Time World Crafting: Generating Structured Game Behaviors from Natural Language with Large Language Models
by: Drake, Austin, et al.
Published: (2025)
by: Drake, Austin, et al.
Published: (2025)
Collaborative Document Editing with Multiple Users and AI Agents
by: Lehmann, Florian, et al.
Published: (2025)
by: Lehmann, Florian, et al.
Published: (2025)
The Adaptation Paradox: Agency vs. Mimicry in Companion Chatbots
by: Brandt, T. James, et al.
Published: (2025)
by: Brandt, T. James, et al.
Published: (2025)
CorpusStudio: Surfacing Emergent Patterns in a Corpus of Prior Work while Writing
by: Dang, Hai, et al.
Published: (2025)
by: Dang, Hai, et al.
Published: (2025)
Similar Items
-
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
by: Prieto, S. A., et al.
Published: (2026) -
Exploring the Effect of Robotic Embodiment and Empathetic Tone of LLMs on Empathy Elicitation
by: Darwesh, Liza, et al.
Published: (2025) -
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024) -
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
by: Lukin, Stephanie M., et al.
Published: (2024) -
Human Agency, Causality, and the Human Computer Interface in High-Stakes Artificial Intelligence
by: Hattab, Georges
Published: (2026)