Speech to Reality: On-Demand Production using Natural Language, 3D Generative AI, and Discrete Robotic Assembly
Fuente:
arXiv
Saved in:
| Main Authors: | Kyaw, Alexander Htet, Smith, Miana, Jeon, Se Hwan, Gershenfeld, Neil |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
by: Prieto, S. A., et al.
Published: (2026)
by: Prieto, S. A., et al.
Published: (2026)
Exploring the Effect of Robotic Embodiment and Empathetic Tone of LLMs on Empathy Elicitation
by: Darwesh, Liza, et al.
Published: (2025)
by: Darwesh, Liza, et al.
Published: (2025)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024)
by: Bonial, Claire, et al.
Published: (2024)
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
by: Lukin, Stephanie M., et al.
Published: (2024)
by: Lukin, Stephanie M., et al.
Published: (2024)
HoloSpot: Intuitive Object Manipulation via Mixed Reality Drag-and-Drop
by: Garcia, Pablo Soler, et al.
Published: (2024)
by: Garcia, Pablo Soler, et al.
Published: (2024)
What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics
by: Wachowiak, Lennart, et al.
Published: (2025)
by: Wachowiak, Lennart, et al.
Published: (2025)
Otherness as a Quality in Designing Expressive Robotic Touch
by: Zhou, Ran, et al.
Published: (2026)
by: Zhou, Ran, et al.
Published: (2026)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
by: Menon, Anjali R., et al.
Published: (2025)
by: Menon, Anjali R., et al.
Published: (2025)
Unconscious and Intentional Human Motion Cues for Expressive Robot-Arm Motion Design
by: Tashiro, Taito, et al.
Published: (2025)
by: Tashiro, Taito, et al.
Published: (2025)
HiSync: Spatio-Temporally Aligning Hand Motion from Wearable IMU and On-Robot Camera for Command Source Identification in Long-Range HRI
by: Zhang, Chengwen, et al.
Published: (2026)
by: Zhang, Chengwen, et al.
Published: (2026)
Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains
by: Calboreanu, Elias
Published: (2026)
by: Calboreanu, Elias
Published: (2026)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
by: Lim, Shoon Kit, et al.
Published: (2025)
by: Lim, Shoon Kit, et al.
Published: (2025)
IRL Dittos: Embodied Multimodal AI Agent Interactions in Open Spaces
by: Lee, Seonghee, et al.
Published: (2025)
by: Lee, Seonghee, et al.
Published: (2025)
Interactive Groupwise Comparison for Reinforcement Learning from Human Feedback
by: Kompatscher, Jan, et al.
Published: (2025)
by: Kompatscher, Jan, et al.
Published: (2025)
BRAVE: Brain-Controlled Prosthetic Arm with Voice Integration and Embodied Learning for Enhanced Mobility
by: Basit, Abdul, et al.
Published: (2025)
by: Basit, Abdul, et al.
Published: (2025)
MAP: Multi-user Personalization with Collaborative LLM-powered Agents
by: Lee, Christine, et al.
Published: (2025)
by: Lee, Christine, et al.
Published: (2025)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Immersive Robot Programming Interface for Human-Guided Automation and Randomized Path Planning
by: Malek, Kaveh, et al.
Published: (2024)
by: Malek, Kaveh, et al.
Published: (2024)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
by: Sharma, Aditya, et al.
Published: (2025)
by: Sharma, Aditya, et al.
Published: (2025)
Diffusion-SAFE: Diffusion-Native Human-to-Robot Driving Handover for Shared Autonomy
by: Fan, Yunxin, et al.
Published: (2025)
by: Fan, Yunxin, et al.
Published: (2025)
M2HRI: An LLM-Driven Multimodal Multi-Agent Framework for Personalized Human-Robot Interaction
by: Hasan, Shaid, et al.
Published: (2026)
by: Hasan, Shaid, et al.
Published: (2026)
GIST: Multimodal Knowledge Extraction and Spatial Grounding via Intelligent Semantic Topology
by: Agrawal, Shivendra, et al.
Published: (2026)
by: Agrawal, Shivendra, et al.
Published: (2026)
A Service Robot's Guide to Interacting with Busy Customers
by: Nukala, Suraj, et al.
Published: (2025)
by: Nukala, Suraj, et al.
Published: (2025)
Conversations with Andrea: Visitors' Opinions on Android Robots in a Museum
by: Heisler, Marcel, et al.
Published: (2025)
by: Heisler, Marcel, et al.
Published: (2025)
Training Human-Robot Teams by Improving Transparency Through a Virtual Spectator Interface
by: Dallas, Sean, et al.
Published: (2025)
by: Dallas, Sean, et al.
Published: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
by: Li, Danyang, et al.
Published: (2025)
by: Li, Danyang, et al.
Published: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
by: Deichler, Anna, et al.
Published: (2025)
by: Deichler, Anna, et al.
Published: (2025)
Aesthetics of Robot-Mediated Applied Drama: A Case Study on REMind
by: Sanoubari, Elaheh, et al.
Published: (2026)
by: Sanoubari, Elaheh, et al.
Published: (2026)
Mixed-Initiative Dialog for Human-Robot Collaborative Manipulation
by: Yu, Albert, et al.
Published: (2025)
by: Yu, Albert, et al.
Published: (2025)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
3HANDS Dataset: Learning from Humans for Generating Naturalistic Handovers with Supernumerary Robotic Limbs
by: Abadian, Artin Saberpour, et al.
Published: (2025)
by: Abadian, Artin Saberpour, et al.
Published: (2025)
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
by: Reza, Mohi, et al.
Published: (2025)
by: Reza, Mohi, et al.
Published: (2025)
CWM: Contrastive World Models for Action Feasibility Learning in Embodied Agent Pipelines
by: Banerjee, Chayan
Published: (2026)
by: Banerjee, Chayan
Published: (2026)
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making
by: Li, Wenbo, et al.
Published: (2025)
by: Li, Wenbo, et al.
Published: (2025)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
by: Chen, Yiteng, et al.
Published: (2025)
by: Chen, Yiteng, et al.
Published: (2025)
Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation
by: Stratton, Andrew, et al.
Published: (2026)
by: Stratton, Andrew, et al.
Published: (2026)
Working with Trouble and Failures in Conversation between Humans and Robots (WTF 2023) & Is CUI Design Ready Yet?
by: Förster, Frank, et al.
Published: (2023)
by: Förster, Frank, et al.
Published: (2023)
AI Pedagogy: Dialogic Social Learning for Artificial Agents
by: Patania, Sabrina, et al.
Published: (2025)
by: Patania, Sabrina, et al.
Published: (2025)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
by: Luijkx, Jelle, et al.
Published: (2025)
by: Luijkx, Jelle, et al.
Published: (2025)
Similar Items
-
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
by: Prieto, S. A., et al.
Published: (2026) -
Exploring the Effect of Robotic Embodiment and Empathetic Tone of LLMs on Empathy Elicitation
by: Darwesh, Liza, et al.
Published: (2025) -
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024) -
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
by: Lukin, Stephanie M., et al.
Published: (2024) -
HoloSpot: Intuitive Object Manipulation via Mixed Reality Drag-and-Drop
by: Garcia, Pablo Soler, et al.
Published: (2024)