Saved in:
| Main Authors: | Fedozzi, Marco Gabriele, Nagai, Yukie, Rea, Francesco, Sciutti, Alessandra |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.08418 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Let people fail! Exploring the influence of explainable virtual and robotic agents in learning-by-doing tasks
by: Matarese, Marco, et al.
Published: (2024)
by: Matarese, Marco, et al.
Published: (2024)
Building Knowledge from Interactions: An LLM-Based Architecture for Adaptive Tutoring and Social Reasoning
by: Garello, Luca, et al.
Published: (2025)
by: Garello, Luca, et al.
Published: (2025)
Game Theory to Study Cooperation in Human-Robot Mixed Groups: Exploring the Potential of the Public Good Game
by: Pusceddu, Giulia, et al.
Published: (2025)
by: Pusceddu, Giulia, et al.
Published: (2025)
To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills
by: Mazzola, Carlo, et al.
Published: (2023)
by: Mazzola, Carlo, et al.
Published: (2023)
Correspondence learning between morphologically different robots via task demonstrations
by: Aktas, Hakan, et al.
Published: (2023)
by: Aktas, Hakan, et al.
Published: (2023)
Cross-Embodied Affordance Transfer through Learning Affordance Equivalences
by: Aktas, Hakan, et al.
Published: (2024)
by: Aktas, Hakan, et al.
Published: (2024)
If They Disagree, Will You Conform? Exploring the Role of Robots' Value Awareness in a Decision-Making Task
by: Pusceddu, Giulia, et al.
Published: (2025)
by: Pusceddu, Giulia, et al.
Published: (2025)
Next-Gen Museum Guides: Autonomous Navigation and Visitor Interaction with an Agentic Robot
by: Garello, Luca, et al.
Published: (2025)
by: Garello, Luca, et al.
Published: (2025)
Expressing and Inferring Action Carefulness in Human-to-Robot Handovers
by: Lastrico, Linda, et al.
Published: (2023)
by: Lastrico, Linda, et al.
Published: (2023)
Temporal Action Representation Learning for Tactical Resource Control and Subsequent Maneuver Generation
by: Jung, Hoseong, et al.
Published: (2026)
by: Jung, Hoseong, et al.
Published: (2026)
The Effects of Selected Object Features on a Pick-and-Place Task: a Human Multimodal Dataset
by: Lastrico, Linda, et al.
Published: (2024)
by: Lastrico, Linda, et al.
Published: (2024)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
M2R2: MultiModal Robotic Representation for Temporal Action Segmentation
by: Sliwowski, Daniel, et al.
Published: (2025)
by: Sliwowski, Daniel, et al.
Published: (2025)
Prediction with Action: Visual Policy Learning via Joint Denoising Process
by: Guo, Yanjiang, et al.
Published: (2024)
by: Guo, Yanjiang, et al.
Published: (2024)
Decision-Focused Learning to Predict Action Costs for Planning
by: Mandi, Jayanta, et al.
Published: (2024)
by: Mandi, Jayanta, et al.
Published: (2024)
Mapping Embodied Affective Touch Strategies on a Humanoid Robot
by: Ren, Qiaoqiao, et al.
Published: (2026)
by: Ren, Qiaoqiao, et al.
Published: (2026)
A Roadmap for Embodied and Social Grounding in LLMs
by: Incao, Sara, et al.
Published: (2024)
by: Incao, Sara, et al.
Published: (2024)
A roadmap for AI in robotics
by: Billard, Aude, et al.
Published: (2025)
by: Billard, Aude, et al.
Published: (2025)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
by: Wang, Taowen, et al.
Published: (2024)
by: Wang, Taowen, et al.
Published: (2024)
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
by: Sun, Yuteng, et al.
Published: (2026)
by: Sun, Yuteng, et al.
Published: (2026)
A Primer on SO(3) Action Representations in Deep Reinforcement Learning
by: Schuck, Martin, et al.
Published: (2025)
by: Schuck, Martin, et al.
Published: (2025)
Context-Aware Human Behavior Prediction Using Multimodal Large Language Models: Challenges and Insights
by: Liu, Yuchen, et al.
Published: (2025)
by: Liu, Yuchen, et al.
Published: (2025)
Premier-TACO is a Few-Shot Policy Learner: Pretraining Multitask Representation via Temporal Action-Driven Contrastive Loss
by: Zheng, Ruijie, et al.
Published: (2024)
by: Zheng, Ruijie, et al.
Published: (2024)
Bridging Scale Discrepancies in Robotic Control via Language-Based Action Representations
by: Zhang, Yuchi, et al.
Published: (2025)
by: Zhang, Yuchi, et al.
Published: (2025)
Low Dimensional State Representation Learning with Robotics Priors in Continuous Action Spaces
by: Botteghi, Nicolò, et al.
Published: (2021)
by: Botteghi, Nicolò, et al.
Published: (2021)
Bridging Bots: from Perception to Action via Multimodal-LMs and Knowledge Graphs
by: Martorana, Margherita, et al.
Published: (2025)
by: Martorana, Margherita, et al.
Published: (2025)
IMRL: Integrating Visual, Physical, Temporal, and Geometric Representations for Enhanced Food Acquisition
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
Exploring Embodied Multimodal Large Models: Development, Datasets, and Future Directions
by: Chen, Shoubin, et al.
Published: (2025)
by: Chen, Shoubin, et al.
Published: (2025)
Understanding Multimodal Failure in Action-Chunking Behavioral Cloning
by: Mazza, Lorenzo, et al.
Published: (2026)
by: Mazza, Lorenzo, et al.
Published: (2026)
Reconciling Spatial and Temporal Abstractions for Goal Representation
by: Zadem, Mehdi, et al.
Published: (2024)
by: Zadem, Mehdi, et al.
Published: (2024)
Exploring Spatial Representation to Enhance LLM Reasoning in Aerial Vision-Language Navigation
by: Gao, Yunpeng, et al.
Published: (2024)
by: Gao, Yunpeng, et al.
Published: (2024)
Action-to-Action Flow Matching
by: Jia, Jindou, et al.
Published: (2026)
by: Jia, Jindou, et al.
Published: (2026)
ETA-VLA: Efficient Token Adaptation via Temporal Fusion and Intra-LLM Sparsification for Vision-Language-Action Models
by: Wang, Yiru, et al.
Published: (2026)
by: Wang, Yiru, et al.
Published: (2026)
Humanoid Agent via Embodied Chain-of-Action Reasoning with Multimodal Foundation Models for Zero-Shot Loco-Manipulation
by: Wen, Congcong, et al.
Published: (2025)
by: Wen, Congcong, et al.
Published: (2025)
Discrete Gaussian Process Representations for Optimising UAV-based Precision Weed Mapping
by: Swindell, Jacob, et al.
Published: (2025)
by: Swindell, Jacob, et al.
Published: (2025)
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
by: Chen, Xinzhe, et al.
Published: (2026)
by: Chen, Xinzhe, et al.
Published: (2026)
Spectral Alignment in Forward-Backward Representations via Temporal Abstraction
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
ActionCodec: What Makes for Good Action Tokenizers
by: Dong, Zibin, et al.
Published: (2026)
by: Dong, Zibin, et al.
Published: (2026)
Action Hallucination in Generative Vision-Language-Action Models
by: Soh, Harold, et al.
Published: (2026)
by: Soh, Harold, et al.
Published: (2026)
Neural Interaction Energy for Multi-Agent Trajectory Prediction
by: Shen, Kaixin, et al.
Published: (2024)
by: Shen, Kaixin, et al.
Published: (2024)
Similar Items
-
Let people fail! Exploring the influence of explainable virtual and robotic agents in learning-by-doing tasks
by: Matarese, Marco, et al.
Published: (2024) -
Building Knowledge from Interactions: An LLM-Based Architecture for Adaptive Tutoring and Social Reasoning
by: Garello, Luca, et al.
Published: (2025) -
Game Theory to Study Cooperation in Human-Robot Mixed Groups: Exploring the Potential of the Public Good Game
by: Pusceddu, Giulia, et al.
Published: (2025) -
To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills
by: Mazzola, Carlo, et al.
Published: (2023) -
Correspondence learning between morphologically different robots via task demonstrations
by: Aktas, Hakan, et al.
Published: (2023)