Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Heng, Liang, Xu, Jiadong, Wang, Yiwen, Li, Xiaoqi, Cai, Muhe, Shen, Yan, Zhu, Juan, Ren, Guanghui, Dong, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
by: Li, Xiaoqi, et al.
Published: (2026)
by: Li, Xiaoqi, et al.
Published: (2026)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
by: Li, Yajie, et al.
Published: (2026)
by: Li, Yajie, et al.
Published: (2026)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
by: Zhu, Juan, et al.
Published: (2026)
by: Zhu, Juan, et al.
Published: (2026)
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)
by: Li, Badi, et al.
Published: (2025)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
Diffusion-Based Imaginative Coordination for Bimanual Manipulation
by: Xu, Huilin, et al.
Published: (2025)
by: Xu, Huilin, et al.
Published: (2025)
Imagine2Servo: Intelligent Visual Servoing with Diffusion-Driven Goal Generation for Robotic Tasks
by: Pathre, Pranjali, et al.
Published: (2024)
by: Pathre, Pranjali, et al.
Published: (2024)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
by: Liu, Chenyv, et al.
Published: (2026)
by: Liu, Chenyv, et al.
Published: (2026)
ImagineNav: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
by: Zhao, Xinxin, et al.
Published: (2024)
by: Zhao, Xinxin, et al.
Published: (2024)
ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
by: Wang, Teng, et al.
Published: (2025)
by: Wang, Teng, et al.
Published: (2025)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
RGBManip: Monocular Image-based Robotic Manipulation through Active Object Pose Estimation
by: An, Boshi, et al.
Published: (2023)
by: An, Boshi, et al.
Published: (2023)
RoboDreamer: Learning Compositional World Models for Robot Imagination
by: Zhou, Siyuan, et al.
Published: (2024)
by: Zhou, Siyuan, et al.
Published: (2024)
Fully Spiking Actor-Critic Neural Network for Robotic Manipulation
by: Zhang, Liwen, et al.
Published: (2025)
by: Zhang, Liwen, et al.
Published: (2025)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
by: Barcellona, Leonardo, et al.
Published: (2024)
by: Barcellona, Leonardo, et al.
Published: (2024)
Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving
by: Sun, Zhengqi, et al.
Published: (2026)
by: Sun, Zhengqi, et al.
Published: (2026)
RAIL: Robot Affordance Imagination with Large Language Models
by: Zhang, Ceng, et al.
Published: (2024)
by: Zhang, Ceng, et al.
Published: (2024)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
by: Su, Yue, et al.
Published: (2024)
by: Su, Yue, et al.
Published: (2024)
Act2Goal: From World Model To General Goal-conditioned Policy
by: Zhou, Pengfei, et al.
Published: (2025)
by: Zhou, Pengfei, et al.
Published: (2025)
When to Trust Imagination: Adaptive Action Execution for World Action Models
by: Wang, Rui, et al.
Published: (2026)
by: Wang, Rui, et al.
Published: (2026)
ImaginationPolicy: Towards Generalizable, Precise and Reliable End-to-End Policy for Robotic Manipulation
by: Lu, Dekun, et al.
Published: (2025)
by: Lu, Dekun, et al.
Published: (2025)
Data-Asymmetric Latent Imagination and Reranking for 3D Robotic Imitation Learning
by: Luo, Lianghao, et al.
Published: (2026)
by: Luo, Lianghao, et al.
Published: (2026)
From Watch to Imagine: Steering Long-horizon Manipulation via Human Demonstration and Future Envisionment
by: Ye, Ke, et al.
Published: (2025)
by: Ye, Ke, et al.
Published: (2025)
Goal-VLA: Image-Generative VLMs as Object-Centric World Models Empowering Zero-shot Robot Manipulation
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
Spatially Generalizable Mobile Manipulation via Adaptive Experience Selection and Dynamic Imagination
by: Zhong, Ping, et al.
Published: (2026)
by: Zhong, Ping, et al.
Published: (2026)
Collision-Inclusive Manipulation Planning for Occluded Object Grasping via Compliant Robot Motions
by: Ren, Kejia, et al.
Published: (2024)
by: Ren, Kejia, et al.
Published: (2024)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
Imagine2touch: Predictive Tactile Sensing for Robotic Manipulation using Efficient Low-Dimensional Signals
by: Ayad, Abdallah, et al.
Published: (2024)
by: Ayad, Abdallah, et al.
Published: (2024)
Physically-Grounded Goal Imagination: Physics-Informed Variational Autoencoder for Self-Supervised Reinforcement Learning
by: Nguyen, Lan Thi Ha, et al.
Published: (2025)
by: Nguyen, Lan Thi Ha, et al.
Published: (2025)
OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
by: Liu, Yushan, et al.
Published: (2026)
by: Liu, Yushan, et al.
Published: (2026)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
by: Bu, Qingwen, et al.
Published: (2025)
by: Bu, Qingwen, et al.
Published: (2025)
H-GAR: A Hierarchical Interaction Framework via Goal-Driven Observation-Action Refinement for Robotic Manipulation
by: Zhu, Yijie, et al.
Published: (2025)
by: Zhu, Yijie, et al.
Published: (2025)
ImagineUAV: Aerial Vision-Language Navigation via World-Action Modeling and Kinodynamic Planning
by: Liu, Xuchen, et al.
Published: (2026)
by: Liu, Xuchen, et al.
Published: (2026)
Instruct2Act: From Human Instruction to Actions Sequencing and Execution via Robot Action Network for Robotic Manipulation
by: Sharma, Archit, et al.
Published: (2026)
by: Sharma, Archit, et al.
Published: (2026)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
by: Zuo, Kuangji, et al.
Published: (2026)
by: Zuo, Kuangji, et al.
Published: (2026)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
by: Wang, Guokang, et al.
Published: (2024)
by: Wang, Guokang, et al.
Published: (2024)
Learning Robotic Policy with Imagined Transition: Mitigating the Trade-off between Robustness and Optimality
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Multimodal Spiking Neural Network for Space Robotic Manipulation
by: Zhang, Liwen, et al.
Published: (2025)
by: Zhang, Liwen, et al.
Published: (2025)
Similar Items
-
AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models
by: Li, Xiaoqi, et al.
Published: (2026) -
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025) -
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
by: Li, Yajie, et al.
Published: (2026) -
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
by: Zhu, Juan, et al.
Published: (2026) -
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)