GravMAD: Grounded Spatial Value Maps Guided Action Diffusion for Generalized 3D Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Yangtao, Chen, Zixuan, Yin, Junhui, Huo, Jing, Tian, Pinzhuo, Shi, Jieqi, Gao, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeCo: Task Decomposition and Skill Composition for Zero-Shot Generalization in Long-Horizon 3D Manipulation
by: Chen, Zixuan, et al.
Published: (2025)
by: Chen, Zixuan, et al.
Published: (2025)
RoboHiMan: A Hierarchical Evaluation Paradigm for Compositional Generalization in Long-Horizon Manipulation
by: Chen, Yangtao, et al.
Published: (2025)
by: Chen, Yangtao, et al.
Published: (2025)
RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation
by: Chen, Zixuan, et al.
Published: (2025)
by: Chen, Zixuan, et al.
Published: (2025)
RoTri-Diff: A Spatial Robot-Object Triadic Interaction-Guided Diffusion Model for Bimanual Manipulation
by: Chen, Zixuan, et al.
Published: (2026)
by: Chen, Zixuan, et al.
Published: (2026)
ManiLong-Shot: Interaction-Aware One-Shot Imitation Learning for Long-Horizon Manipulation
by: Chen, Zixuan, et al.
Published: (2025)
by: Chen, Zixuan, et al.
Published: (2025)
ST-VLA: Enabling 4D-Aware Spatiotemporal Understanding for General Robot Manipulation
by: Wu, You, et al.
Published: (2026)
by: Wu, You, et al.
Published: (2026)
MoMaStage: Skill-State Graph Guided Planning and Closed-Loop Execution for Long-Horizon Indoor Mobile Manipulation
by: Li, Chenxu, et al.
Published: (2026)
by: Li, Chenxu, et al.
Published: (2026)
LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
by: Ding, Hongyu, et al.
Published: (2025)
by: Ding, Hongyu, et al.
Published: (2025)
V-Dreamer: Automating Robotic Simulation and Trajectory Synthesis via Video Generation Priors
by: He, Songjia, et al.
Published: (2026)
by: He, Songjia, et al.
Published: (2026)
INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval
by: Fang, YukTungSamuel, et al.
Published: (2026)
by: Fang, YukTungSamuel, et al.
Published: (2026)
AdaClearGrasp: Learning Adaptive Clearing for Zero-Shot Robust Dexterous Grasping in Densely Cluttered Environments
by: Chen, Zixuan, et al.
Published: (2026)
by: Chen, Zixuan, et al.
Published: (2026)
SEPT: Standard-Definition Map Enhanced Scene Perception and Topology Reasoning for Autonomous Driving
by: Pei, Muleilan, et al.
Published: (2025)
by: Pei, Muleilan, et al.
Published: (2025)
SEA: Semantic Map Prediction for Active Exploration of Uncertain Areas
by: Ding, Hongyu, et al.
Published: (2025)
by: Ding, Hongyu, et al.
Published: (2025)
STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation
by: Tian, Yuxuan, et al.
Published: (2026)
by: Tian, Yuxuan, et al.
Published: (2026)
TopoDiffuser: A Diffusion-Based Multimodal Trajectory Prediction Model with Topometric Maps
by: Xu, Zehui, et al.
Published: (2025)
by: Xu, Zehui, et al.
Published: (2025)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
by: Tu, Ruisen, et al.
Published: (2026)
by: Tu, Ruisen, et al.
Published: (2026)
Subgoal Diffuser: Coarse-to-fine Subgoal Generation to Guide Model Predictive Control for Robot Manipulation
by: Huang, Zixuan, et al.
Published: (2024)
by: Huang, Zixuan, et al.
Published: (2024)
Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation
by: Ding, Hongyu, et al.
Published: (2026)
by: Ding, Hongyu, et al.
Published: (2026)
Perceiving, Reasoning, Adapting: A Dual-Layer Framework for VLM-Guided Precision Robotic Manipulation
by: Jia, Qingxuan, et al.
Published: (2025)
by: Jia, Qingxuan, et al.
Published: (2025)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
ST4VLA: Spatially Guided Training for Vision-Language-Action Models
by: Ye, Jinhui, et al.
Published: (2026)
by: Ye, Jinhui, et al.
Published: (2026)
Latent Action Diffusion for Cross-Embodiment Manipulation
by: Bauer, Erik, et al.
Published: (2025)
by: Bauer, Erik, et al.
Published: (2025)
FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation
by: Shao, Dian, et al.
Published: (2026)
by: Shao, Dian, et al.
Published: (2026)
mindmap: Spatial Memory in Deep Feature Maps for 3D Action Policies
by: Steiner, Remo, et al.
Published: (2025)
by: Steiner, Remo, et al.
Published: (2025)
Learning Spatial-Aware Manipulation Ordering
by: Yan, Yuxiang, et al.
Published: (2025)
by: Yan, Yuxiang, et al.
Published: (2025)
FM-Fusion: Instance-aware Semantic Mapping Boosted by Vision-Language Foundation Models
by: Liu, Chuhao, et al.
Published: (2024)
by: Liu, Chuhao, et al.
Published: (2024)
Multimodal Diffusion Forcing for Forceful Manipulation
by: Huang, Zixuan, et al.
Published: (2025)
by: Huang, Zixuan, et al.
Published: (2025)
Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action
by: Li, Pengteng, et al.
Published: (2026)
by: Li, Pengteng, et al.
Published: (2026)
AIM: Intent-Aware Unified world action Modeling with Spatial Value Maps
by: Fan, Liaoyuan, et al.
Published: (2026)
by: Fan, Liaoyuan, et al.
Published: (2026)
Action-aware Dynamic Pruning for Efficient Vision-Language-Action Manipulation
by: Pei, Xiaohuan, et al.
Published: (2025)
by: Pei, Xiaohuan, et al.
Published: (2025)
VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Efficient Manipulation-Enhanced Semantic Mapping With Uncertainty-Informed Action Selection
by: Dengler, Nils, et al.
Published: (2025)
by: Dengler, Nils, et al.
Published: (2025)
Contact Coverage-Guided Exploration for General-Purpose Dexterous Manipulation
by: Liu, Zixuan, et al.
Published: (2026)
by: Liu, Zixuan, et al.
Published: (2026)
Joint Optimization-based Targetless Extrinsic Calibration for Multiple LiDARs and GNSS-Aided INS of Ground Vehicles
by: Wang, Junhui, et al.
Published: (2025)
by: Wang, Junhui, et al.
Published: (2025)
Trace-Focused Diffusion Policy for Multi-Modal Action Disambiguation in Long-Horizon Robotic Manipulation
by: Hu, Yuxuan, et al.
Published: (2026)
by: Hu, Yuxuan, et al.
Published: (2026)
HACMan++: Spatially-Grounded Motion Primitives for Manipulation
by: Jiang, Bowen, et al.
Published: (2024)
by: Jiang, Bowen, et al.
Published: (2024)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
by: Su, Yue, et al.
Published: (2024)
by: Su, Yue, et al.
Published: (2024)
Self-Guided Action Diffusion
by: Malhotra, Rhea, et al.
Published: (2025)
by: Malhotra, Rhea, et al.
Published: (2025)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
by: Zhu, Juan, et al.
Published: (2026)
by: Zhu, Juan, et al.
Published: (2026)
Similar Items
-
DeCo: Task Decomposition and Skill Composition for Zero-Shot Generalization in Long-Horizon 3D Manipulation
by: Chen, Zixuan, et al.
Published: (2025) -
RoboHiMan: A Hierarchical Evaluation Paradigm for Compositional Generalization in Long-Horizon Manipulation
by: Chen, Yangtao, et al.
Published: (2025) -
RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation
by: Chen, Zixuan, et al.
Published: (2025) -
RoTri-Diff: A Spatial Robot-Object Triadic Interaction-Guided Diffusion Model for Bimanual Manipulation
by: Chen, Zixuan, et al.
Published: (2026) -
ManiLong-Shot: Interaction-Aware One-Shot Imitation Learning for Long-Horizon Manipulation
by: Chen, Zixuan, et al.
Published: (2025)