Diagnose, Correct, and Learn from Manipulation Failures via Visual Symbols
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Xianchao, Zhou, Xinyu, Li, Youcheng, Shi, Jiayou, Li, Tianle, Chen, Liangming, Ren, Lei, Li, Yong-Lu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Actionable Manipulation Recovery via Counterfactual Failure Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026)
von: Li, Dayou, et al.
Veröffentlicht: (2026)
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
von: Li, Yuyang, et al.
Veröffentlicht: (2025)
von: Li, Yuyang, et al.
Veröffentlicht: (2025)
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
von: Xu, Yifu, et al.
Veröffentlicht: (2026)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
von: Li, Kailin, et al.
Veröffentlicht: (2025)
von: Li, Kailin, et al.
Veröffentlicht: (2025)
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
von: Dong, Runpei, et al.
Veröffentlicht: (2026)
von: Dong, Runpei, et al.
Veröffentlicht: (2026)
Learning Manipulation by Predicting Interaction
von: Zeng, Jia, et al.
Veröffentlicht: (2024)
von: Zeng, Jia, et al.
Veröffentlicht: (2024)
VLN-MME: Diagnosing MLLMs as Language-guided Visual Navigation agents
von: Zhao, Xunyi, et al.
Veröffentlicht: (2025)
von: Zhao, Xunyi, et al.
Veröffentlicht: (2025)
Structural Action Transformer for 3D Dexterous Manipulation
von: Lei, Xiaohan, et al.
Veröffentlicht: (2026)
von: Lei, Xiaohan, et al.
Veröffentlicht: (2026)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
von: Li, Puhao, et al.
Veröffentlicht: (2024)
von: Li, Puhao, et al.
Veröffentlicht: (2024)
ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation
von: Qian, Jingjing, et al.
Veröffentlicht: (2026)
von: Qian, Jingjing, et al.
Veröffentlicht: (2026)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
von: Liu, Enguang, et al.
Veröffentlicht: (2025)
von: Liu, Enguang, et al.
Veröffentlicht: (2025)
GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
von: Qian, Jingjing, et al.
Veröffentlicht: (2025)
von: Qian, Jingjing, et al.
Veröffentlicht: (2025)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
Mash, Spread, Slice! Learning to Manipulate Object States via Visual Spatial Progress
von: Mandikal, Priyanka, et al.
Veröffentlicht: (2025)
von: Mandikal, Priyanka, et al.
Veröffentlicht: (2025)
Neuro-Symbolic Manipulation Understanding with Enriched Semantic Event Chains
von: Ziaeetabar, Fatemeh
Veröffentlicht: (2026)
von: Ziaeetabar, Fatemeh
Veröffentlicht: (2026)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
von: Xiong, Chuyan, et al.
Veröffentlicht: (2024)
von: Xiong, Chuyan, et al.
Veröffentlicht: (2024)
Multi-robot autonomous 3D reconstruction using Gaussian splatting with Semantic guidance
von: Zeng, Jing, et al.
Veröffentlicht: (2024)
von: Zeng, Jing, et al.
Veröffentlicht: (2024)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
$χ_{0}$: Resource-Aware Robust Manipulation via Taming Distributional Inconsistencies
von: Yu, Checheng, et al.
Veröffentlicht: (2026)
von: Yu, Checheng, et al.
Veröffentlicht: (2026)
YOCO: You Only Calibrate Once for Accurate Extrinsic Parameter in LiDAR-Camera Systems
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
von: Zeng, Tianle, et al.
Veröffentlicht: (2024)
Think Proprioceptively: Embodied Visual Reasoning for VLA Manipulation
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
TAPTRv2: Attention-based Position Update Improves Tracking Any Point
von: Li, Hongyang, et al.
Veröffentlicht: (2024)
von: Li, Hongyang, et al.
Veröffentlicht: (2024)
DeformMaster: An Interactive Physics-Neural World Model for Deformable Objects from Videos
von: Li, Can, et al.
Veröffentlicht: (2026)
von: Li, Can, et al.
Veröffentlicht: (2026)
Physically Ground Commonsense Knowledge for Articulated Object Manipulation with Analytic Concepts
von: Wei, Jiude, et al.
Veröffentlicht: (2025)
von: Wei, Jiude, et al.
Veröffentlicht: (2025)
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
TAPTR: Tracking Any Point with Transformers as Detection
von: Li, Hongyang, et al.
Veröffentlicht: (2024)
von: Li, Hongyang, et al.
Veröffentlicht: (2024)
Learning Generalizable 3D Manipulation With 10 Demonstrations
von: Ren, Yu, et al.
Veröffentlicht: (2024)
von: Ren, Yu, et al.
Veröffentlicht: (2024)
Prior Does Matter: Visual Navigation via Denoising Diffusion Bridge Models
von: Ren, Hao, et al.
Veröffentlicht: (2025)
von: Ren, Hao, et al.
Veröffentlicht: (2025)
DynaRend: Learning 3D Dynamics via Masked Future Rendering for Robotic Manipulation
von: Tian, Jingyi, et al.
Veröffentlicht: (2025)
von: Tian, Jingyi, et al.
Veröffentlicht: (2025)
STRNet: Visual Navigation with Spatio-Temporal Representation through Dynamic Graph Aggregation
von: Ren, Hao, et al.
Veröffentlicht: (2026)
von: Ren, Hao, et al.
Veröffentlicht: (2026)
ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations
von: Zeng, Qiyuan, et al.
Veröffentlicht: (2025)
von: Zeng, Qiyuan, et al.
Veröffentlicht: (2025)
R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation
von: Xu, Xiuwei, et al.
Veröffentlicht: (2025)
von: Xu, Xiuwei, et al.
Veröffentlicht: (2025)
Scaling Cross-Environment Failure Reasoning Data for Vision-Language Robotic Manipulation
von: Pacaud, Paul, et al.
Veröffentlicht: (2025)
von: Pacaud, Paul, et al.
Veröffentlicht: (2025)
Click to Grasp: Zero-Shot Precise Manipulation via Visual Diffusion Descriptors
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2024)
von: Tsagkas, Nikolaos, et al.
Veröffentlicht: (2024)
SF-Loc: A Visual Mapping and Geo-Localization System based on Sparse Visual Structure Frames
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhou, Yuxuan, et al.
Veröffentlicht: (2024)
DSM: Constructing a Diverse Semantic Map for 3D Visual Grounding
von: Xie, Qinghongbing, et al.
Veröffentlicht: (2025)
von: Xie, Qinghongbing, et al.
Veröffentlicht: (2025)
SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
von: Zhang, Junjie, et al.
Veröffentlicht: (2024)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
von: Liu, Chenyv, et al.
Veröffentlicht: (2026)
von: Liu, Chenyv, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning Actionable Manipulation Recovery via Counterfactual Failure Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026) -
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
von: Li, Yuyang, et al.
Veröffentlicht: (2025) -
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
von: Xu, Yifu, et al.
Veröffentlicht: (2026) -
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
von: Jiang, Zebin, et al.
Veröffentlicht: (2025) -
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
von: Li, Kailin, et al.
Veröffentlicht: (2025)