OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yushan, Sun, Peibo, Li, Shoujie, Xie, Yifan, Zhang, Lingfeng, Chao, Xintao, Dong, Shiyuan, Chen, Fang, Zhang, Xiao-Ping, Ding, Wenbo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AVR: Active Vision-Driven Precise Robot Manipulation with Viewpoint and Focal Length Optimization
by: Liu, Yushan, et al.
Published: (2025)
by: Liu, Yushan, et al.
Published: (2025)
Exo-ViHa: A Cross-Platform Exoskeleton System with Visual and Haptic Feedback for Efficient Dexterous Skill Learning
by: Chao, Xintao, et al.
Published: (2025)
by: Chao, Xintao, et al.
Published: (2025)
JailWAM: Jailbreaking World Action Models in Robot Control
by: Liu, Hanqing, et al.
Published: (2026)
by: Liu, Hanqing, et al.
Published: (2026)
HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models
by: Feng, Qiuxuan, et al.
Published: (2026)
by: Feng, Qiuxuan, et al.
Published: (2026)
SimLiquid: A Simulation‐Based Liquid Perception Pipeline for Robot Liquid Manipulation
by: Yan Huang, et al.
Published: (2025)
by: Yan Huang, et al.
Published: (2025)
Percept-WAM: Perception-Enhanced World-Awareness-Action Model for Robust End-to-End Autonomous Driving
by: Han, Jianhua, et al.
Published: (2025)
by: Han, Jianhua, et al.
Published: (2025)
Latent-WAM: Latent World Action Modeling for End-to-End Autonomous Driving
by: Wang, Linbo, et al.
Published: (2026)
by: Wang, Linbo, et al.
Published: (2026)
Fast-WAM: Do World Action Models Need Test-time Future Imagination?
by: Yuan, Tianyuan, et al.
Published: (2026)
by: Yuan, Tianyuan, et al.
Published: (2026)
FUTURE-VLA: Forecasting Unified Trajectories Under Real-time Execution
by: Fan, Jingjing, et al.
Published: (2026)
by: Fan, Jingjing, et al.
Published: (2026)
Master Micro Residual Correction with Adaptive Tactile Fusion and Force-Mixed Control for Contact-Rich Manipulation
by: Li, Xingting, et al.
Published: (2026)
by: Li, Xingting, et al.
Published: (2026)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
by: Liu, Jinkun, et al.
Published: (2026)
by: Liu, Jinkun, et al.
Published: (2026)
Depth Restoration of Hand-Held Transparent Objects for Human-to-Robot Handover
by: Yu, Ran, et al.
Published: (2024)
by: Yu, Ran, et al.
Published: (2024)
CKT-WAM: Parameter-Efficient Context Knowledge Transfer Between World Action Models
by: Jiang, Yuhua, et al.
Published: (2026)
by: Jiang, Yuhua, et al.
Published: (2026)
DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving
by: Shi, Chen, et al.
Published: (2026)
by: Shi, Chen, et al.
Published: (2026)
Visual-tactile Fusion for Transparent Object Grasping in Complex Backgrounds
by: Li, Shoujie, et al.
Published: (2022)
by: Li, Shoujie, et al.
Published: (2022)
UMIGen: A Unified Framework for Egocentric Point Cloud Generation and Cross-Embodiment Robotic Imitation Learning
by: Huang, Yan, et al.
Published: (2025)
by: Huang, Yan, et al.
Published: (2025)
Dual-modal Tactile E-skin: Enabling Bidirectional Human-Robot Interaction via Integrated Tactile Perception and Feedback
by: Mu, Shilong, et al.
Published: (2024)
by: Mu, Shilong, et al.
Published: (2024)
Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation
by: Zhang, Wenbo, et al.
Published: (2025)
by: Zhang, Wenbo, et al.
Published: (2025)
Chemistry3D: Robotic Interaction Benchmark for Chemistry Experiments
by: Li, Shoujie, et al.
Published: (2024)
by: Li, Shoujie, et al.
Published: (2024)
SSyncOA: Self-synchronizing Object-aligned Watermarking to Resist Cropping-paste Attacks
by: Zhao, Chengxin, et al.
Published: (2024)
by: Zhao, Chengxin, et al.
Published: (2024)
Growing from Exploration: A self-exploring framework for robots based on foundation models
by: Li, Shoujie, et al.
Published: (2024)
by: Li, Shoujie, et al.
Published: (2024)
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
by: Xie, Yifan, et al.
Published: (2026)
by: Xie, Yifan, et al.
Published: (2026)
GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation
by: Shao, Beichen, et al.
Published: (2026)
by: Shao, Beichen, et al.
Published: (2026)
Weather-Conditioned Branch Routing for Robust LiDAR-Radar 3D Object Detection
by: Li, Hongsheng, et al.
Published: (2026)
by: Li, Hongsheng, et al.
Published: (2026)
SandWorm: Event-based Visuotactile Perception with Active Vibration for Screw-Actuated Robot in Granular Media
by: Li, Shoujie, et al.
Published: (2026)
by: Li, Shoujie, et al.
Published: (2026)
Cover Image, Volume 42, Number 6, September 2025
by: Yan Huang, et al.
Published: (2025)
by: Yan Huang, et al.
Published: (2025)
RoboAfford++: A Generative AI-Enhanced Dataset for Multimodal Affordance Learning in Robotic Manipulation and Navigation
by: Hao, Xiaoshuai, et al.
Published: (2025)
by: Hao, Xiaoshuai, et al.
Published: (2025)
When Vision Meets Touch: A Contemporary Review for Visuotactile Sensors from the Signal Processing Perspective
by: Li, Shoujie, et al.
Published: (2024)
by: Li, Shoujie, et al.
Published: (2024)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
Curriculum-based Sensing Reduction in Simulation to Real-World Transfer for In-hand Manipulation
by: Tao, Lingfeng, et al.
Published: (2023)
by: Tao, Lingfeng, et al.
Published: (2023)
VTire: A Bimodal Visuotactile Tire with High-Resolution Sensing Capability
by: Li, Shoujie, et al.
Published: (2025)
by: Li, Shoujie, et al.
Published: (2025)
ALORE: Autonomous Large-Object Rearrangement with a Legged Manipulator
by: Bi, Zhihai, et al.
Published: (2026)
by: Bi, Zhihai, et al.
Published: (2026)
Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control
by: Fu, Xiao, et al.
Published: (2025)
by: Fu, Xiao, et al.
Published: (2025)
MuxHand: A Cable-driven Dexterous Robotic Hand Using Time-division Multiplexing Motors
by: Xu, Jianle, et al.
Published: (2024)
by: Xu, Jianle, et al.
Published: (2024)
RAM-Net: Expressive Linear Attention with Selectively Addressable Memory
by: Xiao, Kaicheng, et al.
Published: (2026)
by: Xiao, Kaicheng, et al.
Published: (2026)
Universal Visuo-Tactile Video Understanding for Embodied Interaction
by: Xie, Yifan, et al.
Published: (2025)
by: Xie, Yifan, et al.
Published: (2025)
EXOT: Exit-aware Object Tracker for Safe Robotic Manipulation of Moving Object
by: Kim, Hyunseo, et al.
Published: (2023)
by: Kim, Hyunseo, et al.
Published: (2023)
$τ_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
by: Zhou, Pengfei, et al.
Published: (2026)
by: Zhou, Pengfei, et al.
Published: (2026)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
by: Li, Yajie, et al.
Published: (2026)
by: Li, Yajie, et al.
Published: (2026)
Similar Items
-
AVR: Active Vision-Driven Precise Robot Manipulation with Viewpoint and Focal Length Optimization
by: Liu, Yushan, et al.
Published: (2025) -
Exo-ViHa: A Cross-Platform Exoskeleton System with Visual and Haptic Feedback for Efficient Dexterous Skill Learning
by: Chao, Xintao, et al.
Published: (2025) -
JailWAM: Jailbreaking World Action Models in Robot Control
by: Liu, Hanqing, et al.
Published: (2026) -
HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models
by: Feng, Qiuxuan, et al.
Published: (2026) -
SimLiquid: A Simulation‐Based Liquid Perception Pipeline for Robot Liquid Manipulation
by: Yan Huang, et al.
Published: (2025)