LaVA-Man: Learning Visual Action Representations for Robot Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Chaoran, Wang, Hengyi, Pang, Yik Lung, Oh, Changjae |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025)
by: Cui, Chenglin, et al.
Published: (2025)
Chain-of-Caption: Training-free improvement of multimodal large language model on referring expression comprehension
by: Pang, Yik Lung, et al.
Published: (2026)
by: Pang, Yik Lung, et al.
Published: (2026)
Sparse multi-view hand-object reconstruction for unseen environments
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
Image Generation as a Visual Planner for Robotic Manipulation
by: Pang, Ye
Published: (2025)
by: Pang, Ye
Published: (2025)
Action-Sketcher: From Reasoning to Action via Visual Sketches for Long-Horizon Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2026)
by: Tan, Huajie, et al.
Published: (2026)
Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation
by: Jiang, Yicheng, et al.
Published: (2026)
by: Jiang, Yicheng, et al.
Published: (2026)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
by: Li, Puhao, et al.
Published: (2024)
by: Li, Puhao, et al.
Published: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
by: Tang, Nan, et al.
Published: (2025)
by: Tang, Nan, et al.
Published: (2025)
Learning Geometrically-Grounded 3D Visual Representations for View-Generalizable Robotic Manipulation
by: Zhang, Di, et al.
Published: (2026)
by: Zhang, Di, et al.
Published: (2026)
Demystifying Action Space Design for Robotic Manipulation Policies
by: Feng, Yuchun, et al.
Published: (2026)
by: Feng, Yuchun, et al.
Published: (2026)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Visual Sculpting: Visually-Aligned Planning Representations for Long-Horizon Robot Clay Sculpting
by: Schaldenbrand, Peter, et al.
Published: (2026)
by: Schaldenbrand, Peter, et al.
Published: (2026)
A Humanoid Visual-Tactile-Action Dataset for Contact-Rich Manipulation
by: Kwon, Eunju, et al.
Published: (2025)
by: Kwon, Eunju, et al.
Published: (2025)
Visual Robotic Manipulation with Depth-Aware Pretraining
by: Wang, Wanying, et al.
Published: (2024)
by: Wang, Wanying, et al.
Published: (2024)
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation
by: Cai, Junhao, et al.
Published: (2026)
by: Cai, Junhao, et al.
Published: (2026)
Bridging Language and Action: A Survey of Language-Conditioned Robot Manipulation
by: Yao, Xiangtong, et al.
Published: (2023)
by: Yao, Xiangtong, et al.
Published: (2023)
LoLA: Long Horizon Latent Action Learning for General Robot Manipulation
by: Wang, Xiaofan, et al.
Published: (2025)
by: Wang, Xiaofan, et al.
Published: (2025)
RoboInter: A Holistic Intermediate Representation Suite Towards Robotic Manipulation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
CLASS: Contrastive Learning via Action Sequence Supervision for Robot Manipulation
by: Lee, Sung-Wook, et al.
Published: (2025)
by: Lee, Sung-Wook, et al.
Published: (2025)
Learning Long-Horizon Robot Manipulation Skills via Privileged Action
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation
by: Xiao, Junjin, et al.
Published: (2026)
by: Xiao, Junjin, et al.
Published: (2026)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
by: Zhu, Xiang, et al.
Published: (2026)
by: Zhu, Xiang, et al.
Published: (2026)
KineSoft: Learning Proprioceptive Manipulation Policies with Soft Robot Hands
by: Yoo, Uksang, et al.
Published: (2025)
by: Yoo, Uksang, et al.
Published: (2025)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
Autoregressive Action Sequence Learning for Robotic Manipulation
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
by: Yu, Wenda, et al.
Published: (2026)
by: Yu, Wenda, et al.
Published: (2026)
Unified Learning of Temporal Task Structure and Action Timing for Bimanual Robot Manipulation
by: Dreher, Christian, et al.
Published: (2026)
by: Dreher, Christian, et al.
Published: (2026)
Toward Visually Realistic Simulation: A Benchmark for Evaluating Robot Manipulation in Simulation
by: Zhu, Yixin, et al.
Published: (2026)
by: Zhu, Yixin, et al.
Published: (2026)
Vi-TacMan: Articulated Object Manipulation via Vision and Touch
by: Cui, Leiyao, et al.
Published: (2025)
by: Cui, Leiyao, et al.
Published: (2025)
Generalist Robot Manipulation beyond Action Labeled Data
by: Spiridonov, Alexander, et al.
Published: (2025)
by: Spiridonov, Alexander, et al.
Published: (2025)
Tac-Man: Tactile-Informed Prior-Free Manipulation of Articulated Objects
by: Zhao, Zihang, et al.
Published: (2024)
by: Zhao, Zihang, et al.
Published: (2024)
Personalization in Human-Robot Interaction through Preference-based Action Representation Learning
by: Wang, Ruiqi, et al.
Published: (2024)
by: Wang, Ruiqi, et al.
Published: (2024)
Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation
by: Zhu, Yihang, et al.
Published: (2025)
by: Zhu, Yihang, et al.
Published: (2025)
VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation
by: Zhou, Huayi, et al.
Published: (2025)
by: Zhou, Huayi, et al.
Published: (2025)
FreeTacMan: Robot-free Visuo-Tactile Data Collection System for Contact-rich Manipulation
by: Wu, Longyan, et al.
Published: (2025)
by: Wu, Longyan, et al.
Published: (2025)
Similar Items
-
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
by: Pang, Yik Lung, et al.
Published: (2024) -
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025) -
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025) -
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025) -
Chain-of-Caption: Training-free improvement of multimodal large language model on referring expression comprehension
by: Pang, Yik Lung, et al.
Published: (2026)