MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Runhao, Guo, Wenkai, Wu, Zhenyu, Wang, Changyuan, Deng, Haoyuan, Weng, Zhenyu, Tan, Yap-Peng, Wang, Ziwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
SafeBimanual: Diffusion-based Trajectory Optimization for Safe Bimanual Manipulation
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025)
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
DAM-VLA: A Dynamic Action Model-Based Vision-Language-Action Framework for Robot Manipulation
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
von: Peng, Xiongfeng, et al.
Veröffentlicht: (2026)
PriorVLA: Prior-Preserving Adaptation for Vision-Language-Action Models
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
von: Guo, Xinyu, et al.
Veröffentlicht: (2026)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
von: Fan, Yiguo, et al.
Veröffentlicht: (2025)
Audio-VLA: Adding Contact Audio Perception to Vision-Language-Action Model for Robotic Manipulation
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
von: Wei, Xiangyi, et al.
Veröffentlicht: (2025)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025)
von: Miao, Cui, et al.
Veröffentlicht: (2025)
AC^2-VLA: Action-Context-Aware Adaptive Computation in Vision-Language-Action Models for Efficient Robotic Manipulation
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
von: Yu, Wenda, et al.
Veröffentlicht: (2026)
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation
von: Xue, Yuquan, et al.
Veröffentlicht: (2025)
von: Xue, Yuquan, et al.
Veröffentlicht: (2025)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
von: Xie, Haozhe, et al.
Veröffentlicht: (2026)
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
von: Liu, Zhenyang, et al.
Veröffentlicht: (2026)
von: Liu, Zhenyang, et al.
Veröffentlicht: (2026)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
Spatial Memory for Out-of-Vision Manipulation in Vision-Language-Action
von: Li, Pengteng, et al.
Veröffentlicht: (2026)
von: Li, Pengteng, et al.
Veröffentlicht: (2026)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
von: Wen, Junjie, et al.
Veröffentlicht: (2024)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
von: Sun, Jianli, et al.
Veröffentlicht: (2026)
von: Sun, Jianli, et al.
Veröffentlicht: (2026)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning
von: Deng, Haoyuan, et al.
Veröffentlicht: (2026)
von: Deng, Haoyuan, et al.
Veröffentlicht: (2026)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
von: Khan, Muhamamd Haris, et al.
Veröffentlicht: (2025)
GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions
von: Huang, Helong, et al.
Veröffentlicht: (2025)
von: Huang, Helong, et al.
Veröffentlicht: (2025)
StereoVLA: Enhancing Vision-Language-Action Models with Stereo Vision
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
von: Deng, Shengliang, et al.
Veröffentlicht: (2025)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
von: Weng, Wuding, et al.
Veröffentlicht: (2026)
von: Weng, Wuding, et al.
Veröffentlicht: (2026)
EveryDayVLA: A Vision-Language-Action Model for Affordable Robotic Manipulation
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
von: Chopra, Samarth, et al.
Veröffentlicht: (2025)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
von: Zhao, Han, et al.
Veröffentlicht: (2025)
von: Zhao, Han, et al.
Veröffentlicht: (2025)
AnchorVLA4D: an Anchor-Based Spatial-Temporal Vision-Language-Action Model for Robotic Manipulation
von: Zhu, Juan, et al.
Veröffentlicht: (2026)
von: Zhu, Juan, et al.
Veröffentlicht: (2026)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models
von: Zhang, Zongzheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zongzheng, et al.
Veröffentlicht: (2025)
Green-VLA: Staged Vision-Language-Action Model for Generalist Robots
von: Apanasevich, I., et al.
Veröffentlicht: (2026)
von: Apanasevich, I., et al.
Veröffentlicht: (2026)
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
von: Guo, Peizheng, et al.
Veröffentlicht: (2026)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
von: Hu, Yucheng, et al.
Veröffentlicht: (2026)
von: Hu, Yucheng, et al.
Veröffentlicht: (2026)
InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation
von: Cai, Junhao, et al.
Veröffentlicht: (2026)
von: Cai, Junhao, et al.
Veröffentlicht: (2026)
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
SignVLA: A Gloss-Free Vision-Language-Action Framework for Real-Time Sign Language-Guided Robotic Manipulation
von: Tan, Xinyu, et al.
Veröffentlicht: (2026)
von: Tan, Xinyu, et al.
Veröffentlicht: (2026)
SA-VLA: Spatially-Aware Flow-Matching for Vision-Language-Action Reinforcement Learning
von: Pan, Xu, et al.
Veröffentlicht: (2026)
von: Pan, Xu, et al.
Veröffentlicht: (2026)
SG-VLA: Learning Spatially-Grounded Vision-Language-Action Models for Mobile Manipulation
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
von: Tu, Ruisen, et al.
Veröffentlicht: (2026)
TacVLA: Contact-Aware Tactile Fusion for Robust Vision-Language-Action Manipulation
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
von: Zhang, Kaidi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
von: Guo, Wenkai, et al.
Veröffentlicht: (2025) -
SafeBimanual: Diffusion-based Trajectory Optimization for Safe Bimanual Manipulation
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025) -
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
von: Wu, Zhenyu, et al.
Veröffentlicht: (2025) -
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
von: Shi, Hao, et al.
Veröffentlicht: (2025) -
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)