RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lei, Huashuo, Song, Wenxuan, Zhang, Huarui, Pei, Jieyuan, Chen, Jiayi, Yan, Haodong, Zhao, Han, Ding, Pengxiang, Zhang, Zhipeng, Huang, Lida, Wang, Donglin, Wang, Yan, Li, Haoang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model
von: Li, Fuhao, et al.
Veröffentlicht: (2025)
von: Li, Fuhao, et al.
Veröffentlicht: (2025)
CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
von: Zhao, Han, et al.
Veröffentlicht: (2025)
von: Zhao, Han, et al.
Veröffentlicht: (2025)
Towards a Unified Understanding of Robot Manipulation: A Comprehensive Survey
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
GeRM: A Generalist Robotic Model with Mixture-of-experts for Quadruped Robot
von: Song, Wenxuan, et al.
Veröffentlicht: (2024)
von: Song, Wenxuan, et al.
Veröffentlicht: (2024)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought
von: Lei, Haodong, et al.
Veröffentlicht: (2026)
von: Lei, Haodong, et al.
Veröffentlicht: (2026)
ProFD: Prompt-Guided Feature Disentangling for Occluded Person Re-Identification
von: Cui, Can, et al.
Veröffentlicht: (2024)
von: Cui, Can, et al.
Veröffentlicht: (2024)
PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
von: Song, Wenxuan, et al.
Veröffentlicht: (2025)
Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport
von: Sun, Mingyang, et al.
Veröffentlicht: (2025)
von: Sun, Mingyang, et al.
Veröffentlicht: (2025)
Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning
von: Sun, Mingyang, et al.
Veröffentlicht: (2025)
von: Sun, Mingyang, et al.
Veröffentlicht: (2025)
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment
von: Zhao, Han, et al.
Veröffentlicht: (2026)
von: Zhao, Han, et al.
Veröffentlicht: (2026)
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
Mem-W: Latent Memory-Native GUI Agents
von: Zhang, Guibin, et al.
Veröffentlicht: (2026)
von: Zhang, Guibin, et al.
Veröffentlicht: (2026)
Embodied Robot Manipulation in the Era of Foundation Models: Planning and Learning Perspectives
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective
von: Wang, Yuyao, et al.
Veröffentlicht: (2026)
von: Wang, Yuyao, et al.
Veröffentlicht: (2026)
EvolMem: A Cognitive-Driven Benchmark for Multi-Session Dialogue Memory
von: Shen, Ye, et al.
Veröffentlicht: (2026)
von: Shen, Ye, et al.
Veröffentlicht: (2026)
MemGUI-Bench: Benchmarking Memory of Mobile GUI Agents in Dynamic Environments
von: Liu, Guangyi, et al.
Veröffentlicht: (2026)
von: Liu, Guangyi, et al.
Veröffentlicht: (2026)
FlowVLA: Visual Chain of Thought-based Motion Reasoning for Vision-Language-Action Models
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
von: Zhong, Zhide, et al.
Veröffentlicht: (2025)
CloneMem: Benchmarking Long-Term Memory for AI Clones
von: Hu, Sen, et al.
Veröffentlicht: (2026)
von: Hu, Sen, et al.
Veröffentlicht: (2026)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
von: Atreya, Pranav, et al.
Veröffentlicht: (2025)
GEVRM: Goal-Expressive Video Generation Model For Robust Visual Manipulation
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
MemFlow: Intent-Driven Memory Orchestration for Small Language Model Agents
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
von: Chen, Jiayi, et al.
Veröffentlicht: (2026)
MemGen: Weaving Generative Latent Memory for Self-Evolving Agents
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
Rethinking Target Label Conditioning in Adversarial Attacks: A 2D Tensor-Guided Generative Approach
von: Liu, Hangyu, et al.
Veröffentlicht: (2025)
von: Liu, Hangyu, et al.
Veröffentlicht: (2025)
Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
von: Bai, Shuanghao, et al.
Veröffentlicht: (2025)
MemEvolve: Meta-Evolution of Agent Memory Systems
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
von: Zhang, Guibin, et al.
Veröffentlicht: (2025)
RoboFAC: A Comprehensive Framework for Robotic Failure Analysis and Correction
von: Ye, Zewei, et al.
Veröffentlicht: (2025)
von: Ye, Zewei, et al.
Veröffentlicht: (2025)
MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models
von: Zhao, Han, et al.
Veröffentlicht: (2025)
von: Zhao, Han, et al.
Veröffentlicht: (2025)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
von: Liu, Chengzhi, et al.
Veröffentlicht: (2026)
von: Liu, Chengzhi, et al.
Veröffentlicht: (2026)
MemVerse: Multimodal Memory for Lifelong Learning Agents
von: Liu, Junming, et al.
Veröffentlicht: (2025)
von: Liu, Junming, et al.
Veröffentlicht: (2025)
Mem-T: Densifying Rewards for Long-Horizon Memory Agents
von: Yue, Yanwei, et al.
Veröffentlicht: (2026)
von: Yue, Yanwei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline
von: Song, Wenxuan, et al.
Veröffentlicht: (2026) -
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
von: Song, Wenxuan, et al.
Veröffentlicht: (2025) -
CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding
von: Song, Wenxuan, et al.
Veröffentlicht: (2025) -
Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process
von: Chen, Jiayi, et al.
Veröffentlicht: (2025) -
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance
von: Song, Wenxuan, et al.
Veröffentlicht: (2026)