ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Yandan, Zeng, Shuang, Lin, Tong, Chang, Xinyuan, Qi, Dekang, Xiao, Junjin, Liu, Haoyun, Chen, Ronghan, Chen, Yuzhi, Huo, Dongjie, Xiong, Feng, Wei, Xing, Ma, Zhiheng, Xu, Mu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
por: Chen, Yuzhi, et al.
Publicado: (2026)
por: Chen, Yuzhi, et al.
Publicado: (2026)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
por: Huo, Dongjie, et al.
Publicado: (2026)
por: Huo, Dongjie, et al.
Publicado: (2026)
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation
por: Xiao, Junjin, et al.
Publicado: (2026)
por: Xiao, Junjin, et al.
Publicado: (2026)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
por: Cai, Zhejia, et al.
Publicado: (2025)
por: Cai, Zhejia, et al.
Publicado: (2025)
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
por: Xiao, Junjin, et al.
Publicado: (2025)
por: Xiao, Junjin, et al.
Publicado: (2025)
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models
por: Tang, Zuojin, et al.
Publicado: (2026)
por: Tang, Zuojin, et al.
Publicado: (2026)
ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
por: Chu, Zedong, et al.
Publicado: (2026)
por: Chu, Zedong, et al.
Publicado: (2026)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
por: Wang, Hongyu, et al.
Publicado: (2025)
por: Wang, Hongyu, et al.
Publicado: (2025)
ABot-OCR Technical Report
por: Jiang, Kaitao, et al.
Publicado: (2026)
por: Jiang, Kaitao, et al.
Publicado: (2026)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
por: Qi, Dekang, et al.
Publicado: (2026)
por: Qi, Dekang, et al.
Publicado: (2026)
JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation
por: Zeng, Shuang, et al.
Publicado: (2025)
por: Zeng, Shuang, et al.
Publicado: (2025)
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching
por: Chen, Jiayi, et al.
Publicado: (2026)
por: Chen, Jiayi, et al.
Publicado: (2026)
Learning Generalizable 3D Manipulation With 10 Demonstrations
por: Ren, Yu, et al.
Publicado: (2024)
por: Ren, Yu, et al.
Publicado: (2024)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
por: Shi, Hao, et al.
Publicado: (2025)
por: Shi, Hao, et al.
Publicado: (2025)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
por: Li, Runhao, et al.
Publicado: (2025)
por: Li, Runhao, et al.
Publicado: (2025)
EveryDayVLA: A Vision-Language-Action Model for Affordable Robotic Manipulation
por: Chopra, Samarth, et al.
Publicado: (2025)
por: Chopra, Samarth, et al.
Publicado: (2025)
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
por: Li, Qixiu, et al.
Publicado: (2024)
por: Li, Qixiu, et al.
Publicado: (2024)
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation
por: Lian, Shijie, et al.
Publicado: (2026)
por: Lian, Shijie, et al.
Publicado: (2026)
VLA-LPAF: Lightweight Perspective-Adaptive Fusion for Vision-Language-Action to Enable More Unconstrained Robotic Manipulation
por: Bian, Jinyue, et al.
Publicado: (2025)
por: Bian, Jinyue, et al.
Publicado: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
por: Wen, Junjie, et al.
Publicado: (2024)
por: Wen, Junjie, et al.
Publicado: (2024)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
por: Xie, Haozhe, et al.
Publicado: (2026)
por: Xie, Haozhe, et al.
Publicado: (2026)
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models
por: Liu, Haoyun, et al.
Publicado: (2026)
por: Liu, Haoyun, et al.
Publicado: (2026)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
por: Yang, Shuai, et al.
Publicado: (2025)
por: Yang, Shuai, et al.
Publicado: (2025)
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
por: Jiang, Yuming, et al.
Publicado: (2025)
por: Jiang, Yuming, et al.
Publicado: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
por: Guo, Heyu, et al.
Publicado: (2025)
por: Guo, Heyu, et al.
Publicado: (2025)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
por: Yang, Tianshuo, et al.
Publicado: (2026)
por: Yang, Tianshuo, et al.
Publicado: (2026)
ST-$π$: Structured SpatioTemporal VLA for Robotic Manipulation
por: Ma, Chuanhao, et al.
Publicado: (2026)
por: Ma, Chuanhao, et al.
Publicado: (2026)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
por: Li, Qiwei, et al.
Publicado: (2026)
por: Li, Qiwei, et al.
Publicado: (2026)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
por: Wang, Qiuyue, et al.
Publicado: (2026)
por: Wang, Qiuyue, et al.
Publicado: (2026)
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
por: Song, Wenxuan, et al.
Publicado: (2025)
por: Song, Wenxuan, et al.
Publicado: (2025)
SeqGrowGraph: Learning Lane Topology as a Chain of Graph Expansions
por: Xie, Mengwei, et al.
Publicado: (2025)
por: Xie, Mengwei, et al.
Publicado: (2025)
VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation
por: Zhou, Hanyu, et al.
Publicado: (2025)
por: Zhou, Hanyu, et al.
Publicado: (2025)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
por: Yin, Zhenhan, et al.
Publicado: (2025)
por: Yin, Zhenhan, et al.
Publicado: (2025)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
por: Ding, Pengxiang, et al.
Publicado: (2023)
por: Ding, Pengxiang, et al.
Publicado: (2023)
A Pragmatic VLA Foundation Model
por: Wu, Wei, et al.
Publicado: (2026)
por: Wu, Wei, et al.
Publicado: (2026)
Marrying NeRF with Feature Matching for One-step Pose Estimation
por: Chen, Ronghan, et al.
Publicado: (2024)
por: Chen, Ronghan, et al.
Publicado: (2024)
SAGA: Surface-Aligned Gaussian Avatar
por: Chen, Ronghan, et al.
Publicado: (2024)
por: Chen, Ronghan, et al.
Publicado: (2024)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
por: Shen, Yichao, et al.
Publicado: (2025)
por: Shen, Yichao, et al.
Publicado: (2025)
Physical Autoregressive Model for Robotic Manipulation without Action Pretraining
por: Song, Zijian, et al.
Publicado: (2025)
por: Song, Zijian, et al.
Publicado: (2025)
Ejemplares similares
-
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
por: Chen, Yuzhi, et al.
Publicado: (2026) -
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
por: Huo, Dongjie, et al.
Publicado: (2026) -
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation
por: Xiao, Junjin, et al.
Publicado: (2026) -
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
por: Cai, Zhejia, et al.
Publicado: (2025) -
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
por: Xiao, Junjin, et al.
Publicado: (2025)