A Dual Process VLA: Efficient Robotic Manipulation Leveraging VLM
Fuente:
arXiv
Guardado en:
| Autores principales: | Han, ByungOk, Kim, Jaehong, Jang, Jinhyeok |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Space-Aware Instruction Tuning: Dataset and Benchmark for Guide Dog Robots Assisting the Visually Impaired
por: Han, ByungOk, et al.
Publicado: (2025)
por: Han, ByungOk, et al.
Publicado: (2025)
SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation
por: Li, Wei, et al.
Publicado: (2025)
por: Li, Wei, et al.
Publicado: (2025)
OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation
por: Cui, Can, et al.
Publicado: (2025)
por: Cui, Can, et al.
Publicado: (2025)
GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
por: Qian, Jingjing, et al.
Publicado: (2025)
por: Qian, Jingjing, et al.
Publicado: (2025)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
por: Guo, Heyu, et al.
Publicado: (2025)
por: Guo, Heyu, et al.
Publicado: (2025)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
por: Wen, Junjie, et al.
Publicado: (2024)
por: Wen, Junjie, et al.
Publicado: (2024)
ST-$π$: Structured SpatioTemporal VLA for Robotic Manipulation
por: Ma, Chuanhao, et al.
Publicado: (2026)
por: Ma, Chuanhao, et al.
Publicado: (2026)
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
por: Jiang, Yuming, et al.
Publicado: (2025)
por: Jiang, Yuming, et al.
Publicado: (2025)
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
por: Hanyu, Taisei, et al.
Publicado: (2025)
por: Hanyu, Taisei, et al.
Publicado: (2025)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
por: Xu, Ran, et al.
Publicado: (2024)
por: Xu, Ran, et al.
Publicado: (2024)
EveryDayVLA: A Vision-Language-Action Model for Affordable Robotic Manipulation
por: Chopra, Samarth, et al.
Publicado: (2025)
por: Chopra, Samarth, et al.
Publicado: (2025)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
por: Wang, Hongyu, et al.
Publicado: (2025)
por: Wang, Hongyu, et al.
Publicado: (2025)
Compressor-VLA: Instruction-Guided Visual Token Compression for Efficient Robotic Manipulation
por: Gao, Juntao, et al.
Publicado: (2025)
por: Gao, Juntao, et al.
Publicado: (2025)
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
por: Shao, Rui, et al.
Publicado: (2025)
por: Shao, Rui, et al.
Publicado: (2025)
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
por: Shi, Hao, et al.
Publicado: (2025)
por: Shi, Hao, et al.
Publicado: (2025)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
por: Li, Runhao, et al.
Publicado: (2025)
por: Li, Runhao, et al.
Publicado: (2025)
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching
por: Chen, Jiayi, et al.
Publicado: (2026)
por: Chen, Jiayi, et al.
Publicado: (2026)
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
por: Ye, Guo, et al.
Publicado: (2025)
por: Ye, Guo, et al.
Publicado: (2025)
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
por: Shen, Yichao, et al.
Publicado: (2025)
por: Shen, Yichao, et al.
Publicado: (2025)
Learning from Oblivion: Predicting Knowledge Overflowed Weights via Retrodiction of Forgetting
por: Jang, Jinhyeok, et al.
Publicado: (2025)
por: Jang, Jinhyeok, et al.
Publicado: (2025)
VERM: Leveraging Foundation Models to Create a Virtual Eye for Efficient 3D Robotic Manipulation
por: Chen, Yixiang, et al.
Publicado: (2025)
por: Chen, Yixiang, et al.
Publicado: (2025)
Think Proprioceptively: Embodied Visual Reasoning for VLA Manipulation
por: Wang, Fangyuan, et al.
Publicado: (2026)
por: Wang, Fangyuan, et al.
Publicado: (2026)
ForceVLA: Enhancing VLA Models with a Force-aware MoE for Contact-rich Manipulation
por: Yu, Jiawen, et al.
Publicado: (2025)
por: Yu, Jiawen, et al.
Publicado: (2025)
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
por: Yang, Yandan, et al.
Publicado: (2026)
por: Yang, Yandan, et al.
Publicado: (2026)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
por: Ding, Pengxiang, et al.
Publicado: (2023)
por: Ding, Pengxiang, et al.
Publicado: (2023)
CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling
por: Li, Hao, et al.
Publicado: (2025)
por: Li, Hao, et al.
Publicado: (2025)
DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
por: Xie, Haozhe, et al.
Publicado: (2026)
por: Xie, Haozhe, et al.
Publicado: (2026)
TriRelVLA: Triadic Relational Structure for Generalizable Embodied Manipulation
por: Zhou, Hanyu, et al.
Publicado: (2026)
por: Zhou, Hanyu, et al.
Publicado: (2026)
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
por: Song, Wenxuan, et al.
Publicado: (2025)
por: Song, Wenxuan, et al.
Publicado: (2025)
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation
por: Lian, Shijie, et al.
Publicado: (2026)
por: Lian, Shijie, et al.
Publicado: (2026)
A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
por: Patel, Shivansh, et al.
Publicado: (2025)
por: Patel, Shivansh, et al.
Publicado: (2025)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
por: Yang, Shuai, et al.
Publicado: (2025)
por: Yang, Shuai, et al.
Publicado: (2025)
ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration
por: Zhu, Minjie, et al.
Publicado: (2025)
por: Zhu, Minjie, et al.
Publicado: (2025)
Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation
por: Wu, Yiming, et al.
Publicado: (2026)
por: Wu, Yiming, et al.
Publicado: (2026)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
por: Wu, Zhenyu, et al.
Publicado: (2025)
por: Wu, Zhenyu, et al.
Publicado: (2025)
FD-VLA: Force-Distilled Vision-Language-Action Model for Contact-Rich Manipulation
por: Zhao, Ruiteng, et al.
Publicado: (2026)
por: Zhao, Ruiteng, et al.
Publicado: (2026)
Robust Instant Policy: Leveraging Student's t-Regression Model for Robust In-context Imitation Learning of Robot Manipulation
por: Oh, Hanbit, et al.
Publicado: (2025)
por: Oh, Hanbit, et al.
Publicado: (2025)
VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
por: Xu, Siyu, et al.
Publicado: (2025)
por: Xu, Siyu, et al.
Publicado: (2025)
ReVLA: Reverting Visual Domain Limitation of Robotic Foundation Models
por: Dey, Sombit, et al.
Publicado: (2024)
por: Dey, Sombit, et al.
Publicado: (2024)
Robotic VLA Benefits from Joint Learning with Motion Image Diffusion
por: Fang, Yu, et al.
Publicado: (2025)
por: Fang, Yu, et al.
Publicado: (2025)
Ejemplares similares
-
Space-Aware Instruction Tuning: Dataset and Benchmark for Guide Dog Robots Assisting the Visually Impaired
por: Han, ByungOk, et al.
Publicado: (2025) -
SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation
por: Li, Wei, et al.
Publicado: (2025) -
OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation
por: Cui, Can, et al.
Publicado: (2025) -
GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation
por: Qian, Jingjing, et al.
Publicado: (2025) -
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
por: Guo, Heyu, et al.
Publicado: (2025)