GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Cheang, Chi-Lam, Chen, Guangzeng, Jing, Ya, Kong, Tao, Li, Hang, Li, Yifeng, Liu, Yuxiao, Wu, Hongtao, Xu, Jiafeng, Yang, Yichu, Zhang, Hanbo, Zhu, Minzhao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GR-3 Technical Report
by: Cheang, Chilam, et al.
Published: (2025)
by: Cheang, Chilam, et al.
Published: (2025)
IRASim: A Fine-Grained World Model for Robot Manipulation
by: Zhu, Fangqi, et al.
Published: (2024)
by: Zhu, Fangqi, et al.
Published: (2024)
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
by: Li, Peiyan, et al.
Published: (2024)
by: Li, Peiyan, et al.
Published: (2024)
Vision-Language Foundation Models as Effective Robot Imitators
by: Li, Xinghang, et al.
Published: (2023)
by: Li, Xinghang, et al.
Published: (2023)
Human-assisted Robotic Policy Refinement via Action Preference Optimization
by: Xia, Wenke, et al.
Published: (2025)
by: Xia, Wenke, et al.
Published: (2025)
GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation
by: Li, Yunfei, et al.
Published: (2025)
by: Li, Yunfei, et al.
Published: (2025)
Chain-of-Action: Trajectory Autoregressive Modeling for Robotic Manipulation
by: Zhang, Wenbo, et al.
Published: (2025)
by: Zhang, Wenbo, et al.
Published: (2025)
Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots
by: Liu, Minghuan, et al.
Published: (2025)
by: Liu, Minghuan, et al.
Published: (2025)
GR-Dexter Technical Report
by: Wen, Ruoshi, et al.
Published: (2025)
by: Wen, Ruoshi, et al.
Published: (2025)
Towards Unified Interactive Visual Grounding in The Wild
by: Xu, Jie, et al.
Published: (2024)
by: Xu, Jie, et al.
Published: (2024)
A Controllable‐Stiffness Tensegrity Robot Joint for Robust Compliant Manipulation
by: Yifeng Hao, et al.
Published: (2025)
by: Yifeng Hao, et al.
Published: (2025)
SInViG: A Self-Evolving Interactive Visual Agent for Human-Robot Interaction
by: Xu, Jie, et al.
Published: (2024)
by: Xu, Jie, et al.
Published: (2024)
Efficient Sensorimotor Learning for Open-world Robot Manipulation
by: Zhu, Yifeng
Published: (2025)
by: Zhu, Yifeng
Published: (2025)
What Matters in Building Vision-Language-Action Models for Generalist Robots
by: Li, Xinghang, et al.
Published: (2024)
by: Li, Xinghang, et al.
Published: (2024)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
by: Wang, Guokang, et al.
Published: (2024)
by: Wang, Guokang, et al.
Published: (2024)
On bounded energy of convolution of fractal measures
by: Yi, Guangzeng
Published: (2024)
by: Yi, Guangzeng
Published: (2024)
OKAMI: Teaching Humanoid Robots Manipulation Skills through Single Video Imitation
by: Li, Jinhan, et al.
Published: (2024)
by: Li, Jinhan, et al.
Published: (2024)
World Model-based Perception for Visual Legged Locomotion
by: Lai, Hang, et al.
Published: (2024)
by: Lai, Hang, et al.
Published: (2024)
Knowledge-driven Augmentation and Retrieval for Integrative Temporal Adaptation
by: Liu, Weisi, et al.
Published: (2026)
by: Liu, Weisi, et al.
Published: (2026)
DGLA Actions: An Application in GR
by: Grady, Ryan
Published: (2025)
by: Grady, Ryan
Published: (2025)
Learning to Localize Actions in Instructional Videos with LLM-Based Multi-Pathway Text-Video Alignment
by: Chen, Yuxiao, et al.
Published: (2024)
by: Chen, Yuxiao, et al.
Published: (2024)
ZeroMimic: Distilling Robotic Manipulation Skills from Web Videos
by: Shi, Junyao, et al.
Published: (2025)
by: Shi, Junyao, et al.
Published: (2025)
Characterizing User Platforms for Video Streaming in Broadband Networks
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
by: Jiang, Guangqi, et al.
Published: (2024)
by: Jiang, Guangqi, et al.
Published: (2024)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulness
by: Cheang, Chi Seng, et al.
Published: (2025)
by: Cheang, Chi Seng, et al.
Published: (2025)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
by: Li, Yajie, et al.
Published: (2026)
by: Li, Yajie, et al.
Published: (2026)
$τ_0$-WM: A Unified Video-Action World Model for Robotic Manipulation
by: Zhou, Pengfei, et al.
Published: (2026)
by: Zhou, Pengfei, et al.
Published: (2026)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
by: Miao, Cui, et al.
Published: (2025)
by: Miao, Cui, et al.
Published: (2025)
Demystifying Action Space Design for Robotic Manipulation Policies
by: Feng, Yuchun, et al.
Published: (2026)
by: Feng, Yuchun, et al.
Published: (2026)
EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow
by: Chen, Yixiang, et al.
Published: (2025)
by: Chen, Yixiang, et al.
Published: (2025)
Time-Unified Diffusion Policy with Action Discrimination for Robotic Manipulation
by: Niu, Ye, et al.
Published: (2025)
by: Niu, Ye, et al.
Published: (2025)
Large cliques in extremal incidence configurations
by: Orponen, Tuomas, et al.
Published: (2024)
by: Orponen, Tuomas, et al.
Published: (2024)
Blow up analysis for a parabolic MEMS problem, I: Hölder estimate
by: Wang, Kelei, et al.
Published: (2024)
by: Wang, Kelei, et al.
Published: (2024)
What Makes Good Instruction-Tuning Data? An In-Context Learning Perspective
by: Han, Guangzeng, et al.
Published: (2026)
by: Han, Guangzeng, et al.
Published: (2026)
Scaling World Model for Hierarchical Manipulation Policies
by: Long, Qian, et al.
Published: (2026)
by: Long, Qian, et al.
Published: (2026)
STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation
by: Tian, Yuxuan, et al.
Published: (2026)
by: Tian, Yuxuan, et al.
Published: (2026)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
ConLA: Contrastive Latent Action Learning from Human Videos for Robotic Manipulation
by: Dai, Weisheng, et al.
Published: (2026)
by: Dai, Weisheng, et al.
Published: (2026)
Similar Items
-
GR-3 Technical Report
by: Cheang, Chilam, et al.
Published: (2025) -
IRASim: A Fine-Grained World Model for Robot Manipulation
by: Zhu, Fangqi, et al.
Published: (2024) -
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
by: Li, Peiyan, et al.
Published: (2024) -
Vision-Language Foundation Models as Effective Robot Imitators
by: Li, Xinghang, et al.
Published: (2023) -
Human-assisted Robotic Policy Refinement via Action Preference Optimization
by: Xia, Wenke, et al.
Published: (2025)