GigaBrain-0.5M*: a VLA That Learns From World Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | GigaBrain Team, Wang, Boyuan, Li, Bohan, Ni, Chaojun, Huang, Guan, Zhao, Guosheng, Li, Hao, Li, Jie, Lv, Jindi, Liu, Jingyu, Feng, Lv, Yu, Mingming, Li, Peng, Deng, Qiuping, Liu, Tianze, Zhou, Xinyu, Chen, Xinze, Wang, Xiaofeng, Wang, Yang, Li, Yifan, Nie, Yifei, Li, Yilong, Zhou, Yukun, Ye, Yun, Liu, Zhichao, Zhu, Zheng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
by: GigaBrain Team, et al.
Published: (2025)
by: GigaBrain Team, et al.
Published: (2025)
GigaWorld-Policy: An Efficient Action-Centered World--Action Model
by: Ye, Angen, et al.
Published: (2026)
by: Ye, Angen, et al.
Published: (2026)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
by: GigaWorld Team, et al.
Published: (2025)
by: GigaWorld Team, et al.
Published: (2025)
GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning
by: Bao, Xiaoyi, et al.
Published: (2025)
by: Bao, Xiaoyi, et al.
Published: (2025)
ViVa: A Video-Generative Value Model for Robot Reinforcement Learning
by: Lv, Jindi, et al.
Published: (2026)
by: Lv, Jindi, et al.
Published: (2026)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
by: Wang, Boyuan, et al.
Published: (2025)
by: Wang, Boyuan, et al.
Published: (2025)
Multi-Path Collaborative Reasoning via Reinforcement Learning
by: Lv, Jindi, et al.
Published: (2025)
by: Lv, Jindi, et al.
Published: (2025)
SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead
by: Ni, Chaojun, et al.
Published: (2025)
by: Ni, Chaojun, et al.
Published: (2025)
Ferret: An Efficient Online Continual Learning Framework under Varying Memory Constraints
by: Zhou, Yuhao, et al.
Published: (2025)
by: Zhou, Yuhao, et al.
Published: (2025)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
by: Ni, Chaojun, et al.
Published: (2025)
by: Ni, Chaojun, et al.
Published: (2025)
GPS: Distilling Compact Memories via Grid-based Patch Sampling for Efficient Online Class-Incremental Learning
by: Ma, Mingchuan, et al.
Published: (2025)
by: Ma, Mingchuan, et al.
Published: (2025)
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis
by: Lang, Xiaolei, et al.
Published: (2026)
by: Lang, Xiaolei, et al.
Published: (2026)
Local Embeddedness and Growth of Non‐local Firms
by: Zhongda Li, et al.
Published: (2025)
by: Zhongda Li, et al.
Published: (2025)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
by: Li, Haoyun, et al.
Published: (2025)
by: Li, Haoyun, et al.
Published: (2025)
Say More with Less: Understanding Prompt Learning Behaviors through Gist Compression
by: Li, Xinze, et al.
Published: (2024)
by: Li, Xinze, et al.
Published: (2024)
Deploying Models to Non-participating Clients in Federated Learning without Fine-tuning: A Hypernetwork-based Approach
by: Zhou, Yuhao, et al.
Published: (2025)
by: Zhou, Yuhao, et al.
Published: (2025)
The Impact of 2D and 3D Gamified VR on Learning American Sign Language
by: Wang, Jindi, et al.
Published: (2024)
by: Wang, Jindi, et al.
Published: (2024)
Gas‐Phase F‐Atom Migration Reactions of Perfluoroalkyl and Polyfluoroalkyl Sulfonic/Sulfinic Anions
by: Qingqin Liu, et al.
Published: (2026)
by: Qingqin Liu, et al.
Published: (2026)
Personalized News Recommendation with Multi-granularity Candidate-aware User Modeling
by: Li, Qiang, et al.
Published: (2025)
by: Li, Qiang, et al.
Published: (2025)
ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations
by: Zhou, Yuhao, et al.
Published: (2026)
by: Zhou, Yuhao, et al.
Published: (2026)
Melatonin Inhibits CD4+T Cell Apoptosis via the Bcl‐2/BAX Pathway and Improves Survival Rates in Mice With Sepsis
by: Zhenggong Li, et al.
Published: (2025)
by: Zhenggong Li, et al.
Published: (2025)
Entire Space Counterfactual Learning for Reliable Content Recommendations
by: Wang, Hao, et al.
Published: (2022)
by: Wang, Hao, et al.
Published: (2022)
Astrolabe: A Content-Addressable Hypergraph for Semantic Knowledge Management
by: Li, Xinze
Published: (2026)
by: Li, Xinze
Published: (2026)
Lecture Notes on Comparison Geometry
by: Li, Xinze
Published: (2024)
by: Li, Xinze
Published: (2024)
Incorporating Inductive Biases to Energy-based Generative Models
by: Li, Yukun, et al.
Published: (2025)
by: Li, Yukun, et al.
Published: (2025)
Underwater Image Enhancement with Cascaded Contrastive Learning
by: Liu, Yi, et al.
Published: (2024)
by: Liu, Yi, et al.
Published: (2024)
DualPG‐DTA: A Large Language Model‐Powered Graph Neural Network Framework for Enhanced Drug‐Target Affinity Prediction and Discovery of Novel CDK9 Inhibitors Exhibiting In Vivo Anti‐Leukemia Activity
by: Yihao Chen, et al.
Published: (2026)
by: Yihao Chen, et al.
Published: (2026)
On the Reduction of the Spherical Point-in-Polygon Problem for Antipode-Excluding Spherical Polygons
by: Li, Ziqiang, et al.
Published: (2023)
by: Li, Ziqiang, et al.
Published: (2023)
Similarity signature curves for forming periodic orbits in the Lorenz system
by: Li, Jindi, et al.
Published: (2022)
by: Li, Jindi, et al.
Published: (2022)
ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction
by: Ni, Chaojun, et al.
Published: (2025)
by: Ni, Chaojun, et al.
Published: (2025)
CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human
by: Sun, Nan, et al.
Published: (2025)
by: Sun, Nan, et al.
Published: (2025)
RoleMAG: Learning Neighbor Roles in Multimodal Graphs
by: Zuo, Yilong, et al.
Published: (2026)
by: Zuo, Yilong, et al.
Published: (2026)
Electrodeposited Vanadium Pentoxide/Polypyrrole as a Highly Efficient and Low Cost Electrode Material for Microsupercapacitor Applications
by: Zilong Wang, et al.
Published: (2026)
by: Zilong Wang, et al.
Published: (2026)
Self‐Powered Lower‐Limb Motion Detection with a Piezo‐Electromagnetic Generator
by: Pingchang Wang, et al.
Published: (2025)
by: Pingchang Wang, et al.
Published: (2025)
DriveDreamer4D: World Models Are Effective Data Machines for 4D Driving Scene Representation
by: Zhao, Guosheng, et al.
Published: (2024)
by: Zhao, Guosheng, et al.
Published: (2024)
Graph-based Confidence Calibration for Large Language Models
by: Li, Yukun, et al.
Published: (2024)
by: Li, Yukun, et al.
Published: (2024)
HyperNAS: Enhancing Architecture Representation for NAS Predictor via Hypernetwork
by: Lv, Jindi, et al.
Published: (2025)
by: Lv, Jindi, et al.
Published: (2025)
DarkShot: Lighting Dark Images with Low-Compute and High-Quality
by: Zheng, Jiazhang, et al.
Published: (2023)
by: Zheng, Jiazhang, et al.
Published: (2023)
Enhancing Diffusion-based Point Cloud Generation with Smoothness Constraint
by: Li, Yukun, et al.
Published: (2024)
by: Li, Yukun, et al.
Published: (2024)
The LLN and CLT for the statistical ensembles of discrete integrable Hamiltonian systems
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Similar Items
-
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
by: GigaBrain Team, et al.
Published: (2025) -
GigaWorld-Policy: An Efficient Action-Centered World--Action Model
by: Ye, Angen, et al.
Published: (2026) -
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
by: GigaWorld Team, et al.
Published: (2025) -
GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning
by: Bao, Xiaoyi, et al.
Published: (2025) -
ViVa: A Video-Generative Value Model for Robot Reinforcement Learning
by: Lv, Jindi, et al.
Published: (2026)