OAT: Ordered Action Tokenization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Chaoqi, Han, Xiaoshen, Gao, Jiawei, Zhao, Yue, Chen, Haonan, Du, Yilun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-Modal Manipulation via Multi-Modal Policy Consensus
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
Localized Graph-Based Neural Dynamics Models for Terrain Manipulation
von: Liu, Chaoqi, et al.
Veröffentlicht: (2025)
von: Liu, Chaoqi, et al.
Veröffentlicht: (2025)
Grounding Video Models to Actions through Goal Conditioned Exploration
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
Autoregressive Action Sequence Learning for Robotic Manipulation
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
Solving New Tasks by Adapting Internet Video Knowledge
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
von: Nie, Buqing, et al.
Veröffentlicht: (2025)
Plan First, Diffuse Later: Extrinsic Graph Guidance for Long-Horizon Diffusion Planning
von: Hassidof, Yaniv, et al.
Veröffentlicht: (2026)
von: Hassidof, Yaniv, et al.
Veröffentlicht: (2026)
Flexible Multitask Learning with Factorized Diffusion Policy
von: Liu, Chaoqi, et al.
Veröffentlicht: (2025)
von: Liu, Chaoqi, et al.
Veröffentlicht: (2025)
Generative Trajectory Stitching through Diffusion Composition
von: Luo, Yunhao, et al.
Veröffentlicht: (2025)
von: Luo, Yunhao, et al.
Veröffentlicht: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
von: Zeng, Qixin, et al.
Veröffentlicht: (2026)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
Deconfounded Lifelong Learning for Autonomous Driving via Dynamic Knowledge Spaces
von: Du, Jiayuan, et al.
Veröffentlicht: (2026)
von: Du, Jiayuan, et al.
Veröffentlicht: (2026)
Extremum-Seeking Action Selection for Accelerating Policy Optimization
von: Chang, Ya-Chien, et al.
Veröffentlicht: (2024)
von: Chang, Ya-Chien, et al.
Veröffentlicht: (2024)
Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling
von: Liu, Yuejiang, et al.
Veröffentlicht: (2024)
von: Liu, Yuejiang, et al.
Veröffentlicht: (2024)
Compositional Generative Modeling: A Single Model is Not All You Need
von: Du, Yilun, et al.
Veröffentlicht: (2024)
von: Du, Yilun, et al.
Veröffentlicht: (2024)
Tool-as-Interface: Learning Robot Policies from Observing Human Tool Use
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
Few-Shot Task Learning through Inverse Generative Modeling
von: Netanyahu, Aviv, et al.
Veröffentlicht: (2024)
von: Netanyahu, Aviv, et al.
Veröffentlicht: (2024)
Hybrid Diffusion for Simultaneous Symbolic and Continuous Planning
von: Høeg, Sigmund Hennum, et al.
Veröffentlicht: (2025)
von: Høeg, Sigmund Hennum, et al.
Veröffentlicht: (2025)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2025)
Learning Long-Context Diffusion Policies via Past-Token Prediction
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
von: Torne, Marcel, et al.
Veröffentlicht: (2025)
Equivariant Action Sampling for Reinforcement Learning and Planning
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
von: Zhao, Linfeng, et al.
Veröffentlicht: (2024)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
von: Bu, Qingwen, et al.
Veröffentlicht: (2025)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
von: Chen, Shirui, et al.
Veröffentlicht: (2026)
Dual Action Policy for Robust Sim-to-Real Reinforcement Learning
von: Terence, Ng Wen Zheng, et al.
Veröffentlicht: (2024)
von: Terence, Ng Wen Zheng, et al.
Veröffentlicht: (2024)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
von: Yunis, David, et al.
Veröffentlicht: (2023)
von: Yunis, David, et al.
Veröffentlicht: (2023)
Behavior Generation with Latent Actions
von: Lee, Seungjae, et al.
Veröffentlicht: (2024)
von: Lee, Seungjae, et al.
Veröffentlicht: (2024)
Reinforcement Learning with Action Chunking
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
Learning Human-Like RL Agents Through Trajectory Optimization With Action Quantization
von: Guo, Jian-Ting, et al.
Veröffentlicht: (2025)
von: Guo, Jian-Ting, et al.
Veröffentlicht: (2025)
Efficient Multi-Task Reinforcement Learning via Task-Specific Action Correction
von: Feng, Jinyuan, et al.
Veröffentlicht: (2024)
von: Feng, Jinyuan, et al.
Veröffentlicht: (2024)
ExACT: An End-to-End Autonomous Excavator System Using Action Chunking With Transformers
von: Chen, Liangliang, et al.
Veröffentlicht: (2024)
von: Chen, Liangliang, et al.
Veröffentlicht: (2024)
Continuous Reasoning for Vision-Language-Action
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
Unsupervised Learning of Effective Actions in Robotics
von: Zaric, Marko, et al.
Veröffentlicht: (2024)
von: Zaric, Marko, et al.
Veröffentlicht: (2024)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
von: Yue, Yang, et al.
Veröffentlicht: (2024)
von: Yue, Yang, et al.
Veröffentlicht: (2024)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
von: Wu, Yilin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-Modal Manipulation via Multi-Modal Policy Consensus
von: Chen, Haonan, et al.
Veröffentlicht: (2025) -
Localized Graph-Based Neural Dynamics Models for Terrain Manipulation
von: Liu, Chaoqi, et al.
Veröffentlicht: (2025) -
Grounding Video Models to Actions through Goal Conditioned Exploration
von: Luo, Yunhao, et al.
Veröffentlicht: (2024) -
Autoregressive Action Sequence Learning for Robotic Manipulation
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024) -
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)