CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Zhefei, Ding, Pengxiang, Lyu, Shangke, Huang, Siteng, Sun, Mingyang, Zhao, Wei, Fan, Zhaoxin, Wang, Donglin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Score and Distribution Matching Policy: Advanced Accelerated Visuomotor Policies via Matched Distillation
by: Jia, Bofang, et al.
Published: (2024)
by: Jia, Bofang, et al.
Published: (2024)
Robust Online Residual Refinement via Koopman-Guided Dynamics Modeling
by: Gong, Zhefei, et al.
Published: (2025)
by: Gong, Zhefei, et al.
Published: (2025)
Learning Robotic Policy with Imagined Transition: Mitigating the Trade-off between Robustness and Optimality
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Discovering Self-Protective Falling Policy for Humanoid Robot via Deep Reinforcement Learning
by: Shi, Diyuan, et al.
Published: (2025)
by: Shi, Diyuan, et al.
Published: (2025)
Dynamic Adaptive Legged Locomotion Policy via Decoupling Reaction Force Control and Gait Control
by: Wang, Renjie, et al.
Published: (2025)
by: Wang, Renjie, et al.
Published: (2025)
QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning
by: Tong, Xinyang, et al.
Published: (2024)
by: Tong, Xinyang, et al.
Published: (2024)
Unveiling the Potential of Vision-Language-Action Models with Open-Ended Multimodal Instructions
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
GEVRM: Goal-Expressive Video Generation Model For Robust Visual Manipulation
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
by: Lin, Minghui, et al.
Published: (2025)
by: Lin, Minghui, et al.
Published: (2025)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
GeRM: A Generalist Robotic Model with Mixture-of-experts for Quadruped Robot
by: Song, Wenxuan, et al.
Published: (2024)
by: Song, Wenxuan, et al.
Published: (2024)
CUBic: Coordinated Unified Bimanual Perception and Control Framework
by: Wang, Xingyu, et al.
Published: (2026)
by: Wang, Xingyu, et al.
Published: (2026)
Integrating Trajectory Optimization and Reinforcement Learning for Quadrupedal Jumping with Terrain-Adaptive Landing
by: Wang, Renjie, et al.
Published: (2025)
by: Wang, Renjie, et al.
Published: (2025)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
by: Fan, Yiguo, et al.
Published: (2025)
by: Fan, Yiguo, et al.
Published: (2025)
VAMPO: Policy Optimization for Improving Visual Dynamics in Video Action Models
by: Ge, Zirui, et al.
Published: (2026)
by: Ge, Zirui, et al.
Published: (2026)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
by: Ma, Jiahua, et al.
Published: (2025)
by: Ma, Jiahua, et al.
Published: (2025)
FreqPolicy: Frequency Autoregressive Visuomotor Policy with Continuous Tokens
by: Zhong, Yiming, et al.
Published: (2025)
by: Zhong, Yiming, et al.
Published: (2025)
QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
by: Ding, Pengxiang, et al.
Published: (2023)
by: Ding, Pengxiang, et al.
Published: (2023)
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
by: Chi, Cheng, et al.
Published: (2023)
by: Chi, Cheng, et al.
Published: (2023)
Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution
by: Sun, Zhanyi, et al.
Published: (2025)
by: Sun, Zhanyi, et al.
Published: (2025)
CMR: Contractive Mapping Embeddings for Robust Humanoid Locomotion on Unstructured Terrains
by: Zeng, Qixin, et al.
Published: (2026)
by: Zeng, Qixin, et al.
Published: (2026)
3D Equivariant Visuomotor Policy Learning via Spherical Projection
by: Hu, Boce, et al.
Published: (2025)
by: Hu, Boce, et al.
Published: (2025)
CEI: A Unified Interface for Cross-Embodiment Visuomotor Policy Learning in 3D Space
by: Wu, Tong, et al.
Published: (2026)
by: Wu, Tong, et al.
Published: (2026)
Learning Visuomotor Policy for Multi-Robot Laser Tag Game
by: Li, Kai, et al.
Published: (2026)
by: Li, Kai, et al.
Published: (2026)
History-Aware Visuomotor Policy Learning via Point Tracking
by: Chen, Jingjing, et al.
Published: (2025)
by: Chen, Jingjing, et al.
Published: (2025)
Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
Imagination at Inference: Synthesizing In-Hand Views for Robust Visuomotor Policy Inference
by: Ding, Haoran, et al.
Published: (2025)
by: Ding, Haoran, et al.
Published: (2025)
ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
by: Zhao, Han, et al.
Published: (2025)
by: Zhao, Han, et al.
Published: (2025)
MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
L1 Sample Flow for Efficient Visuomotor Learning
by: Song, Weixi, et al.
Published: (2025)
by: Song, Weixi, et al.
Published: (2025)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
by: Wang, Zhendong, et al.
Published: (2024)
by: Wang, Zhendong, et al.
Published: (2024)
STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction
by: Li, Jinhao, et al.
Published: (2026)
by: Li, Jinhao, et al.
Published: (2026)
KineDex: Learning Tactile-Informed Visuomotor Policies via Kinesthetic Teaching for Dexterous Manipulation
by: Zhang, Di, et al.
Published: (2025)
by: Zhang, Di, et al.
Published: (2025)
Falcon: Fast Visuomotor Policies via Partial Denoising
by: Chen, Haojun, et al.
Published: (2025)
by: Chen, Haojun, et al.
Published: (2025)
Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
by: Kim, Moo Jin, et al.
Published: (2026)
by: Kim, Moo Jin, et al.
Published: (2026)
Unlock Reliable Skill Inference for Quadruped Adaptive Behavior by Skill Graph
by: Zhang, Hongyin, et al.
Published: (2023)
by: Zhang, Hongyin, et al.
Published: (2023)
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
by: Lyu, Mingyang, et al.
Published: (2025)
by: Lyu, Mingyang, et al.
Published: (2025)
Similar Items
-
Score and Distribution Matching Policy: Advanced Accelerated Visuomotor Policies via Matched Distillation
by: Jia, Bofang, et al.
Published: (2024) -
Robust Online Residual Refinement via Koopman-Guided Dynamics Modeling
by: Gong, Zhefei, et al.
Published: (2025) -
Learning Robotic Policy with Imagined Transition: Mitigating the Trade-off between Robustness and Optimality
by: Xiao, Wei, et al.
Published: (2025) -
Discovering Self-Protective Falling Policy for Humanoid Robot via Deep Reinforcement Learning
by: Shi, Diyuan, et al.
Published: (2025) -
Dynamic Adaptive Legged Locomotion Policy via Decoupling Reaction Force Control and Gait Control
by: Wang, Renjie, et al.
Published: (2025)