Dense Policy: Bidirectional Autoregressive Learning of Actions
Fuente:
arXiv
Saved in:
| Main Authors: | Su, Yue, Zhan, Xinyu, Fang, Hongjie, Xue, Han, Fang, Hao-Shu, Li, Yong-Lu, Lu, Cewu, Yang, Lixin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
by: Xu, Yifu, et al.
Published: (2026)
by: Xu, Yifu, et al.
Published: (2026)
AnyDexGrasp: General Dexterous Grasping for Different Hands with Human-level Learning Efficiency
by: Fang, Hao-Shu, et al.
Published: (2025)
by: Fang, Hao-Shu, et al.
Published: (2025)
Graspness Discovery in Clutters for Fast and Accurate Grasp Detection
by: Wang, Chenxi, et al.
Published: (2024)
by: Wang, Chenxi, et al.
Published: (2024)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
by: Wang, Xinkai, et al.
Published: (2026)
by: Wang, Xinkai, et al.
Published: (2026)
Motion Before Action: Diffusing Object Motion as Manipulation Condition
by: Su, Yue, et al.
Published: (2024)
by: Su, Yue, et al.
Published: (2024)
DeformPAM: Data-Efficient Learning for Long-horizon Deformable Object Manipulation via Preference-based Action Alignment
by: Chen, Wendi, et al.
Published: (2024)
by: Chen, Wendi, et al.
Published: (2024)
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction
by: Xiong, Kai, et al.
Published: (2026)
by: Xiong, Kai, et al.
Published: (2026)
Digital Gene: Learning about the Physical World through Analytic Concepts
by: Sun, Jianhua, et al.
Published: (2025)
by: Sun, Jianhua, et al.
Published: (2025)
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
by: Wang, Yating, et al.
Published: (2025)
by: Wang, Yating, et al.
Published: (2025)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
by: Zhao, Baining, et al.
Published: (2026)
by: Zhao, Baining, et al.
Published: (2026)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
by: Zhang, Chubin, et al.
Published: (2026)
by: Zhang, Chubin, et al.
Published: (2026)
TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning
by: Liu, Jiacheng, et al.
Published: (2025)
by: Liu, Jiacheng, et al.
Published: (2025)
Kalib: Easy Hand-Eye Calibration with Reference Point Tracking
by: Tang, Tutian, et al.
Published: (2024)
by: Tang, Tutian, et al.
Published: (2024)
Physically Ground Commonsense Knowledge for Articulated Object Manipulation with Analytic Concepts
by: Wei, Jiude, et al.
Published: (2025)
by: Wei, Jiude, et al.
Published: (2025)
GarmentTracking: Category-Level Garment Pose Tracking
by: Xue, Han, et al.
Published: (2023)
by: Xue, Han, et al.
Published: (2023)
FASTer: Toward Efficient Autoregressive Vision Language Action Modeling via Neural Action Tokenization
by: Liu, Yicheng, et al.
Published: (2025)
by: Liu, Yicheng, et al.
Published: (2025)
Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy
by: Xu, Kechun, et al.
Published: (2025)
by: Xu, Kechun, et al.
Published: (2025)
Autoregressive Meta-Actions for Unified Controllable Trajectory Generation
by: Zhao, Jianbo, et al.
Published: (2025)
by: Zhao, Jianbo, et al.
Published: (2025)
MS-MANO: Enabling Hand Pose Tracking with Biomechanical Constraints
by: Xie, Pengfei, et al.
Published: (2024)
by: Xie, Pengfei, et al.
Published: (2024)
CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction
by: Gong, Zhefei, et al.
Published: (2024)
by: Gong, Zhefei, et al.
Published: (2024)
FoAR: Force-Aware Reactive Policy for Contact-Rich Robotic Manipulation
by: He, Zihao, et al.
Published: (2024)
by: He, Zihao, et al.
Published: (2024)
Demystifying Action Space Design for Robotic Manipulation Policies
by: Feng, Yuchun, et al.
Published: (2026)
by: Feng, Yuchun, et al.
Published: (2026)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
by: Ma, Jiahua, et al.
Published: (2025)
by: Ma, Jiahua, et al.
Published: (2025)
ManiPose: A Comprehensive Benchmark for Pose-aware Object Manipulation in Robotics
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models
by: Ruan, Bo-Kai, et al.
Published: (2026)
by: Ruan, Bo-Kai, et al.
Published: (2026)
Grounding Actions in Camera Space: Observation-Centric Vision-Language-Action Policy
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
ImDy: Human Inverse Dynamics from Imitated Observations
by: Liu, Xinpeng, et al.
Published: (2024)
by: Liu, Xinpeng, et al.
Published: (2024)
RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion
by: Liang, Xiujian, et al.
Published: (2025)
by: Liang, Xiujian, et al.
Published: (2025)
Discovering Conceptual Knowledge with Analytic Ontology Templates for Articulated Objects
by: Sun, Jianhua, et al.
Published: (2024)
by: Sun, Jianhua, et al.
Published: (2024)
RFTrans: Leveraging Refractive Flow of Transparent Objects for Surface Normal Estimation and Manipulation
by: Tang, Tutian, et al.
Published: (2023)
by: Tang, Tutian, et al.
Published: (2023)
Learning Part-Aware Dense 3D Feature Field for Generalizable Articulated Object Manipulation
by: Chen, Yue, et al.
Published: (2026)
by: Chen, Yue, et al.
Published: (2026)
GAMMA: Generalizable Articulation Modeling and Manipulation for Articulated Objects
by: Yu, Qiaojun, et al.
Published: (2023)
by: Yu, Qiaojun, et al.
Published: (2023)
CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation
by: Xia, Shangning, et al.
Published: (2024)
by: Xia, Shangning, et al.
Published: (2024)
RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective
by: Wang, Chenxi, et al.
Published: (2024)
by: Wang, Chenxi, et al.
Published: (2024)
Towards Effective Utilization of Mixed-Quality Demonstrations in Robotic Manipulation via Segment-Level Selection and Optimization
by: Chen, Jingjing, et al.
Published: (2024)
by: Chen, Jingjing, et al.
Published: (2024)
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
by: Liu, Jiaming, et al.
Published: (2025)
by: Liu, Jiaming, et al.
Published: (2025)
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
by: Xue, Han, et al.
Published: (2026)
by: Xue, Han, et al.
Published: (2026)
DynVLA: Learning World Dynamics for Action Reasoning in Autonomous Driving
by: Shang, Shuyao, et al.
Published: (2026)
by: Shang, Shuyao, et al.
Published: (2026)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
by: Jin, Yang, et al.
Published: (2024)
by: Jin, Yang, et al.
Published: (2024)
OccupancyDETR: Using DETR for Mixed Dense-sparse 3D Occupancy Prediction
by: Jia, Yupeng, et al.
Published: (2023)
by: Jia, Yupeng, et al.
Published: (2023)
Similar Items
-
LIDEA: Human-to-Robot Imitation Learning via Implicit Feature Distillation and Explicit Geometry Alignment
by: Xu, Yifu, et al.
Published: (2026) -
AnyDexGrasp: General Dexterous Grasping for Different Hands with Human-level Learning Efficiency
by: Fang, Hao-Shu, et al.
Published: (2025) -
Graspness Discovery in Clutters for Fast and Accurate Grasp Detection
by: Wang, Chenxi, et al.
Published: (2024) -
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
by: Wang, Xinkai, et al.
Published: (2026) -
Motion Before Action: Diffusing Object Motion as Manipulation Condition
by: Su, Yue, et al.
Published: (2024)