Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Junjin, Li, Dongyang, Yang, Yandan, Zeng, Shuang, Lin, Tong, Chang, Xinyuan, Xiong, Feng, Xu, Mu, Wei, Xing, Ma, Zhiheng, Zhang, Qing, Zheng, Wei-Shi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
by: Yang, Yandan, et al.
Published: (2026)
by: Yang, Yandan, et al.
Published: (2026)
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
by: Chen, Yuzhi, et al.
Published: (2026)
by: Chen, Yuzhi, et al.
Published: (2026)
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
by: Xiao, Junjin, et al.
Published: (2025)
by: Xiao, Junjin, et al.
Published: (2025)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
by: Cai, Zhejia, et al.
Published: (2025)
by: Cai, Zhejia, et al.
Published: (2025)
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models
by: Tang, Zuojin, et al.
Published: (2026)
by: Tang, Zuojin, et al.
Published: (2026)
PriorDrive: Enhancing Online HD Mapping with Unified Vector Priors
by: Zeng, Shuang, et al.
Published: (2024)
by: Zeng, Shuang, et al.
Published: (2024)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
by: Huo, Dongjie, et al.
Published: (2026)
by: Huo, Dongjie, et al.
Published: (2026)
NECA: Neural Customizable Human Avatar
by: Xiao, Junjin, et al.
Published: (2024)
by: Xiao, Junjin, et al.
Published: (2024)
SeqGrowGraph: Learning Lane Topology as a Chain of Graph Expansions
by: Xie, Mengwei, et al.
Published: (2025)
by: Xie, Mengwei, et al.
Published: (2025)
RoGSplat: Learning Robust Generalizable Human Gaussian Splatting from Sparse Multi-View Images
by: Xiao, Junjin, et al.
Published: (2025)
by: Xiao, Junjin, et al.
Published: (2025)
Learning Explicit Continuous Motion Representation for Dynamic Gaussian Splatting from Monocular Videos
by: Zhang, Xuankai, et al.
Published: (2026)
by: Zhang, Xuankai, et al.
Published: (2026)
JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation
by: Zeng, Shuang, et al.
Published: (2025)
by: Zeng, Shuang, et al.
Published: (2025)
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models
by: Liu, Haoyun, et al.
Published: (2026)
by: Liu, Haoyun, et al.
Published: (2026)
Dynamic Gaussian Splatting from Defocused and Motion-blurred Monocular Videos
by: Zhang, Xuankai, et al.
Published: (2025)
by: Zhang, Xuankai, et al.
Published: (2025)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
by: Qi, Dekang, et al.
Published: (2026)
by: Qi, Dekang, et al.
Published: (2026)
iManip: Skill-Incremental Learning for Robotic Manipulation
by: Zheng, Zexin, et al.
Published: (2025)
by: Zheng, Zexin, et al.
Published: (2025)
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
by: Zeng, Shuang, et al.
Published: (2025)
by: Zeng, Shuang, et al.
Published: (2025)
Multimodal Action Quality Assessment
by: Zeng, Ling-An, et al.
Published: (2024)
by: Zeng, Ling-An, et al.
Published: (2024)
ConLA: Contrastive Latent Action Learning from Human Videos for Robotic Manipulation
by: Dai, Weisheng, et al.
Published: (2026)
by: Dai, Weisheng, et al.
Published: (2026)
Rethinking Bimanual Robotic Manipulation: Learning with Decoupled Interaction Framework
by: Jiang, Jian-Jian, et al.
Published: (2025)
by: Jiang, Jian-Jian, et al.
Published: (2025)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
by: Li, Yajie, et al.
Published: (2026)
by: Li, Yajie, et al.
Published: (2026)
Action-Prior Denoising for Smooth Real-Time Chunking
by: Liu, Dongyang, et al.
Published: (2026)
by: Liu, Dongyang, et al.
Published: (2026)
What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion
by: Yue, Zhengrong, et al.
Published: (2026)
by: Yue, Zhengrong, et al.
Published: (2026)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
by: Dai, Tingjun, et al.
Published: (2026)
by: Dai, Tingjun, et al.
Published: (2026)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
Demystifying Action Space Design for Robotic Manipulation Policies
by: Feng, Yuchun, et al.
Published: (2026)
by: Feng, Yuchun, et al.
Published: (2026)
Autoregressive Action Sequence Learning for Robotic Manipulation
by: Zhang, Xinyu, et al.
Published: (2024)
by: Zhang, Xinyu, et al.
Published: (2024)
LoLA: Long Horizon Latent Action Learning for General Robot Manipulation
by: Wang, Xiaofan, et al.
Published: (2025)
by: Wang, Xiaofan, et al.
Published: (2025)
Parallel compressive super-resolution imaging with wide field-of-view based on physics enhanced network
by: Jin, Xiao-Peng, et al.
Published: (2023)
by: Jin, Xiao-Peng, et al.
Published: (2023)
Modeling Elastic-Body Dynamics of Robotic Fish Using a Variational Framework
by: Chen, Zhiheng, et al.
Published: (2025)
by: Chen, Zhiheng, et al.
Published: (2025)
DiffPoint: Single and Multi-view Point Cloud Reconstruction with ViT Based Diffusion Model
by: Feng, Yu, et al.
Published: (2024)
by: Feng, Yu, et al.
Published: (2024)
Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic Manipulation
by: Mu, Tong, et al.
Published: (2025)
by: Mu, Tong, et al.
Published: (2025)
Driving by the Rules: A Benchmark for Integrating Traffic Sign Regulations into Vectorized HD Map
by: Chang, Xinyuan, et al.
Published: (2024)
by: Chang, Xinyuan, et al.
Published: (2024)
Action-Geometry Prediction with 3D Geometric Prior for Bimanual Manipulation
by: Xu, Chongyang, et al.
Published: (2026)
by: Xu, Chongyang, et al.
Published: (2026)
FACTO: Function-space Adaptive Constrained Trajectory Optimization for Robotic Manipulators
by: Feng, Yichang, et al.
Published: (2026)
by: Feng, Yichang, et al.
Published: (2026)
Latent Conditioned Loco-Manipulation Using Motion Priors
by: Stępień, Maciej, et al.
Published: (2025)
by: Stępień, Maciej, et al.
Published: (2025)
One-Policy-Fits-All: Geometry-Aware Action Latents for Cross-Embodiment Manipulation
by: Mu, Juncheng, et al.
Published: (2026)
by: Mu, Juncheng, et al.
Published: (2026)
TypeTele: Releasing Dexterity in Teleoperation by Dexterous Manipulation Types
by: Lin, Yuhao, et al.
Published: (2025)
by: Lin, Yuhao, et al.
Published: (2025)
Hyperbolic Multiview Pretraining for Robotic Manipulation
by: Yang, Jin, et al.
Published: (2026)
by: Yang, Jin, et al.
Published: (2026)
Latent Action Diffusion for Cross-Embodiment Manipulation
by: Bauer, Erik, et al.
Published: (2025)
by: Bauer, Erik, et al.
Published: (2025)
Similar Items
-
ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning
by: Yang, Yandan, et al.
Published: (2026) -
ABot-PhysWorld: Interactive World Foundation Model for Robotic Manipulation with Physics Alignment
by: Chen, Yuzhi, et al.
Published: (2026) -
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
by: Xiao, Junjin, et al.
Published: (2025) -
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
by: Cai, Zhejia, et al.
Published: (2025) -
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models
by: Tang, Zuojin, et al.
Published: (2026)