UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jianke, Hu, Yucheng, Guo, Yanjiang, Chen, Xiaoyu, Liu, Yichen, Chen, Wenna, Lu, Chaochao, Chen, Jianyu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prediction with Action: Visual Policy Learning via Joint Denoising Process
by: Guo, Yanjiang, et al.
Published: (2024)
by: Guo, Yanjiang, et al.
Published: (2024)
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
by: Hu, Yucheng, et al.
Published: (2024)
by: Hu, Yucheng, et al.
Published: (2024)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
by: Zhang, Jianke, et al.
Published: (2024)
by: Zhang, Jianke, et al.
Published: (2024)
Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation?
by: Zhang, Zhongru, et al.
Published: (2026)
by: Zhang, Zhongru, et al.
Published: (2026)
Improving Vision-Language-Action Model with Online Reinforcement Learning
by: Guo, Yanjiang, et al.
Published: (2025)
by: Guo, Yanjiang, et al.
Published: (2025)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
by: Hu, Yucheng, et al.
Published: (2026)
by: Hu, Yucheng, et al.
Published: (2026)
UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent
by: Zhang, Jianke, et al.
Published: (2025)
by: Zhang, Jianke, et al.
Published: (2025)
Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt
by: Zhu, Xiang, et al.
Published: (2025)
by: Zhu, Xiang, et al.
Published: (2025)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
HARP-VLA: Human-Robot Aligned Representation Learning for Vision-Language-Action Model
by: Zhu, Xiang, et al.
Published: (2026)
by: Zhu, Xiang, et al.
Published: (2026)
Ctrl-World: A Controllable Generative World Model for Robot Manipulation
by: Guo, Yanjiang, et al.
Published: (2025)
by: Guo, Yanjiang, et al.
Published: (2025)
dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model
by: Li, Yaxuan, et al.
Published: (2026)
by: Li, Yaxuan, et al.
Published: (2026)
VLAW: Iterative Co-Improvement of Vision-Language-Action Policy and World Model
by: Guo, Yanjiang, et al.
Published: (2026)
by: Guo, Yanjiang, et al.
Published: (2026)
UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling
by: Chen, Boyu, et al.
Published: (2026)
by: Chen, Boyu, et al.
Published: (2026)
Advancing Humanoid Locomotion: Mastering Challenging Terrains with Denoising World Model Learning
by: Gu, Xinyang, et al.
Published: (2024)
by: Gu, Xinyang, et al.
Published: (2024)
DoReMi: Grounding Language Model by Detecting and Recovering from Plan-Execution Misalignment
by: Guo, Yanjiang, et al.
Published: (2023)
by: Guo, Yanjiang, et al.
Published: (2023)
UniTacHand: Unified Spatio-Tactile Representation for Human to Robotic Hand Skill Transfer
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
UniGraspTransformer: Simplified Policy Distillation for Scalable Dexterous Robotic Grasping
by: Wang, Wenbo, et al.
Published: (2024)
by: Wang, Wenbo, et al.
Published: (2024)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
UniQuad: A Unified and Versatile Quadrotor Platform Series for UAV Research and Application
by: Zhang, Yichen, et al.
Published: (2024)
by: Zhang, Yichen, et al.
Published: (2024)
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
by: Lei, Kun, et al.
Published: (2023)
by: Lei, Kun, et al.
Published: (2023)
UniBiDex: A Unified Teleoperation Framework for Robotic Bimanual Dexterous Manipulation
by: Li, Zhongxuan, et al.
Published: (2026)
by: Li, Zhongxuan, et al.
Published: (2026)
UniDWM: Towards a Unified Driving World Model via Multifaceted Representation Learning
by: Liu, Shuai, et al.
Published: (2026)
by: Liu, Shuai, et al.
Published: (2026)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
by: Chen, Xiaoyu, et al.
Published: (2025)
by: Chen, Xiaoyu, et al.
Published: (2025)
GLUE: Global-Local Unified Encoding for Imitation Learning via Key-Patch Tracking
by: Chen, Ye, et al.
Published: (2025)
by: Chen, Ye, et al.
Published: (2025)
UniCon: A Unified System for Efficient Robot Learning Transfers
by: Lin, Yunfeng, et al.
Published: (2026)
by: Lin, Yunfeng, et al.
Published: (2026)
Efficient and Generalized end-to-end Autonomous Driving System with Latent Deep Reinforcement Learning and Demonstrations
by: Tang, Zuojin, et al.
Published: (2024)
by: Tang, Zuojin, et al.
Published: (2024)
UniForce: A Unified Latent Force Model for Robot Manipulation with Diverse Tactile Sensors
by: Chen, Zhuo, et al.
Published: (2026)
by: Chen, Zhuo, et al.
Published: (2026)
Robotic Scene Cloning:Advancing Zero-Shot Robotic Scene Adaptation in Manipulation via Visual Prompt Editing
by: Huang, Binyuan, et al.
Published: (2026)
by: Huang, Binyuan, et al.
Published: (2026)
Discovering Robotic Interaction Modes with Discrete Representation Learning
by: Wang, Liquan, et al.
Published: (2024)
by: Wang, Liquan, et al.
Published: (2024)
A Contact-Safe Reinforcement Learning Framework for Contact-Rich Robot Manipulation
by: Zhu, Xiang, et al.
Published: (2022)
by: Zhu, Xiang, et al.
Published: (2022)
UAM: A Dual-Stream Perspective on Forgetting in VLA Training
by: Zhang, Jianke, et al.
Published: (2026)
by: Zhang, Jianke, et al.
Published: (2026)
UniVTAC: A Unified Simulation Platform for Visuo-Tactile Manipulation Data Generation, Learning, and Benchmarking
by: Chen, Baijun, et al.
Published: (2026)
by: Chen, Baijun, et al.
Published: (2026)
Control Your Robot: A Unified System for Robot Control and Policy Deployment
by: Nian, Tian, et al.
Published: (2025)
by: Nian, Tian, et al.
Published: (2025)
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
UniLCD: Unified Local-Cloud Decision-Making via Reinforcement Learning
by: Sengupta, Kathakoli, et al.
Published: (2024)
by: Sengupta, Kathakoli, et al.
Published: (2024)
TurboADMM: A Structure-Exploiting Parallel Solver for Multi-Agent Trajectory Optimization
by: Chen, Yucheng
Published: (2026)
by: Chen, Yucheng
Published: (2026)
RoboUniView: Visual-Language Model with Unified View Representation for Robotic Manipulation
by: Liu, Fanfan, et al.
Published: (2024)
by: Liu, Fanfan, et al.
Published: (2024)
UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning
by: Ye, Haoming, et al.
Published: (2025)
by: Ye, Haoming, et al.
Published: (2025)
Similar Items
-
Prediction with Action: Visual Policy Learning via Joint Denoising Process
by: Guo, Yanjiang, et al.
Published: (2024) -
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations
by: Hu, Yucheng, et al.
Published: (2024) -
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
by: Zhang, Jianke, et al.
Published: (2024) -
Veo-Act: How Far Can Frontier Video Models Advance Generalizable Robot Manipulation?
by: Zhang, Zhongru, et al.
Published: (2026) -
Improving Vision-Language-Action Model with Online Reinforcement Learning
by: Guo, Yanjiang, et al.
Published: (2025)