Grounding Bodily Awareness in Visual Representations for Efficient Policy Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Junlin, Lin, Zhiyun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
by: Tsagkas, Nikolaos, et al.
Published: (2025)
by: Tsagkas, Nikolaos, et al.
Published: (2025)
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control
by: Park, Seongmin, et al.
Published: (2024)
by: Park, Seongmin, et al.
Published: (2024)
Next-Future: Sample-Efficient Policy Learning for Robotic-Arm Tasks
by: Özgür, Fikrican, et al.
Published: (2025)
by: Özgür, Fikrican, et al.
Published: (2025)
Control-oriented Clustering of Visual Latent Representation
by: Qi, Han, et al.
Published: (2024)
by: Qi, Han, et al.
Published: (2024)
M2CURL: Sample-Efficient Multimodal Reinforcement Learning via Self-Supervised Representation Learning for Robotic Manipulation
by: Lygerakis, Fotios, et al.
Published: (2024)
by: Lygerakis, Fotios, et al.
Published: (2024)
Visual Representation Learning with Stochastic Frame Prediction
by: Jang, Huiwon, et al.
Published: (2024)
by: Jang, Huiwon, et al.
Published: (2024)
Transferable Tactile Transformers for Representation Learning Across Diverse Sensors and Tasks
by: Zhao, Jialiang, et al.
Published: (2024)
by: Zhao, Jialiang, et al.
Published: (2024)
Scaling Proprioceptive-Visual Learning with Heterogeneous Pre-trained Transformers
by: Wang, Lirui, et al.
Published: (2024)
by: Wang, Lirui, et al.
Published: (2024)
Reinforcement Learning of Dolly-In Filming Using a Ground-Based Robot
by: Lorimer, Philip, et al.
Published: (2025)
by: Lorimer, Philip, et al.
Published: (2025)
Dense Policy: Bidirectional Autoregressive Learning of Actions
by: Su, Yue, et al.
Published: (2025)
by: Su, Yue, et al.
Published: (2025)
ET-SEED: Efficient Trajectory-Level SE(3) Equivariant Diffusion Policy
by: Tie, Chenrui, et al.
Published: (2024)
by: Tie, Chenrui, et al.
Published: (2024)
Video2Reward: Generating Reward Function from Videos for Legged Robot Behavior Learning
by: Zeng, Runhao, et al.
Published: (2024)
by: Zeng, Runhao, et al.
Published: (2024)
Planning-Guided Diffusion Policy Learning for Generalizable Contact-Rich Bimanual Manipulation
by: Li, Xuanlin, et al.
Published: (2024)
by: Li, Xuanlin, et al.
Published: (2024)
EfficientFlow: Efficient Equivariant Flow Policy Learning for Embodied AI
by: Chang, Jianlei, et al.
Published: (2025)
by: Chang, Jianlei, et al.
Published: (2025)
Mobi-$π$: Mobilizing Your Robot Learning Policy
by: Yang, Jingyun, et al.
Published: (2025)
by: Yang, Jingyun, et al.
Published: (2025)
Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos
by: Zhai, Albert J., et al.
Published: (2026)
by: Zhai, Albert J., et al.
Published: (2026)
Compressor-VLA: Instruction-Guided Visual Token Compression for Efficient Robotic Manipulation
by: Gao, Juntao, et al.
Published: (2025)
by: Gao, Juntao, et al.
Published: (2025)
Subtask-Aware Visual Reward Learning from Segmented Demonstrations
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
EP-Diffuser: An Efficient Diffusion Model for Traffic Scene Generation and Prediction via Polynomial Representations
by: Yao, Yue, et al.
Published: (2025)
by: Yao, Yue, et al.
Published: (2025)
From Single Images to Motion Policies via Video-Generation Environment Representations
by: Zhi, Weiming, et al.
Published: (2025)
by: Zhi, Weiming, et al.
Published: (2025)
CARIL: Confidence-Aware Regression in Imitation Learning for Autonomous Driving
by: Delavari, Elahe, et al.
Published: (2025)
by: Delavari, Elahe, et al.
Published: (2025)
Informed Reinforcement Learning for Situation-Aware Traffic Rule Exceptions
by: Bogdoll, Daniel, et al.
Published: (2024)
by: Bogdoll, Daniel, et al.
Published: (2024)
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
by: Almuzairee, Abdulaziz, et al.
Published: (2026)
A Recipe for Unbounded Data Augmentation in Visual Reinforcement Learning
by: Almuzairee, Abdulaziz, et al.
Published: (2024)
by: Almuzairee, Abdulaziz, et al.
Published: (2024)
Multi-Stage Manipulation with Demonstration-Augmented Reward, Policy, and World Model Learning
by: Escoriza, Adrià López, et al.
Published: (2025)
by: Escoriza, Adrià López, et al.
Published: (2025)
Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
by: Lin, Haitao, et al.
Published: (2026)
by: Lin, Haitao, et al.
Published: (2026)
Pre-trained Visual Dynamics Representations for Efficient Policy Learning
by: Luo, Hao, et al.
Published: (2024)
by: Luo, Hao, et al.
Published: (2024)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
by: Xie, Yichen, et al.
Published: (2026)
by: Xie, Yichen, et al.
Published: (2026)
SPA: 3D Spatial-Awareness Enables Effective Embodied Representation
by: Zhu, Haoyi, et al.
Published: (2024)
by: Zhu, Haoyi, et al.
Published: (2024)
Crossway Diffusion: Improving Diffusion-based Visuomotor Policy via Self-supervised Learning
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient
by: You, Haoxiang, et al.
Published: (2026)
by: You, Haoxiang, et al.
Published: (2026)
ReasonDrive: Efficient Visual Question Answering for Autonomous Vehicles with Reasoning-Enhanced Small Vision-Language Models
by: Chahe, Amirhosein, et al.
Published: (2025)
by: Chahe, Amirhosein, et al.
Published: (2025)
Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen)
by: Ren, Yifei, et al.
Published: (2025)
by: Ren, Yifei, et al.
Published: (2025)
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
by: Krauss, Henrik, et al.
Published: (2025)
by: Krauss, Henrik, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience
by: Chane-Sane, Elliot, et al.
Published: (2024)
by: Chane-Sane, Elliot, et al.
Published: (2024)
AnyTouch: Learning Unified Static-Dynamic Representation across Multiple Visuo-tactile Sensors
by: Feng, Ruoxuan, et al.
Published: (2025)
by: Feng, Ruoxuan, et al.
Published: (2025)
ILeSiA: Interactive Learning of Robot Situational Awareness from Camera Input
by: Vanc, Petr, et al.
Published: (2024)
by: Vanc, Petr, et al.
Published: (2024)
Similar Items
-
3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
by: Ze, Yanjie, et al.
Published: (2024) -
The Temporal Trap: Entanglement in Pre-Trained Visual Representations for Visuomotor Policy Learning
by: Tsagkas, Nikolaos, et al.
Published: (2025) -
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control
by: Park, Seongmin, et al.
Published: (2024) -
Next-Future: Sample-Efficient Policy Learning for Robotic-Arm Tasks
by: Özgür, Fikrican, et al.
Published: (2025) -
Control-oriented Clustering of Visual Latent Representation
by: Qi, Han, et al.
Published: (2024)