Conservative Offline Robot Policy Learning via Posterior-Transition Reweighting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Wanpeng, Luo, Hao, Zheng, Sipeng, Feng, Yicheng, Xu, Haiweng, Xi, Ziheng, Xu, Chaoyi, Yuan, Haoqi, Lu, Zongqing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models
von: Xu, Haiweng, et al.
Veröffentlicht: (2026)
von: Xu, Haiweng, et al.
Veröffentlicht: (2026)
Being-H0.5: Scaling Human-Centric Robot Learning for Cross-Embodiment Generalization
von: Luo, Hao, et al.
Veröffentlicht: (2026)
von: Luo, Hao, et al.
Veröffentlicht: (2026)
Joint-Aligned Latent Action: Towards Scalable VLA Pretraining in the Wild
von: Luo, Hao, et al.
Veröffentlicht: (2026)
von: Luo, Hao, et al.
Veröffentlicht: (2026)
Being-H0.7: A Latent World-Action Model from Egocentric Videos
von: Luo, Hao, et al.
Veröffentlicht: (2026)
von: Luo, Hao, et al.
Veröffentlicht: (2026)
Rethinking Visual-Language-Action Model Scaling: Alignment, Mixture, and Regularization
von: Wang, Ye, et al.
Veröffentlicht: (2026)
von: Wang, Ye, et al.
Veröffentlicht: (2026)
Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos
von: Feng, Yicheng, et al.
Veröffentlicht: (2025)
von: Feng, Yicheng, et al.
Veröffentlicht: (2025)
DiG-Flow: Discrepancy-Guided Flow Matching for Robust VLA Models
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos
von: Luo, Hao, et al.
Veröffentlicht: (2025)
von: Luo, Hao, et al.
Veröffentlicht: (2025)
UniTacHand: Unified Spatio-Tactile Representation for Human to Robotic Hand Skill Transfer
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
von: Zhang, Chi, et al.
Veröffentlicht: (2025)
Universal Dexterous Functional Grasping via Demonstration-Editing Reinforcement Learning
von: Mao, Chuan, et al.
Veröffentlicht: (2025)
von: Mao, Chuan, et al.
Veröffentlicht: (2025)
DemoGrasp: Universal Dexterous Grasping from a Single Demonstration
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
DemoHLM: From One Demonstration to Generalizable Humanoid Loco-Manipulation
von: Fu, Yuhui, et al.
Veröffentlicht: (2025)
von: Fu, Yuhui, et al.
Veröffentlicht: (2025)
X-DiffVLA: X-Embodied Diffusion Action Heads for Vision-Language-Action Models
von: Li, Boyu, et al.
Veröffentlicht: (2026)
von: Li, Boyu, et al.
Veröffentlicht: (2026)
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2025)
Unified Multimodal Understanding via Byte-Pair Visual Encoding
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2025)
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping
von: Huang, Ziye, et al.
Veröffentlicht: (2024)
von: Huang, Ziye, et al.
Veröffentlicht: (2024)
Learning Diverse Bimanual Dexterous Manipulation Skills from Human Demonstrations
von: Zhou, Bohan, et al.
Veröffentlicht: (2024)
von: Zhou, Bohan, et al.
Veröffentlicht: (2024)
Cross-Embodiment Dexterous Grasping with Reinforcement Learning
von: Yuan, Haoqi, et al.
Veröffentlicht: (2024)
von: Yuan, Haoqi, et al.
Veröffentlicht: (2024)
VideoOrion: Tokenizing Object Dynamics in Videos
von: Feng, Yicheng, et al.
Veröffentlicht: (2024)
von: Feng, Yicheng, et al.
Veröffentlicht: (2024)
From Pixels to Tokens: Byte-Pair Encoding on Quantized Visual Modalities
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2024)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2024)
Towards Proprioception-Aware Embodied Planning for Dual-Arm Humanoid Robots
von: Li, Boyu, et al.
Veröffentlicht: (2025)
von: Li, Boyu, et al.
Veröffentlicht: (2025)
RL from Physical Feedback: Aligning Large Motion Models with Humanoid Control
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
von: Yue, Junpeng, et al.
Veröffentlicht: (2025)
Guided Policy Optimization under Partial Observability
von: Li, Yueheng, et al.
Veröffentlicht: (2025)
von: Li, Yueheng, et al.
Veröffentlicht: (2025)
SPAFormer: Sequential 3D Part Assembly with Transformers
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
COSBO: Conservative Offline Simulation-Based Policy Optimization
von: Kargar, Eshagh, et al.
Veröffentlicht: (2024)
von: Kargar, Eshagh, et al.
Veröffentlicht: (2024)
Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
UniCode: Learning a Unified Codebook for Multimodal Large Language Models
von: Zheng, Sipeng, et al.
Veröffentlicht: (2024)
von: Zheng, Sipeng, et al.
Veröffentlicht: (2024)
AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
von: Zhang, Wanpeng, et al.
Veröffentlicht: (2023)
DualTHOR: A Dual-Arm Humanoid Simulation Platform for Contingency-Aware Planning
von: Li, Boyu, et al.
Veröffentlicht: (2025)
von: Li, Boyu, et al.
Veröffentlicht: (2025)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
von: Xu, Charles, et al.
Veröffentlicht: (2024)
von: Xu, Charles, et al.
Veröffentlicht: (2024)
UniGraspTransformer: Simplified Policy Distillation for Scalable Dexterous Robotic Grasping
von: Wang, Wenbo, et al.
Veröffentlicht: (2024)
von: Wang, Wenbo, et al.
Veröffentlicht: (2024)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
von: Liu, Tenglong, et al.
Veröffentlicht: (2024)
From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
Policy Contrastive Decoding for Robotic Foundation Models
von: Wu, Shihan, et al.
Veröffentlicht: (2025)
von: Wu, Shihan, et al.
Veröffentlicht: (2025)
Beyond Static Datasets: Robust Offline Policy Optimization via Vetted Synthetic Transitions
von: Agand, Pedram, et al.
Veröffentlicht: (2026)
von: Agand, Pedram, et al.
Veröffentlicht: (2026)
Adaptive Visual Perception for Robotic Construction Process: A Multi-Robot Coordination Framework
von: Xu, Jia, et al.
Veröffentlicht: (2024)
von: Xu, Jia, et al.
Veröffentlicht: (2024)
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
von: Lei, Kun, et al.
Veröffentlicht: (2023)
von: Lei, Kun, et al.
Veröffentlicht: (2023)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
Reference-Steering via Data-Driven Predictive Control for Hyper-Accurate Robotic Flying-Hopping Locomotion
von: Zeng, Yicheng, et al.
Veröffentlicht: (2024)
von: Zeng, Yicheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models
von: Xu, Haiweng, et al.
Veröffentlicht: (2026) -
Being-H0.5: Scaling Human-Centric Robot Learning for Cross-Embodiment Generalization
von: Luo, Hao, et al.
Veröffentlicht: (2026) -
Joint-Aligned Latent Action: Towards Scalable VLA Pretraining in the Wild
von: Luo, Hao, et al.
Veröffentlicht: (2026) -
Being-H0.7: A Latent World-Action Model from Egocentric Videos
von: Luo, Hao, et al.
Veröffentlicht: (2026) -
Rethinking Visual-Language-Action Model Scaling: Alignment, Mixture, and Regularization
von: Wang, Ye, et al.
Veröffentlicht: (2026)