Guardado en:
| Autores principales: | Ren, Hanxiang, Zhou, Pei, Zhou, Xunzhe, Yang, Yanchao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.20856 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
por: Zhou, Pei, et al.
Publicado: (2025)
por: Zhou, Pei, et al.
Publicado: (2025)
Decoupled Action Expert: Confining Task Knowledge to the Conditioning Pathway
por: Zhou, Jian, et al.
Publicado: (2025)
por: Zhou, Jian, et al.
Publicado: (2025)
InfoCon: Concept Discovery with Generative and Discriminative Informativeness
por: Liu, Ruizhe, et al.
Publicado: (2024)
por: Liu, Ruizhe, et al.
Publicado: (2024)
MaxMI: A Maximal Mutual Information Criterion for Manipulation Concept Discovery
por: Zhou, Pei, et al.
Publicado: (2024)
por: Zhou, Pei, et al.
Publicado: (2024)
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
por: Höpner, Niklas, et al.
Publicado: (2025)
por: Höpner, Niklas, et al.
Publicado: (2025)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
por: Zhang, Jesse, et al.
Publicado: (2023)
por: Zhang, Jesse, et al.
Publicado: (2023)
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
por: Zhou, Zihan, et al.
Publicado: (2025)
por: Zhou, Zihan, et al.
Publicado: (2025)
GenDexHand: Generative Simulation for Dexterous Hands
por: Chen, Feng, et al.
Publicado: (2025)
por: Chen, Feng, et al.
Publicado: (2025)
How to Provably Improve Return Conditioned Supervised Learning?
por: Liu, Zhishuai, et al.
Publicado: (2025)
por: Liu, Zhishuai, et al.
Publicado: (2025)
Flattening Hierarchies with Policy Bootstrapping
por: Zhou, John L., et al.
Publicado: (2025)
por: Zhou, John L., et al.
Publicado: (2025)
Learning Human-Humanoid Coordination for Collaborative Object Carrying
por: Du, Yushi, et al.
Publicado: (2025)
por: Du, Yushi, et al.
Publicado: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
por: Patel, Bhrij, et al.
Publicado: (2023)
por: Patel, Bhrij, et al.
Publicado: (2023)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
por: Nie, Buqing, et al.
Publicado: (2025)
por: Nie, Buqing, et al.
Publicado: (2025)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
por: Koo, Juil, et al.
Publicado: (2026)
por: Koo, Juil, et al.
Publicado: (2026)
Multi-Agent Reinforcement Learning for Unmanned Aerial Vehicle Coordination by Multi-Critic Policy Gradient Optimization
por: Alon, Yoav, et al.
Publicado: (2020)
por: Alon, Yoav, et al.
Publicado: (2020)
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms
por: Zhou, Zehao
Publicado: (2024)
por: Zhou, Zehao
Publicado: (2024)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
por: Ma, Yunchang, et al.
Publicado: (2025)
por: Ma, Yunchang, et al.
Publicado: (2025)
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Implicit Maximum Likelihood Estimation for Real-time Generative Model Predictive Control
por: Lee, Grayson, et al.
Publicado: (2026)
por: Lee, Grayson, et al.
Publicado: (2026)
A Comprehensive Survey of Cross-Domain Policy Transfer for Embodied Agents
por: Niu, Haoyi, et al.
Publicado: (2024)
por: Niu, Haoyi, et al.
Publicado: (2024)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
por: Cai, Shizhe, et al.
Publicado: (2025)
por: Cai, Shizhe, et al.
Publicado: (2025)
Variational Distillation of Diffusion Policies into Mixture of Experts
por: Zhou, Hongyi, et al.
Publicado: (2024)
por: Zhou, Hongyi, et al.
Publicado: (2024)
RAPTOR: A Foundation Policy for Quadrotor Control
por: Eschmann, Jonas, et al.
Publicado: (2025)
por: Eschmann, Jonas, et al.
Publicado: (2025)
Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
por: Kapoor, Aditya, et al.
Publicado: (2024)
por: Kapoor, Aditya, et al.
Publicado: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
por: Xue, Han, et al.
Publicado: (2025)
por: Xue, Han, et al.
Publicado: (2025)
ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning
por: Chen, Wendi, et al.
Publicado: (2025)
por: Chen, Wendi, et al.
Publicado: (2025)
Imagination Policy: Using Generative Point Cloud Models for Learning Manipulation Policies
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
SMAT: Staged Multi-Agent Training for Co-Adaptive Exoskeleton Control
por: Yuan, Yifei, et al.
Publicado: (2026)
por: Yuan, Yifei, et al.
Publicado: (2026)
Safe Exploration via Policy Priors
por: Wendl, Manuel, et al.
Publicado: (2026)
por: Wendl, Manuel, et al.
Publicado: (2026)
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
por: Duan, Yuanlin, et al.
Publicado: (2024)
por: Duan, Yuanlin, et al.
Publicado: (2024)
SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration
por: Jin, Yang, et al.
Publicado: (2025)
por: Jin, Yang, et al.
Publicado: (2025)
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
por: Zhang, Yixue, et al.
Publicado: (2026)
por: Zhang, Yixue, et al.
Publicado: (2026)
Discrete Variational Autoencoding via Policy Search
por: Drolet, Michael, et al.
Publicado: (2025)
por: Drolet, Michael, et al.
Publicado: (2025)
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
por: Chen, Yi, et al.
Publicado: (2026)
por: Chen, Yi, et al.
Publicado: (2026)
A Multi-Fidelity Control Variate Approach for Policy Gradient Estimation
por: Liu, Xinjie, et al.
Publicado: (2025)
por: Liu, Xinjie, et al.
Publicado: (2025)
RoboPocket: Improve Robot Policies Instantly with Your Phone
por: Fang, Junjie, et al.
Publicado: (2026)
por: Fang, Junjie, et al.
Publicado: (2026)
On the Evaluation of Generative Robotic Simulations
por: Chen, Feng, et al.
Publicado: (2024)
por: Chen, Feng, et al.
Publicado: (2024)
Guiding Data Collection via Factored Scaling Curves
por: Zha, Lihan, et al.
Publicado: (2025)
por: Zha, Lihan, et al.
Publicado: (2025)
IMLE Policy: Fast and Sample Efficient Visuomotor Policy Learning via Implicit Maximum Likelihood Estimation
por: Rana, Krishan, et al.
Publicado: (2025)
por: Rana, Krishan, et al.
Publicado: (2025)
Ejemplares similares
-
Hyper-GoalNet: Goal-Conditioned Manipulation Policy Learning with HyperNetworks
por: Zhou, Pei, et al.
Publicado: (2025) -
Decoupled Action Expert: Confining Task Knowledge to the Conditioning Pathway
por: Zhou, Jian, et al.
Publicado: (2025) -
InfoCon: Concept Discovery with Generative and Discriminative Informativeness
por: Liu, Ruizhe, et al.
Publicado: (2024) -
MaxMI: A Maximal Mutual Information Criterion for Manipulation Concept Discovery
por: Zhou, Pei, et al.
Publicado: (2024) -
Data Augmentation for Instruction Following Policies via Trajectory Segmentation
por: Höpner, Niklas, et al.
Publicado: (2025)