PCHC: Enabling Preference Conditioned Humanoid Control via Multi-Objective Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Huanyu, Wang, Dewei, Wang, Xinmiao, Liu, Xinzhe, Liu, Peng, Bai, Chenjia, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Adaptive Humanoid Control via Multi-Behavior Distillation and Reinforced Fine-Tuning
by: Zhao, Yingnan, et al.
Published: (2025)
by: Zhao, Yingnan, et al.
Published: (2025)
MoRE: Mixture of Residual Experts for Humanoid Lifelike Gaits Learning on Complex Terrains
by: Wang, Dewei, et al.
Published: (2025)
by: Wang, Dewei, et al.
Published: (2025)
HUSKY: Humanoid Skateboarding System via Physics-Aware Whole-Body Control
by: Han, Jinrui, et al.
Published: (2026)
by: Han, Jinrui, et al.
Published: (2026)
X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation
by: Wang, Dewei, et al.
Published: (2026)
by: Wang, Dewei, et al.
Published: (2026)
Adversarial Locomotion and Motion Imitation for Humanoid Policy Learning
by: Shi, Jiyuan, et al.
Published: (2025)
by: Shi, Jiyuan, et al.
Published: (2025)
KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills
by: Xie, Weiji, et al.
Published: (2025)
by: Xie, Weiji, et al.
Published: (2025)
Pro-HOI: Perceptive Root-guided Humanoid-Object Interaction
by: Lin, Yuhang, et al.
Published: (2026)
by: Lin, Yuhang, et al.
Published: (2026)
Learning Soccer Skills for Humanoid Robots: A Progressive Perception-Action Framework
by: Kong, Jipeng, et al.
Published: (2026)
by: Kong, Jipeng, et al.
Published: (2026)
Humanoid Whole-Body Locomotion on Narrow Terrain via Dynamic Balance and Reinforcement Learning
by: Xie, Weiji, et al.
Published: (2025)
by: Xie, Weiji, et al.
Published: (2025)
TextOp: Real-time Interactive Text-Driven Humanoid Robot Motion Generation and Control
by: Xie, Weiji, et al.
Published: (2026)
by: Xie, Weiji, et al.
Published: (2026)
Preference Aligned Diffusion Planner for Quadrupedal Locomotion Control
by: Yuan, Xinyi, et al.
Published: (2024)
by: Yuan, Xinyi, et al.
Published: (2024)
Preference-Conditioned Multi-Objective RL for Integrated Command Tracking and Force Compliance in Humanoid Locomotion
by: Leng, Tingxuan, et al.
Published: (2025)
by: Leng, Tingxuan, et al.
Published: (2025)
Gait-Conditioned Reinforcement Learning with Multi-Phase Curriculum for Humanoid Locomotion
by: Peng, Tianhu, et al.
Published: (2025)
by: Peng, Tianhu, et al.
Published: (2025)
Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface
by: Wang, Dewei, et al.
Published: (2025)
by: Wang, Dewei, et al.
Published: (2025)
InterReal: A Unified Physics-Based Imitation Framework for Learning Human-Object Interaction Skills
by: Liang, Dayang, et al.
Published: (2026)
by: Liang, Dayang, et al.
Published: (2026)
Humanoid Whole-Body Badminton via Multi-Stage Reinforcement Learning
by: Liu, Chenhao, et al.
Published: (2025)
by: Liu, Chenhao, et al.
Published: (2025)
VLP: Vision-Language Preference Learning for Embodied Manipulation
by: Liu, Runze, et al.
Published: (2025)
by: Liu, Runze, et al.
Published: (2025)
HoRD: Robust Humanoid Control via History-Conditioned Reinforcement Learning and Online Distillation
by: Wang, Puyue, et al.
Published: (2026)
by: Wang, Puyue, et al.
Published: (2026)
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2020)
by: Bai, Chenjia, et al.
Published: (2020)
DeCoNav: Dialog enhanced Long-Horizon Collaborative Vision-Language Navigation
by: Zhou, Sunyao, et al.
Published: (2026)
by: Zhou, Sunyao, et al.
Published: (2026)
KungfuBot2: Learning Versatile Motion Skills for Humanoid Whole-Body Control
by: Han, Jinrui, et al.
Published: (2025)
by: Han, Jinrui, et al.
Published: (2025)
HALO:Closing Sim-to-Real Gap for Heavy-loaded Humanoid Agile Motion Skills via Differentiable Simulation
by: Wang, Xingyi, et al.
Published: (2026)
by: Wang, Xingyi, et al.
Published: (2026)
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction
by: Fan, Chenyou, et al.
Published: (2025)
by: Fan, Chenyou, et al.
Published: (2025)
Regularized Conditional Diffusion Model for Multi-Task Preference Alignment
by: Yu, Xudong, et al.
Published: (2024)
by: Yu, Xudong, et al.
Published: (2024)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Beyond Short-Horizon: VQ-Memory for Robust Long-Horizon Manipulation in Non-Markovian Simulation Benchmarks
by: Wang, Honghui, et al.
Published: (2026)
by: Wang, Honghui, et al.
Published: (2026)
RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting
by: Xin, Yucheng, et al.
Published: (2026)
by: Xin, Yucheng, et al.
Published: (2026)
Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
HumanoidGen: Data Generation for Bimanual Dexterous Manipulation via LLM Reasoning
by: Jing, Zhi, et al.
Published: (2025)
by: Jing, Zhi, et al.
Published: (2025)
Chasing Stability: Humanoid Running via Control Lyapunov Function Guided Reinforcement Learning
by: Olkin, Zachary, et al.
Published: (2025)
by: Olkin, Zachary, et al.
Published: (2025)
Towards Reliable LLM-based Robot Planning via Combined Uncertainty Estimation
by: Yin, Shiyuan, et al.
Published: (2025)
by: Yin, Shiyuan, et al.
Published: (2025)
ZeroWBC: Learning Natural Visuomotor Humanoid Control Directly from Human Egocentric Video
by: Yang, Haoran, et al.
Published: (2026)
by: Yang, Haoran, et al.
Published: (2026)
ECO: Energy-Constrained Optimization with Reinforcement Learning for Humanoid Walking
by: Huang, Weidong, et al.
Published: (2026)
by: Huang, Weidong, et al.
Published: (2026)
Learn Weightlessness: Imitate Non-Self-Stabilizing Motions on Humanoid Robot
by: Xin, Yucheng, et al.
Published: (2026)
by: Xin, Yucheng, et al.
Published: (2026)
RobotDancing: Residual-Action Reinforcement Learning Enables Robust Long-Horizon Humanoid Motion Tracking
by: Sun, Zhenguo, et al.
Published: (2025)
by: Sun, Zhenguo, et al.
Published: (2025)
Do You Have Freestyle? Expressive Humanoid Locomotion via Audio Control
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Learning to Evolve: Multi-modal Interactive Fields for Robust Humanoid Navigation in Dynamic Environments
by: Jiang, Peifeng, et al.
Published: (2026)
by: Jiang, Peifeng, et al.
Published: (2026)
Mechanical Intelligence-Aware Curriculum Reinforcement Learning for Humanoids with Parallel Actuation
by: Tanaka, Yusuke, et al.
Published: (2025)
by: Tanaka, Yusuke, et al.
Published: (2025)
Scaling Tasks, Not Samples: Mastering Humanoid Control through Multi-Task Model-Based Reinforcement Learning
by: Liu, Shaohuai, et al.
Published: (2026)
by: Liu, Shaohuai, et al.
Published: (2026)
Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach
by: Yang, Siyuan, et al.
Published: (2025)
by: Yang, Siyuan, et al.
Published: (2025)
Similar Items
-
Towards Adaptive Humanoid Control via Multi-Behavior Distillation and Reinforced Fine-Tuning
by: Zhao, Yingnan, et al.
Published: (2025) -
MoRE: Mixture of Residual Experts for Humanoid Lifelike Gaits Learning on Complex Terrains
by: Wang, Dewei, et al.
Published: (2025) -
HUSKY: Humanoid Skateboarding System via Physics-Aware Whole-Body Control
by: Han, Jinrui, et al.
Published: (2026) -
X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation
by: Wang, Dewei, et al.
Published: (2026) -
Adversarial Locomotion and Motion Imitation for Humanoid Policy Learning
by: Shi, Jiyuan, et al.
Published: (2025)