Learning Emergent Gaits with Decentralized Phase Oscillators: on the role of Observations, Rewards, and Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Jenny, Heim, Steve, Jeon, Se Hwan, Kim, Sangbae |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FLD: Fourier Latent Dynamics for Structured Motion Representation and Learning
by: Li, Chenhao, et al.
Published: (2024)
by: Li, Chenhao, et al.
Published: (2024)
Learning Humanoid Arm Motion via Centroidal Momentum Regularized Multi-Agent Reinforcement Learning
by: Lee, Ho Jae, et al.
Published: (2025)
by: Lee, Ho Jae, et al.
Published: (2025)
Residual MPC: Blending Reinforcement Learning with GPU-Parallelized Model Predictive Control
by: Jeon, Se Hwan, et al.
Published: (2025)
by: Jeon, Se Hwan, et al.
Published: (2025)
CusADi: A GPU Parallelization Framework for Symbolic Expressions and Optimal Control
by: Jeon, Se Hwan, et al.
Published: (2024)
by: Jeon, Se Hwan, et al.
Published: (2024)
Policy Learning from Large Vision-Language Model Feedback without Reward Modeling
by: Luu, Tung M., et al.
Published: (2025)
by: Luu, Tung M., et al.
Published: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
Adaptive Querying for Reward Learning from Human Feedback
by: Anand, Yashwanthi, et al.
Published: (2024)
by: Anand, Yashwanthi, et al.
Published: (2024)
Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control
by: Lee, Ho Jae, et al.
Published: (2026)
by: Lee, Ho Jae, et al.
Published: (2026)
STRIDE: Automating Reward Design, Deep Reinforcement Learning Training and Feedback Optimization in Humanoid Robotics Locomotion
by: Wu, Zhenwei, et al.
Published: (2025)
by: Wu, Zhenwei, et al.
Published: (2025)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
by: Hwang, Minjune, et al.
Published: (2026)
by: Hwang, Minjune, et al.
Published: (2026)
Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion
by: Bohlinger, Nico, et al.
Published: (2025)
by: Bohlinger, Nico, et al.
Published: (2025)
Safe Value Functions
by: Massiani, Pierre-François, et al.
Published: (2021)
by: Massiani, Pierre-François, et al.
Published: (2021)
Average-Reward Maximum Entropy Reinforcement Learning for Underactuated Double Pendulum Tasks
by: Choe, Jean Seong Bjorn, et al.
Published: (2024)
by: Choe, Jean Seong Bjorn, et al.
Published: (2024)
Gaitor: Learning a Unified Representation Across Gaits for Real-World Quadruped Locomotion
by: Mitchell, Alexander L., et al.
Published: (2024)
by: Mitchell, Alexander L., et al.
Published: (2024)
Learn to Swim: Data-Driven LSTM Hydrodynamic Model for Quadruped Robot Gait Optimization
by: Han, Fei, et al.
Published: (2025)
by: Han, Fei, et al.
Published: (2025)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
by: Feng, Youhe, et al.
Published: (2026)
by: Feng, Youhe, et al.
Published: (2026)
CoordLight: Learning Decentralized Coordination for Network-Wide Traffic Signal Control
by: Zhang, Yifeng, et al.
Published: (2026)
by: Zhang, Yifeng, et al.
Published: (2026)
Enhancing Rating-Based Reinforcement Learning to Effectively Leverage Feedback from Large Vision-Language Models
by: Luu, Tung Minh, et al.
Published: (2025)
by: Luu, Tung Minh, et al.
Published: (2025)
Curriculum Reinforcement Learning for Complex Reward Functions
by: Freitag, Kilian, et al.
Published: (2024)
by: Freitag, Kilian, et al.
Published: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
by: Biza, Ondrej, et al.
Published: (2024)
by: Biza, Ondrej, et al.
Published: (2024)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
by: Vasan, Gautham, et al.
Published: (2024)
by: Vasan, Gautham, et al.
Published: (2024)
Probabilistic Safety Guarantee for Stochastic Control Systems Using Average Reward MDPs
by: Omidi, Saber, et al.
Published: (2025)
by: Omidi, Saber, et al.
Published: (2025)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
by: Miranda, Victor R. F., et al.
Published: (2022)
by: Miranda, Victor R. F., et al.
Published: (2022)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
by: Krack, Pierre, et al.
Published: (2026)
by: Krack, Pierre, et al.
Published: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
The Dark Side of Rich Rewards: Understanding and Mitigating Noise in VLM Rewards
by: Huang, Sukai, et al.
Published: (2024)
by: Huang, Sukai, et al.
Published: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
by: Tang, Nan, et al.
Published: (2025)
by: Tang, Nan, et al.
Published: (2025)
SCDP: Learning Humanoid Locomotion from Partial Observations via Mixed-Observation Distillation
by: Carroll, Milo, et al.
Published: (2026)
by: Carroll, Milo, et al.
Published: (2026)
Evidence of an Emergent "Self" in Continual Robot Learning
by: Jhunjhunwala, Adidev, et al.
Published: (2026)
by: Jhunjhunwala, Adidev, et al.
Published: (2026)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
by: Zheng, Qinqing, et al.
Published: (2024)
by: Zheng, Qinqing, et al.
Published: (2024)
Robot Policy Learning with Temporal Optimal Transport Reward
by: Fu, Yuwei, et al.
Published: (2024)
by: Fu, Yuwei, et al.
Published: (2024)
Improving Generalization Ability of Robotic Imitation Learning by Resolving Causal Confusion in Observations
by: Chen, Yifei, et al.
Published: (2025)
by: Chen, Yifei, et al.
Published: (2025)
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics
by: Chen, Letian, et al.
Published: (2024)
by: Chen, Letian, et al.
Published: (2024)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
by: Venugopal, Aravind, et al.
Published: (2026)
by: Venugopal, Aravind, et al.
Published: (2026)
Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
by: Freitag, Kilian, et al.
Published: (2026)
by: Freitag, Kilian, et al.
Published: (2026)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
Teaching Robots to Handle Nuclear Waste: A Teleoperation-Based Learning Approach<
by: Lee, Joong-Ku, et al.
Published: (2025)
by: Lee, Joong-Ku, et al.
Published: (2025)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
Intentionally-underestimated Value Function at Terminal State for Temporal-difference Learning with Mis-designed Reward
by: Kobayashi, Taisuke
Published: (2023)
by: Kobayashi, Taisuke
Published: (2023)
Decentralized Gaussian Process Classification and an Application in Subsea Robotics
by: Gao, Yifei, et al.
Published: (2025)
by: Gao, Yifei, et al.
Published: (2025)
Similar Items
-
FLD: Fourier Latent Dynamics for Structured Motion Representation and Learning
by: Li, Chenhao, et al.
Published: (2024) -
Learning Humanoid Arm Motion via Centroidal Momentum Regularized Multi-Agent Reinforcement Learning
by: Lee, Ho Jae, et al.
Published: (2025) -
Residual MPC: Blending Reinforcement Learning with GPU-Parallelized Model Predictive Control
by: Jeon, Se Hwan, et al.
Published: (2025) -
CusADi: A GPU Parallelization Framework for Symbolic Expressions and Optimal Control
by: Jeon, Se Hwan, et al.
Published: (2024) -
Policy Learning from Large Vision-Language Model Feedback without Reward Modeling
by: Luu, Tung M., et al.
Published: (2025)