Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
Fuente:
arXiv
Salvato in:
| Autori principali: | Romio, Gabriel, Melchiades, Mateus Begnini, da Silva, Bruno Castro, Ramos, Gabriel de Oliveira |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
di: Li, Keyu, et al.
Pubblicazione: (2021)
di: Li, Keyu, et al.
Pubblicazione: (2021)
MRHER: Model-based Relay Hindsight Experience Replay for Sequential Object Manipulation Tasks with Sparse Rewards
di: Huang, Yuming, et al.
Pubblicazione: (2023)
di: Huang, Yuming, et al.
Pubblicazione: (2023)
ROER: Regularized Optimal Experience Replay
di: Li, Changling, et al.
Pubblicazione: (2024)
di: Li, Changling, et al.
Pubblicazione: (2024)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
di: Yunis, David, et al.
Pubblicazione: (2023)
di: Yunis, David, et al.
Pubblicazione: (2023)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
di: Diaz-Bone, Leander, et al.
Pubblicazione: (2025)
di: Diaz-Bone, Leander, et al.
Pubblicazione: (2025)
Cluster-based Sampling in Hindsight Experience Replay for Robotic Tasks (Student Abstract)
di: Kim, Taeyoung, et al.
Pubblicazione: (2022)
di: Kim, Taeyoung, et al.
Pubblicazione: (2022)
D-SPEAR: Dual-Stream Prioritized Experience Adaptive Replay for Stable Reinforcement Learning in Robotic Manipulation
di: Zhang, Yu, et al.
Pubblicazione: (2026)
di: Zhang, Yu, et al.
Pubblicazione: (2026)
Generalized Back-Stepping Experience Replay in Sparse-Reward Environments
di: Lyu, Guwen, et al.
Pubblicazione: (2024)
di: Lyu, Guwen, et al.
Pubblicazione: (2024)
Adaptable Hindsight Experience Replay for Search-Based Learning
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
di: Patel, Bhrij, et al.
Pubblicazione: (2023)
di: Patel, Bhrij, et al.
Pubblicazione: (2023)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
di: Zargarbashi, Fatemeh, et al.
Pubblicazione: (2024)
di: Zargarbashi, Fatemeh, et al.
Pubblicazione: (2024)
Learning Constraint Network from Demonstrations via Positive-Unlabeled Learning with Memory Replay
di: Peng, Baiyu, et al.
Pubblicazione: (2024)
di: Peng, Baiyu, et al.
Pubblicazione: (2024)
Hindsight Preference Replay Improves Preference-Conditioned Multi-Objective Reinforcement Learning
di: Shianifar, Jonaid, et al.
Pubblicazione: (2026)
di: Shianifar, Jonaid, et al.
Pubblicazione: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
di: Ishihara, Yu, et al.
Pubblicazione: (2025)
di: Ishihara, Yu, et al.
Pubblicazione: (2025)
Feature Aggregation with Latent Generative Replay for Federated Continual Learning of Socially Appropriate Robot Behaviours
di: Churamani, Nikhil, et al.
Pubblicazione: (2024)
di: Churamani, Nikhil, et al.
Pubblicazione: (2024)
Diffusion-Reward Adversarial Imitation Learning
di: Lai, Chun-Mao, et al.
Pubblicazione: (2024)
di: Lai, Chun-Mao, et al.
Pubblicazione: (2024)
Reward-Punishment Reinforcement Learning with Maximum Entropy
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
di: Wang, Jiexin, et al.
Pubblicazione: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
di: Liu, Yuyang, et al.
Pubblicazione: (2025)
di: Liu, Yuyang, et al.
Pubblicazione: (2025)
Adaptive Querying for Reward Learning from Human Feedback
di: Anand, Yashwanthi, et al.
Pubblicazione: (2024)
di: Anand, Yashwanthi, et al.
Pubblicazione: (2024)
Residual Reward Models for Preference-based Reinforcement Learning
di: Cao, Chenyang, et al.
Pubblicazione: (2025)
di: Cao, Chenyang, et al.
Pubblicazione: (2025)
Robot Policy Learning with Temporal Optimal Transport Reward
di: Fu, Yuwei, et al.
Pubblicazione: (2024)
di: Fu, Yuwei, et al.
Pubblicazione: (2024)
Batch Active Learning of Reward Functions from Human Preferences
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
di: Bıyık, Erdem, et al.
Pubblicazione: (2024)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
di: Mu, Tongzhou, et al.
Pubblicazione: (2024)
di: Mu, Tongzhou, et al.
Pubblicazione: (2024)
CaRL: Learning Scalable Planning Policies with Simple Rewards
di: Jaeger, Bernhard, et al.
Pubblicazione: (2025)
di: Jaeger, Bernhard, et al.
Pubblicazione: (2025)
A Generalized Acquisition Function for Preference-based Reward Learning
di: Ellis, Evan, et al.
Pubblicazione: (2024)
di: Ellis, Evan, et al.
Pubblicazione: (2024)
Reward Learning from Suboptimal Demonstrations with Applications in Surgical Electrocautery
di: Karimi, Zohre, et al.
Pubblicazione: (2024)
di: Karimi, Zohre, et al.
Pubblicazione: (2024)
SoftMimic: Learning Compliant Whole-body Control from Examples
di: Margolis, Gabriel B., et al.
Pubblicazione: (2025)
di: Margolis, Gabriel B., et al.
Pubblicazione: (2025)
CodeIt: Self-Improving Language Models with Prioritized Hindsight Replay
di: Butt, Natasha, et al.
Pubblicazione: (2024)
di: Butt, Natasha, et al.
Pubblicazione: (2024)
Hindsight-Anchored Policy Optimization: Turning Failure into Feedback in Sparse Reward Settings
di: Wu, Yuning, et al.
Pubblicazione: (2026)
di: Wu, Yuning, et al.
Pubblicazione: (2026)
Contact Energy Based Hindsight Experience Prioritization
di: Sayar, Erdi, et al.
Pubblicazione: (2023)
di: Sayar, Erdi, et al.
Pubblicazione: (2023)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
di: Tabakov, Stefan, et al.
Pubblicazione: (2025)
di: Tabakov, Stefan, et al.
Pubblicazione: (2025)
Can We Really Learn One Representation to Optimize All Rewards?
di: Zheng, Chongyi, et al.
Pubblicazione: (2026)
di: Zheng, Chongyi, et al.
Pubblicazione: (2026)
A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving
di: Abouelazm, Ahmed, et al.
Pubblicazione: (2024)
di: Abouelazm, Ahmed, et al.
Pubblicazione: (2024)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
di: Hwang, Minjune, et al.
Pubblicazione: (2026)
di: Hwang, Minjune, et al.
Pubblicazione: (2026)
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
di: Gumbsch, Christian, et al.
Pubblicazione: (2026)
di: Gumbsch, Christian, et al.
Pubblicazione: (2026)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
di: Zhang, Chen Bo Calvin, et al.
Pubblicazione: (2024)
di: Zhang, Chen Bo Calvin, et al.
Pubblicazione: (2024)
Diffusion Meets Options: Hierarchical Generative Skill Composition for Temporally-Extended Tasks
di: Feng, Zeyu, et al.
Pubblicazione: (2024)
di: Feng, Zeyu, et al.
Pubblicazione: (2024)
Learning to Recover: Dynamic Reward Shaping with Wheel-Leg Coordination for Fallen Robots
di: Deng, Boyuan, et al.
Pubblicazione: (2025)
di: Deng, Boyuan, et al.
Pubblicazione: (2025)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
di: Lee, Vint, et al.
Pubblicazione: (2023)
di: Lee, Vint, et al.
Pubblicazione: (2023)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
di: Guo, Yihong, et al.
Pubblicazione: (2024)
di: Guo, Yihong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
di: Li, Keyu, et al.
Pubblicazione: (2021) -
MRHER: Model-based Relay Hindsight Experience Replay for Sequential Object Manipulation Tasks with Sparse Rewards
di: Huang, Yuming, et al.
Pubblicazione: (2023) -
ROER: Regularized Optimal Experience Replay
di: Li, Changling, et al.
Pubblicazione: (2024) -
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
di: Yunis, David, et al.
Pubblicazione: (2023) -
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
di: Diaz-Bone, Leander, et al.
Pubblicazione: (2025)