Deep Reinforcement Learning with Hybrid Intrinsic Reward Model
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Mingqi, Li, Bo, Jin, Xin, Zeng, Wenjun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Data Exploitation in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2024)
by: Yuan, Mingqi, et al.
Published: (2024)
Goal-Driven Reward by Video Diffusion Models for Reinforcement Learning
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
RLLTE: Long-Term Evolution Project of Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2023)
by: Yuan, Mingqi, et al.
Published: (2023)
Plasticine: Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
PvP: Data-Efficient Humanoid Robot Learning with Proprioceptive-Privileged Contrastive Representations
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
Hierarchical Procedural Framework for Low-latency Robot-Assisted Hand-Object Interaction
by: Yuan, Mingqi, et al.
Published: (2024)
by: Yuan, Mingqi, et al.
Published: (2024)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
by: Wang, Qi, et al.
Published: (2023)
by: Wang, Qi, et al.
Published: (2023)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
by: Nguyen, Viet Bac, et al.
Published: (2026)
by: Nguyen, Viet Bac, et al.
Published: (2026)
Open-World Reinforcement Learning over Long Short-Term Imagination
by: Li, Jiajian, et al.
Published: (2024)
by: Li, Jiajian, et al.
Published: (2024)
Intrinsic Vicarious Conditioning for Deep Reinforcement Learning
by: Sanchez, Rodney A, et al.
Published: (2026)
by: Sanchez, Rodney A, et al.
Published: (2026)
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
by: Quadros, André, et al.
Published: (2025)
by: Quadros, André, et al.
Published: (2025)
Effective Reward Specification in Deep Reinforcement Learning
by: Roy, Julien
Published: (2024)
by: Roy, Julien
Published: (2024)
Reward Models in Deep Reinforcement Learning: A Survey
by: Yu, Rui, et al.
Published: (2025)
by: Yu, Rui, et al.
Published: (2025)
Beyond Simple Sum of Delayed Rewards: Non-Markovian Reward Modeling for Reinforcement Learning
by: Tang, Yuting, et al.
Published: (2024)
by: Tang, Yuting, et al.
Published: (2024)
Black Box Meta-Learning Intrinsic Rewards
by: Pappalardo, Octavio, et al.
Published: (2024)
by: Pappalardo, Octavio, et al.
Published: (2024)
Continual Policy Distillation from Distributed Reinforcement Learning Teachers
by: Li, Yuxuan, et al.
Published: (2026)
by: Li, Yuxuan, et al.
Published: (2026)
Reinforcement Learning from Bagged Reward
by: Tang, Yuting, et al.
Published: (2024)
by: Tang, Yuting, et al.
Published: (2024)
LERO: LLM-driven Evolutionary framework with Hybrid Rewards and Enhanced Observation for Multi-Agent Reinforcement Learning
by: Wei, Yuan, et al.
Published: (2025)
by: Wei, Yuan, et al.
Published: (2025)
Federated Multi-Agent Deep Reinforcement Learning Approach via Physics-Informed Reward for Multi-Microgrid Energy Management
by: Li, Yuanzheng, et al.
Published: (2022)
by: Li, Yuanzheng, et al.
Published: (2022)
Analyzing and Bridging the Gap between Maximizing Total Reward and Discounted Reward in Deep Reinforcement Learning
by: Yin, Shuyu, et al.
Published: (2024)
by: Yin, Shuyu, et al.
Published: (2024)
Shrinking the Variance: Shrinkage Baselines for Reinforcement Learning with Verifiable Rewards
by: Zeng, Guanning, et al.
Published: (2025)
by: Zeng, Guanning, et al.
Published: (2025)
Mutual Information as Intrinsic Reward of Reinforcement Learning Agents for On-demand Ride Pooling
by: Zhang, Xianjie, et al.
Published: (2023)
by: Zhang, Xianjie, et al.
Published: (2023)
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning
by: Wan, Zhenglin, et al.
Published: (2025)
by: Wan, Zhenglin, et al.
Published: (2025)
MIR: Efficient Exploration in Episodic Multi-Agent Reinforcement Learning via Mutual Intrinsic Reward
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
by: Miranda, Victor R. F., et al.
Published: (2022)
by: Miranda, Victor R. F., et al.
Published: (2022)
Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards
by: Huang, Yu, et al.
Published: (2026)
by: Huang, Yu, et al.
Published: (2026)
BCR-DRL: Behavior- and Context-aware Reward for Deep Reinforcement Learning in Human-AI Coordination
by: Hao, Xin, et al.
Published: (2024)
by: Hao, Xin, et al.
Published: (2024)
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Mol-AIR: Molecular Reinforcement Learning with Adaptive Intrinsic Rewards for Goal-directed Molecular Generation
by: Park, Jinyeong, et al.
Published: (2024)
by: Park, Jinyeong, et al.
Published: (2024)
Reward-Conditioned Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2026)
by: Nauman, Michal, et al.
Published: (2026)
Hybrid Reinforcement: When Reward Is Sparse, It's Better to Be Dense
by: Tao, Leitian, et al.
Published: (2025)
by: Tao, Leitian, et al.
Published: (2025)
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise
by: Wu, Xuefei, et al.
Published: (2025)
by: Wu, Xuefei, et al.
Published: (2025)
Intrinsic Reward Policy Optimization for Sparse-Reward Environments
by: Cho, Minjae, et al.
Published: (2026)
by: Cho, Minjae, et al.
Published: (2026)
Vision Transformers that Never Stop Learning
by: Sun, Caihao, et al.
Published: (2026)
by: Sun, Caihao, et al.
Published: (2026)
Token Hidden Reward: Steering Exploration-Exploitation in Group Relative Deep Reinforcement Learning
by: Deng, Wenlong, et al.
Published: (2025)
by: Deng, Wenlong, et al.
Published: (2025)
Entropy Centroids as Intrinsic Rewards for Test-Time Scaling
by: Zhao, Wenshuo, et al.
Published: (2026)
by: Zhao, Wenshuo, et al.
Published: (2026)
ELO-Rated Sequence Rewards: Advancing Reinforcement Learning Models
by: Ju, Qi, et al.
Published: (2024)
by: Ju, Qi, et al.
Published: (2024)
Reward Design for Reinforcement Learning Agents
by: Devidze, Rati
Published: (2025)
by: Devidze, Rati
Published: (2025)
Similar Items
-
Adaptive Data Exploitation in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025) -
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025) -
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2024) -
Goal-Driven Reward by Video Diffusion Models for Reinforcement Learning
by: Wang, Qi, et al.
Published: (2025) -
RLLTE: Long-Term Evolution Project of Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2023)