The impact of intrinsic rewards on exploration in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kayal, Aya, Pignatelli, Eduardo, Toni, Laura |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023)
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
NAVIX: Scaling MiniGrid Environments with JAX
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024)
Near-Optimal Sample Complexity in Reward-Free Kernel-Based Reinforcement Learning
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
Episodic Reinforcement Learning with Expanded State-reward Space
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
von: Liang, Dayang, et al.
Veröffentlicht: (2024)
Reinforcement Learning Using known Invariances
von: Cioba, Alexandru, et al.
Veröffentlicht: (2025)
von: Cioba, Alexandru, et al.
Veröffentlicht: (2025)
EVAL: EigenVector-based Average-reward Learning
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Self-rewarding correction for mathematical reasoning
von: Xiong, Wei, et al.
Veröffentlicht: (2025)
von: Xiong, Wei, et al.
Veröffentlicht: (2025)
Noise-based reward-modulated learning
von: Fernández, Jesús García, et al.
Veröffentlicht: (2025)
von: Fernández, Jesús García, et al.
Veröffentlicht: (2025)
Active teacher selection for reward learning
von: Freedman, Rachel, et al.
Veröffentlicht: (2023)
von: Freedman, Rachel, et al.
Veröffentlicht: (2023)
Hyperbolic Residual Quantization: Discrete Representations for Data with Latent Hierarchies
von: Piękos, Piotr, et al.
Veröffentlicht: (2025)
von: Piękos, Piotr, et al.
Veröffentlicht: (2025)
Combining Deep Architectures for Information Gain estimation and Reinforcement Learning for multiagent field exploration
von: Masiero, Emanuele, et al.
Veröffentlicht: (2025)
von: Masiero, Emanuele, et al.
Veröffentlicht: (2025)
Heterogeneous Graph Structure Learning through the Lens of Data-generating Processes
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
Streaming Looking Ahead with Token-level Self-reward
von: Zhang, Hongming, et al.
Veröffentlicht: (2025)
von: Zhang, Hongming, et al.
Veröffentlicht: (2025)
Neuro-Argumentative Learning with Case-Based Reasoning
von: Gould, Adam, et al.
Veröffentlicht: (2025)
von: Gould, Adam, et al.
Veröffentlicht: (2025)
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2026)
Risk-averse Total-reward MDPs with ERM and EVaR
von: Su, Xihong, et al.
Veröffentlicht: (2024)
von: Su, Xihong, et al.
Veröffentlicht: (2024)
reward-lens: A Mechanistic Interpretability Library for Reward Models
von: Nadaf, Mohammed Suhail B
Veröffentlicht: (2026)
von: Nadaf, Mohammed Suhail B
Veröffentlicht: (2026)
Learning controllable dynamics through informative exploration
von: Loxley, Peter N., et al.
Veröffentlicht: (2025)
von: Loxley, Peter N., et al.
Veröffentlicht: (2025)
From In Silico to In Vitro: Evaluating Molecule Generative Models for Hit Generation
von: Osman, Nagham, et al.
Veröffentlicht: (2025)
von: Osman, Nagham, et al.
Veröffentlicht: (2025)
Bures-Wasserstein Flow Matching for Graph Generation
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
von: Jiang, Keyue, et al.
Veröffentlicht: (2025)
Shapley-PC: Constraint-based Causal Structure Learning with a Shapley Inspired Framework
von: Russo, Fabrizio, et al.
Veröffentlicht: (2023)
von: Russo, Fabrizio, et al.
Veröffentlicht: (2023)
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals
von: Molinghen, Yannick, et al.
Veröffentlicht: (2025)
von: Molinghen, Yannick, et al.
Veröffentlicht: (2025)
Leveraging LLMs for reward function design in reinforcement learning control tasks
von: Cardenoso, Franklin, et al.
Veröffentlicht: (2025)
von: Cardenoso, Franklin, et al.
Veröffentlicht: (2025)
Robust Counterfactual Explanations in Machine Learning: A Survey
von: Jiang, Junqi, et al.
Veröffentlicht: (2024)
von: Jiang, Junqi, et al.
Veröffentlicht: (2024)
Bayesian Optimization from Human Feedback: Near-Optimal Regret Bounds
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
BLEUBERI: BLEU is a surprisingly effective reward for instruction following
von: Chang, Yapei, et al.
Veröffentlicht: (2025)
von: Chang, Yapei, et al.
Veröffentlicht: (2025)
Automating Curriculum Learning for Reinforcement Learning using a Skill-Based Bayesian Network
von: Hsiao, Vincent, et al.
Veröffentlicht: (2025)
von: Hsiao, Vincent, et al.
Veröffentlicht: (2025)
Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
Epistemically-guided forward-backward exploration
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
Learning to Optimize for Reinforcement Learning
von: Lan, Qingfeng, et al.
Veröffentlicht: (2023)
von: Lan, Qingfeng, et al.
Veröffentlicht: (2023)
Agent-centric learning: from external reward maximization to internal knowledge curation
von: Zhou, Hanqi, et al.
Veröffentlicht: (2025)
von: Zhou, Hanqi, et al.
Veröffentlicht: (2025)
Is there Value in Reinforcement Learning?
von: Fox, Lior, et al.
Veröffentlicht: (2025)
von: Fox, Lior, et al.
Veröffentlicht: (2025)
Reinforcement Learning: An Overview
von: Murphy, Kevin
Veröffentlicht: (2024)
von: Murphy, Kevin
Veröffentlicht: (2024)
Introduction to Reinforcement Learning
von: Ghasemi, Majid, et al.
Veröffentlicht: (2024)
von: Ghasemi, Majid, et al.
Veröffentlicht: (2024)
Highway Reinforcement Learning
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
von: Wang, Yuhui, et al.
Veröffentlicht: (2024)
Experiential Reinforcement Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
von: Shi, Taiwei, et al.
Veröffentlicht: (2026)
Towards Interpretable Deep Reinforcement Learning Models via Inverse Reinforcement Learning
von: Xie, Sean, et al.
Veröffentlicht: (2022)
von: Xie, Sean, et al.
Veröffentlicht: (2022)
Object-Centric Neuro-Argumentative Learning
von: Jacob, Abdul Rahman, et al.
Veröffentlicht: (2025)
von: Jacob, Abdul Rahman, et al.
Veröffentlicht: (2025)
HelpSteer2: Open-source dataset for training top-performing reward models
von: Wang, Zhilin, et al.
Veröffentlicht: (2024)
von: Wang, Zhilin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Survey of Temporal Credit Assignment in Deep Reinforcement Learning
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2023) -
Assessing the Zero-Shot Capabilities of LLMs for Action Evaluation in RL
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024) -
NAVIX: Scaling MiniGrid Environments with JAX
von: Pignatelli, Eduardo, et al.
Veröffentlicht: (2024) -
Near-Optimal Sample Complexity in Reward-Free Kernel-Based Reinforcement Learning
von: Kayal, Aya, et al.
Veröffentlicht: (2025) -
Episodic Reinforcement Learning with Expanded State-reward Space
von: Liang, Dayang, et al.
Veröffentlicht: (2024)