On the Sample Efficiency of Abstractions and Potential-Based Reward Shaping in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Canonaco, Giuseppe, Ardon, Leo, Pozanco, Alberto, Borrajo, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Projection Abstractions in Planning Under the Lenses of Abstractions for MDPs
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2024)
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2024)
Semantic Partial Grounding via LLMs
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2026)
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2026)
On Learning Action Costs from Input Plans
von: Morales, Marianela, et al.
Veröffentlicht: (2024)
von: Morales, Marianela, et al.
Veröffentlicht: (2024)
Learning Robust Reward Machines from Noisy Labels
von: Parac, Roko, et al.
Veröffentlicht: (2024)
von: Parac, Roko, et al.
Veröffentlicht: (2024)
Subgoal-based Reward Shaping to Improve Efficiency in Reinforcement Learning
von: Okudo, Takato, et al.
Veröffentlicht: (2021)
von: Okudo, Takato, et al.
Veröffentlicht: (2021)
On Computing Plans with Uniform Action Costs
von: Pozanco, Alberto, et al.
Veröffentlicht: (2024)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2024)
Counterfactual Reasoning in Automated Planning
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
von: Azran, Guy, et al.
Veröffentlicht: (2023)
von: Azran, Guy, et al.
Veröffentlicht: (2023)
Contrastive Abstraction for Reinforcement Learning
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
Generalising Planning Environment Redesign
von: Pozanco, Alberto, et al.
Veröffentlicht: (2024)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2024)
Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
von: Ma, Haozhe, et al.
Veröffentlicht: (2024)
Unveiling Interesting Insights: Monte Carlo Tree Search for Knowledge Discovery
von: Totis, Pietro, et al.
Veröffentlicht: (2025)
von: Totis, Pietro, et al.
Veröffentlicht: (2025)
Entropy-informed Decoding: Adaptive Information-Driven Branching
von: Evans, Benjamin Patrick, et al.
Veröffentlicht: (2026)
von: Evans, Benjamin Patrick, et al.
Veröffentlicht: (2026)
Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach
von: Li, Wenyun, et al.
Veröffentlicht: (2025)
von: Li, Wenyun, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Symbolic Reward Machines
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
von: Krug, Thomas, et al.
Veröffentlicht: (2026)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
von: Wibault, Clarisse, et al.
Veröffentlicht: (2026)
Learning Markov State Abstractions for Deep Reinforcement Learning
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
Tactical Decision Making for Autonomous Trucks by Deep Reinforcement Learning with Total Cost of Operation Based Reward
von: Pathare, Deepthi, et al.
Veröffentlicht: (2024)
von: Pathare, Deepthi, et al.
Veröffentlicht: (2024)
Planning with Minimal Disruption
von: Pozanco, Alberto, et al.
Veröffentlicht: (2025)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2025)
A Planning Compilation to Reason about Goal Achievement at Planning Time
von: Pozanco, Alberto, et al.
Veröffentlicht: (2025)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2025)
Planning Task Shielding: Detecting and Repairing Flaws in Planning Tasks through Turning them Unsolvable
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
von: Pozanco, Alberto, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Stochastic Reward Machines
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
von: Corazza, Jan, et al.
Veröffentlicht: (2025)
Attention-Based Reward Shaping for Sparse and Delayed Rewards
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
von: Holmes, Ian, et al.
Veröffentlicht: (2025)
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
Enhancing Inverse Reinforcement Learning through Encoding Dynamic Information in Reward Shaping
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
von: Metcalf, Katherine, et al.
Veröffentlicht: (2024)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
von: Lee, Hosung, et al.
Veröffentlicht: (2024)
von: Lee, Hosung, et al.
Veröffentlicht: (2024)
Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions
von: Nayyar, Rashmeet Kaur, et al.
Veröffentlicht: (2025)
von: Nayyar, Rashmeet Kaur, et al.
Veröffentlicht: (2025)
Realizable Abstractions: Near-Optimal Hierarchical Reinforcement Learning
von: Cipollone, Roberto, et al.
Veröffentlicht: (2025)
von: Cipollone, Roberto, et al.
Veröffentlicht: (2025)
Search-Based Adversarial Estimates for Improving Sample Efficiency in Off-Policy Reinforcement Learning
von: Malato, Federico, et al.
Veröffentlicht: (2025)
von: Malato, Federico, et al.
Veröffentlicht: (2025)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
Decoupled Hierarchical Reinforcement Learning with State Abstraction for Discrete Grids
von: Xiao, Qingyu, et al.
Veröffentlicht: (2025)
von: Xiao, Qingyu, et al.
Veröffentlicht: (2025)
Honesty to Subterfuge: In-Context Reinforcement Learning Can Make Honest Models Reward Hack
von: McKee-Reid, Leo, et al.
Veröffentlicht: (2024)
von: McKee-Reid, Leo, et al.
Veröffentlicht: (2024)
LEASE: Offline Preference-based Reinforcement Learning with High Sample Efficiency
von: Liu, Xiao-Yin, et al.
Veröffentlicht: (2024)
von: Liu, Xiao-Yin, et al.
Veröffentlicht: (2024)
Generalizing Behavior via Inverse Reinforcement Learning with Closed-Form Reward Centroids
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
Speculative Sampling with Reinforcement Learning
von: Wang, Chenan, et al.
Veröffentlicht: (2026)
von: Wang, Chenan, et al.
Veröffentlicht: (2026)
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
von: Zhao, Mingde, et al.
Veröffentlicht: (2023)
von: Zhao, Mingde, et al.
Veröffentlicht: (2023)
Goal-Oriented Skill Abstraction for Offline Multi-Task Reinforcement Learning
von: He, Jinmin, et al.
Veröffentlicht: (2025)
von: He, Jinmin, et al.
Veröffentlicht: (2025)
Bootstrapped Reward Shaping
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
Boosting Sample Efficiency and Generalization in Multi-agent Reinforcement Learning via Equivariance
von: McClellan, Joshua, et al.
Veröffentlicht: (2024)
von: McClellan, Joshua, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Projection Abstractions in Planning Under the Lenses of Abstractions for MDPs
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2024) -
Semantic Partial Grounding via LLMs
von: Canonaco, Giuseppe, et al.
Veröffentlicht: (2026) -
On Learning Action Costs from Input Plans
von: Morales, Marianela, et al.
Veröffentlicht: (2024) -
Learning Robust Reward Machines from Noisy Labels
von: Parac, Roko, et al.
Veröffentlicht: (2024) -
Subgoal-based Reward Shaping to Improve Efficiency in Reinforcement Learning
von: Okudo, Takato, et al.
Veröffentlicht: (2021)