Potential-Based Intrinsic Motivation: Preserving Optimality With Complex, Non-Markovian Shaping Rewards
Fuente:
arXiv
Saved in:
| Main Authors: | Forbes, Grant C., Villalobos-Arias, Leonardo, Wang, Jianxun, Jhala, Arnav, Roberts, David L. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Action-Dependent Optimality-Preserving Reward Shaping
by: Forbes, Grant C., et al.
Published: (2025)
by: Forbes, Grant C., et al.
Published: (2025)
Potential-Based Reward Shaping For Intrinsic Motivation
by: Forbes, Grant C., et al.
Published: (2024)
by: Forbes, Grant C., et al.
Published: (2024)
Minding Motivation: The Effect of Intrinsic Motivation on Agent Behaviors
by: Villalobos-Arias, Leonardo, et al.
Published: (2025)
by: Villalobos-Arias, Leonardo, et al.
Published: (2025)
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
by: Quadros, André, et al.
Published: (2025)
by: Quadros, André, et al.
Published: (2025)
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)
by: Khorasani, Sadegh, et al.
Published: (2025)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026)
by: Tang, Wenjie, et al.
Published: (2026)
PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management
by: Zaregarizi, Shadmehr, et al.
Published: (2026)
by: Zaregarizi, Shadmehr, et al.
Published: (2026)
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
by: McCane, Brendan
Published: (2026)
by: McCane, Brendan
Published: (2026)
Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
by: Parris, William
Published: (2026)
by: Parris, William
Published: (2026)
Cost and Reward Infused Metric Elicitation
by: Bhateja, Chethan, et al.
Published: (2025)
by: Bhateja, Chethan, et al.
Published: (2025)
GPz: Non-stationary sparse Gaussian processes for heteroscedastic uncertainty estimation in photometric redshifts
by: Almosallam, Ibrahim A., et al.
Published: (2016)
by: Almosallam, Ibrahim A., et al.
Published: (2016)
Enhancing Heterogeneous Multi-Agent Cooperation in Decentralized MARL via GNN-driven Intrinsic Rewards
by: Monon, Jahir Sadik, et al.
Published: (2024)
by: Monon, Jahir Sadik, et al.
Published: (2024)
Rewarded Region Replay (R3) for Policy Learning with Discrete Action Space
by: Li, Bangzheng, et al.
Published: (2024)
by: Li, Bangzheng, et al.
Published: (2024)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
by: Singha, Disha
Published: (2026)
by: Singha, Disha
Published: (2026)
Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia
by: Smith, Chandler, et al.
Published: (2025)
by: Smith, Chandler, et al.
Published: (2025)
PRPO: Aligning Process Reward with Outcome Reward in Policy Optimization
by: Ding, Ruiyi, et al.
Published: (2026)
by: Ding, Ruiyi, et al.
Published: (2026)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
2Mamba2Furious: Linear in Complexity, Competitive in Accuracy
by: Mongaras, Gabriel, et al.
Published: (2026)
by: Mongaras, Gabriel, et al.
Published: (2026)
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
by: Calvo, Anherutowa
Published: (2026)
by: Calvo, Anherutowa
Published: (2026)
A Boundary-Aware Non-parametric Granular-Ball Classifier Based on Minimum Description Length
by: Xian, Zeqiang, et al.
Published: (2026)
by: Xian, Zeqiang, et al.
Published: (2026)
Difference Rewards Policy Gradients
by: Castellini, Jacopo, et al.
Published: (2020)
by: Castellini, Jacopo, et al.
Published: (2020)
Understanding Variational Autoencoders with Intrinsic Dimension and Information Imbalance
by: Camboulin, Charles, et al.
Published: (2024)
by: Camboulin, Charles, et al.
Published: (2024)
X-Factor: Quality Is a Dataset-Intrinsic Property
by: Couch, Josiah, et al.
Published: (2025)
by: Couch, Josiah, et al.
Published: (2025)
Measuring In-Context Computation Complexity via Hidden State Prediction
by: Herrmann, Vincent, et al.
Published: (2025)
by: Herrmann, Vincent, et al.
Published: (2025)
Getting ViT in Shape: Scaling Laws for Compute-Optimal Model Design
by: Alabdulmohsin, Ibrahim, et al.
Published: (2023)
by: Alabdulmohsin, Ibrahim, et al.
Published: (2023)
Semi-Supervised Learning for AVO Inversion with Strong Spatial Feature Constraints
by: Liu, Yingtian, et al.
Published: (2025)
by: Liu, Yingtian, et al.
Published: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
by: Fadli, Samih
Published: (2025)
by: Fadli, Samih
Published: (2025)
Deep Complex-valued Radial Basis Function Neural Networks and Parameter Selection
by: Soares, Jonathan A., et al.
Published: (2024)
by: Soares, Jonathan A., et al.
Published: (2024)
Conservative Bias in Multi-Teacher Learning: Why Agents Prefer Low-Reward Advisors
by: Mesto, Maher, et al.
Published: (2025)
by: Mesto, Maher, et al.
Published: (2025)
Enhanced Protein Intrinsic Disorder Prediction Through Dual-View Multiscale Features and Multi-objective Evolutionary Algorithm
by: Wang, Shaokuan, et al.
Published: (2026)
by: Wang, Shaokuan, et al.
Published: (2026)
NFR: Neural Feature-Guided Non-Rigid Shape Registration
by: Chen, Zhangquan, et al.
Published: (2025)
by: Chen, Zhangquan, et al.
Published: (2025)
Exploring Neural Granger Causality with xLSTMs: Unveiling Temporal Dependencies in Complex Data
by: Poonia, Harsh, et al.
Published: (2025)
by: Poonia, Harsh, et al.
Published: (2025)
CoxSE: Exploring the Potential of Self-Explaining Neural Networks with Cox Proportional Hazards Model for Survival Analysis
by: Alabdallah, Abdallah, et al.
Published: (2024)
by: Alabdallah, Abdallah, et al.
Published: (2024)
Prompting Neural-Guided Equation Discovery Based on Residuals
by: Brugger, Jannis, et al.
Published: (2025)
by: Brugger, Jannis, et al.
Published: (2025)
Extrinsicaly Rewarded Soft Q Imitation Learning with Discriminator
by: Furuyama, Ryoma, et al.
Published: (2024)
by: Furuyama, Ryoma, et al.
Published: (2024)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
by: Pal, Biplab, et al.
Published: (2026)
by: Pal, Biplab, et al.
Published: (2026)
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
by: Zhou, Yangyang, et al.
Published: (2026)
by: Zhou, Yangyang, et al.
Published: (2026)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
by: Belcamino, Valerio, et al.
Published: (2026)
by: Belcamino, Valerio, et al.
Published: (2026)
TimeCatcher: A Variational Framework for Volatility-Aware Forecasting of Non-Stationary Time Series
by: Chen, Zhiyu, et al.
Published: (2026)
by: Chen, Zhiyu, et al.
Published: (2026)
Similar Items
-
Action-Dependent Optimality-Preserving Reward Shaping
by: Forbes, Grant C., et al.
Published: (2025) -
Potential-Based Reward Shaping For Intrinsic Motivation
by: Forbes, Grant C., et al.
Published: (2024) -
Minding Motivation: The Effect of Intrinsic Motivation on Agent Behaviors
by: Villalobos-Arias, Leonardo, et al.
Published: (2025) -
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
by: Quadros, André, et al.
Published: (2025) -
Fusing Rewards and Preferences in Reinforcement Learning
by: Khorasani, Sadegh, et al.
Published: (2025)