Information-theoretic analysis of world models in optimal reward maximizers
Fuente:
arXiv
Saved in:
| Main Authors: | Harwood, Alfred, Faustino, Jose, Altair, Alex |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What should be observed for optimal reward in POMDPs?
by: Konsta, Alyzia-Maria, et al.
Published: (2024)
by: Konsta, Alyzia-Maria, et al.
Published: (2024)
Agent-centric learning: from external reward maximization to internal knowledge curation
by: Zhou, Hanqi, et al.
Published: (2025)
by: Zhou, Hanqi, et al.
Published: (2025)
TripScore: Benchmarking and rewarding real-world travel planning with fine-grained evaluation
by: Qu, Yincen, et al.
Published: (2025)
by: Qu, Yincen, et al.
Published: (2025)
Towards better dense rewards in Reinforcement Learning Applications
by: Zhang, Shuyuan
Published: (2025)
by: Zhang, Shuyuan
Published: (2025)
Reasoning over mathematical objects: on-policy reward modeling and test time aggregation
by: Aggarwal, Pranjal, et al.
Published: (2026)
by: Aggarwal, Pranjal, et al.
Published: (2026)
Active teacher selection for reward learning
by: Freedman, Rachel, et al.
Published: (2023)
by: Freedman, Rachel, et al.
Published: (2023)
Self-rewarding correction for mathematical reasoning
by: Xiong, Wei, et al.
Published: (2025)
by: Xiong, Wei, et al.
Published: (2025)
Noise-based reward-modulated learning
by: Fernández, Jesús García, et al.
Published: (2025)
by: Fernández, Jesús García, et al.
Published: (2025)
Order-theoretic models for decision-making: Learning, optimization, complexity and computation
by: Hack, Pedro
Published: (2024)
by: Hack, Pedro
Published: (2024)
GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
by: Liu, Shih-Yang, et al.
Published: (2026)
by: Liu, Shih-Yang, et al.
Published: (2026)
Preference learning in shades of gray: Interpretable and bias-aware reward modeling for human preferences
by: Oprea, Simona-Vasilica, et al.
Published: (2026)
by: Oprea, Simona-Vasilica, et al.
Published: (2026)
Class-Conditional self-reward mechanism for improved Text-to-Image models
by: Ghazouali, Safouane El, et al.
Published: (2024)
by: Ghazouali, Safouane El, et al.
Published: (2024)
The impact of intrinsic rewards on exploration in Reinforcement Learning
by: Kayal, Aya, et al.
Published: (2025)
by: Kayal, Aya, et al.
Published: (2025)
EVAL: EigenVector-based Average-reward Learning
by: Adamczyk, Jacob, et al.
Published: (2025)
by: Adamczyk, Jacob, et al.
Published: (2025)
Streaming Looking Ahead with Token-level Self-reward
by: Zhang, Hongming, et al.
Published: (2025)
by: Zhang, Hongming, et al.
Published: (2025)
Episodic Reinforcement Learning with Expanded State-reward Space
by: Liang, Dayang, et al.
Published: (2024)
by: Liang, Dayang, et al.
Published: (2024)
Self-supervised network distillation: an effective approach to exploration in sparse reward environments
by: Pecháč, Matej, et al.
Published: (2023)
by: Pecháč, Matej, et al.
Published: (2023)
R-ParVI: Particle-based variational inference through lens of rewards
by: Huang, Yongchao
Published: (2025)
by: Huang, Yongchao
Published: (2025)
Deep Reinforcement Learning with anticipatory reward in LSTM for Collision Avoidance of Mobile Robots
by: Poulet, Olivier, et al.
Published: (2025)
by: Poulet, Olivier, et al.
Published: (2025)
Bayesian Flow Networks
by: Graves, Alex, et al.
Published: (2023)
by: Graves, Alex, et al.
Published: (2023)
HelpSteer2: Open-source dataset for training top-performing reward models
by: Wang, Zhilin, et al.
Published: (2024)
by: Wang, Zhilin, et al.
Published: (2024)
Information-theoretic Distinctions Between Deception and Confusion
by: Young, Robin
Published: (2025)
by: Young, Robin
Published: (2025)
reward-lens: A Mechanistic Interpretability Library for Reward Models
by: Nadaf, Mohammed Suhail B
Published: (2026)
by: Nadaf, Mohammed Suhail B
Published: (2026)
MOSLIM:Align with diverse preferences in prompts through reward classification
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
Risk-averse Total-reward MDPs with ERM and EVaR
by: Su, Xihong, et al.
Published: (2024)
by: Su, Xihong, et al.
Published: (2024)
Synthesizing world models for bilevel planning
by: Ahmed, Zergham, et al.
Published: (2025)
by: Ahmed, Zergham, et al.
Published: (2025)
Assessing model error in counterfactual worlds
by: Howerton, Emily, et al.
Published: (2025)
by: Howerton, Emily, et al.
Published: (2025)
LLM world models are mental: Output layer evidence of brittle world model use in LLM mechanical reasoning
by: Robertson, Cole, et al.
Published: (2025)
by: Robertson, Cole, et al.
Published: (2025)
Towards an Inferentialist Account of Information Through Proof-theoretic Semantics
by: Collinson, Matthew, et al.
Published: (2026)
by: Collinson, Matthew, et al.
Published: (2026)
Just Say What You Want: Only-prompting Self-rewarding Online Preference Optimization
by: Xu, Ruijie, et al.
Published: (2024)
by: Xu, Ruijie, et al.
Published: (2024)
Reasoning with maximal consistent signatures
by: Thimm, Matthias, et al.
Published: (2024)
by: Thimm, Matthias, et al.
Published: (2024)
On the Weaknesses of Backdoor-based Model Watermarking: An Information-theoretic Perspective
by: Hu, Aoting, et al.
Published: (2024)
by: Hu, Aoting, et al.
Published: (2024)
Leveraging LLMs for reward function design in reinforcement learning control tasks
by: Cardenoso, Franklin, et al.
Published: (2025)
by: Cardenoso, Franklin, et al.
Published: (2025)
Information-theoretic Bayesian Optimization: Survey and Tutorial
by: Garrido-Merchán, Eduardo C.
Published: (2025)
by: Garrido-Merchán, Eduardo C.
Published: (2025)
BLEUBERI: BLEU is a surprisingly effective reward for instruction following
by: Chang, Yapei, et al.
Published: (2025)
by: Chang, Yapei, et al.
Published: (2025)
The geometry of invariant learning: an information-theoretic analysis of data augmentation and generalization
by: Bouyahia, Abdelali, et al.
Published: (2026)
by: Bouyahia, Abdelali, et al.
Published: (2026)
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals
by: Molinghen, Yannick, et al.
Published: (2025)
by: Molinghen, Yannick, et al.
Published: (2025)
From monoliths to modules: Decomposing transducers for efficient world modelling
by: Boyd, Alexander, et al.
Published: (2025)
by: Boyd, Alexander, et al.
Published: (2025)
A Pre-training Framework for Relational Data with Information-theoretic Principles
by: Truong, Quang, et al.
Published: (2025)
by: Truong, Quang, et al.
Published: (2025)
On minimizers and convolutional filters: theoretical connections and applications to genome analysis
by: Yu, Yun William
Published: (2021)
by: Yu, Yun William
Published: (2021)
Similar Items
-
What should be observed for optimal reward in POMDPs?
by: Konsta, Alyzia-Maria, et al.
Published: (2024) -
Agent-centric learning: from external reward maximization to internal knowledge curation
by: Zhou, Hanqi, et al.
Published: (2025) -
TripScore: Benchmarking and rewarding real-world travel planning with fine-grained evaluation
by: Qu, Yincen, et al.
Published: (2025) -
Towards better dense rewards in Reinforcement Learning Applications
by: Zhang, Shuyuan
Published: (2025) -
Reasoning over mathematical objects: on-policy reward modeling and test time aggregation
by: Aggarwal, Pranjal, et al.
Published: (2026)