Informativeness of Reward Functions in Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Devidze, Rati, Kamalaruban, Parameswaran, Singla, Adish |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reward Design for Reinforcement Learning Agents
by: Devidze, Rati
Published: (2025)
by: Devidze, Rati
Published: (2025)
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
by: Tzannetos, Georgios, et al.
Published: (2024)
by: Tzannetos, Georgios, et al.
Published: (2024)
Curriculum Design for Trajectory-Constrained Agent: Compressing Chain-of-Thought Tokens in LLMs
by: Tzannetos, Georgios, et al.
Published: (2025)
by: Tzannetos, Georgios, et al.
Published: (2025)
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback
by: Nika, Andi, et al.
Published: (2026)
by: Nika, Andi, et al.
Published: (2026)
Corruption Robust Offline Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2024)
by: Mandal, Debmalya, et al.
Published: (2024)
Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences
by: Nika, Andi, et al.
Published: (2024)
by: Nika, Andi, et al.
Published: (2024)
Policy Teaching via Data Poisoning in Learning from Human Preferences
by: Nika, Andi, et al.
Published: (2025)
by: Nika, Andi, et al.
Published: (2025)
Inference-Time Personalized Alignment with a Few User Preference Queries
by: Pădurean, Victor-Alexandru, et al.
Published: (2025)
by: Pădurean, Victor-Alexandru, et al.
Published: (2025)
Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms
by: Nöther, Jonathan, et al.
Published: (2025)
by: Nöther, Jonathan, et al.
Published: (2025)
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints
by: Nöther, Jonathan, et al.
Published: (2025)
by: Nöther, Jonathan, et al.
Published: (2025)
Adversarially Robust Decision Transformer
by: Tang, Xiaohang, et al.
Published: (2024)
by: Tang, Xiaohang, et al.
Published: (2024)
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation
by: Kotalwar, Nachiket, et al.
Published: (2024)
by: Kotalwar, Nachiket, et al.
Published: (2024)
Learning Embeddings for Sequential Tasks Using Population of Agents
by: Mahajan, Mridul, et al.
Published: (2023)
by: Mahajan, Mridul, et al.
Published: (2023)
Towards Generalizable Agents in Text-Based Educational Environments: A Study of Integrating RL with LLMs
by: Radmehr, Bahar, et al.
Published: (2024)
by: Radmehr, Bahar, et al.
Published: (2024)
MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems
by: Nöther, Jonathan, et al.
Published: (2026)
by: Nöther, Jonathan, et al.
Published: (2026)
Formal Models of Active Learning from Contrastive Examples
by: Mansouri, Farnam, et al.
Published: (2025)
by: Mansouri, Farnam, et al.
Published: (2025)
Emergent Bias and Fairness in Multi-Agent Decision Systems
by: Madigan, Maeve, et al.
Published: (2025)
by: Madigan, Maeve, et al.
Published: (2025)
Corruption-Robust Offline Two-Player Zero-Sum Markov Games
by: Nika, Andi, et al.
Published: (2024)
by: Nika, Andi, et al.
Published: (2024)
Evaluating Fairness in Transaction Fraud Models: Fairness Metrics, Bias Audits, and Challenges
by: Kamalaruban, Parameswaran, et al.
Published: (2024)
by: Kamalaruban, Parameswaran, et al.
Published: (2024)
Neural Task Synthesis for Visual Programming
by: Pădurean, Victor-Alexandru, et al.
Published: (2023)
by: Pădurean, Victor-Alexandru, et al.
Published: (2023)
Learning Half-Spaces from Perturbed Contrastive Examples
by: Ravari, Aryan Alavi Razavi, et al.
Published: (2026)
by: Ravari, Aryan Alavi Razavi, et al.
Published: (2026)
Fairness-Aware Low-Rank Adaptation Under Demographic Privacy Constraints
by: Kamalaruban, Parameswaran, et al.
Published: (2025)
by: Kamalaruban, Parameswaran, et al.
Published: (2025)
Curriculum Reinforcement Learning for Complex Reward Functions
by: Freitag, Kilian, et al.
Published: (2024)
by: Freitag, Kilian, et al.
Published: (2024)
Learning Personalized Decision Support Policies
by: Bhatt, Umang, et al.
Published: (2023)
by: Bhatt, Umang, et al.
Published: (2023)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
Optimal Decision Making Under Strategic Behavior
by: Tsirtsis, Stratis, et al.
Published: (2019)
by: Tsirtsis, Stratis, et al.
Published: (2019)
DMAP: A Distribution Map for Text
by: Kempton, Tom, et al.
Published: (2026)
by: Kempton, Tom, et al.
Published: (2026)
Reward-Conditioned Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2026)
by: Nauman, Michal, et al.
Published: (2026)
Parent-Guided Semantic Reward Model (PGSRM): Embedding-Based Reward Functions for Reinforcement Learning of Transformer Language Models
by: Plashchinsky, Alexandr
Published: (2025)
by: Plashchinsky, Alexandr
Published: (2025)
Reinforcement Learning from Bagged Reward
by: Tang, Yuting, et al.
Published: (2024)
by: Tang, Yuting, et al.
Published: (2024)
The Value of Reward Lookahead in Reinforcement Learning
by: Merlis, Nadav, et al.
Published: (2024)
by: Merlis, Nadav, et al.
Published: (2024)
To the Max: Reinventing Reward in Reinforcement Learning
by: Veviurko, Grigorii, et al.
Published: (2024)
by: Veviurko, Grigorii, et al.
Published: (2024)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
by: Frans, Kevin, et al.
Published: (2024)
by: Frans, Kevin, et al.
Published: (2024)
The Distributional Reward Critic Framework for Reinforcement Learning Under Perturbed Rewards
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards
by: Huang, Yu, et al.
Published: (2026)
by: Huang, Yu, et al.
Published: (2026)
A Review of Reward Functions for Reinforcement Learning in the context of Autonomous Driving
by: Abouelazm, Ahmed, et al.
Published: (2024)
by: Abouelazm, Ahmed, et al.
Published: (2024)
Code as Reward: Empowering Reinforcement Learning with VLMs
by: Venuto, David, et al.
Published: (2024)
by: Venuto, David, et al.
Published: (2024)
Effective Reward Specification in Deep Reinforcement Learning
by: Roy, Julien
Published: (2024)
by: Roy, Julien
Published: (2024)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Adapting the Behavior of Reinforcement Learning Agents to Changing Action Spaces and Reward Functions
by: de la Rosa, Raul, et al.
Published: (2026)
by: de la Rosa, Raul, et al.
Published: (2026)
Similar Items
-
Reward Design for Reinforcement Learning Agents
by: Devidze, Rati
Published: (2025) -
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
by: Tzannetos, Georgios, et al.
Published: (2024) -
Curriculum Design for Trajectory-Constrained Agent: Compressing Chain-of-Thought Tokens in LLMs
by: Tzannetos, Georgios, et al.
Published: (2025) -
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback
by: Nika, Andi, et al.
Published: (2026) -
Corruption Robust Offline Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2024)