Integrating Policy Summaries with Reward Decomposition for Explaining Reinforcement Learning Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Septon, Yael, Huber, Tobias, André, Elisabeth, Amir, Ofra |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
"Trust me on this" Explaining Agent Behavior to a Human Terminator
by: Menkes, Uri, et al.
Published: (2025)
by: Menkes, Uri, et al.
Published: (2025)
A Model for Intelligible Interaction Between Agents That Predict and Explain
by: Baskar, A., et al.
Published: (2023)
by: Baskar, A., et al.
Published: (2023)
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
by: Zhang, Chen, et al.
Published: (2024)
by: Zhang, Chen, et al.
Published: (2024)
The AffectToolbox: Affect Analysis for Everyone
by: Mertes, Silvan, et al.
Published: (2024)
by: Mertes, Silvan, et al.
Published: (2024)
Explaining AI Without Code: A User Study on Explainable AI
by: Abarca, Natalia, et al.
Published: (2025)
by: Abarca, Natalia, et al.
Published: (2025)
Investigating an Intelligent System to Monitor \& Explain Abnormal Activity Patterns of Older Adults
by: Lee, Min Hun, et al.
Published: (2025)
by: Lee, Min Hun, et al.
Published: (2025)
Multilingual Dyadic Interaction Corpus NoXi+J: Toward Understanding Asian-European Non-verbal Cultural Characteristics and their Influences on Engagement
by: Funk, Marius, et al.
Published: (2024)
by: Funk, Marius, et al.
Published: (2024)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty
by: Aryan, Prakash, et al.
Published: (2026)
by: Aryan, Prakash, et al.
Published: (2026)
Towards Interactive Reinforcement Learning with Intrinsic Feedback
by: Poole, Benjamin, et al.
Published: (2021)
by: Poole, Benjamin, et al.
Published: (2021)
Abstracted Trajectory Visualization for Explainability in Reinforcement Learning
by: Takagi, Yoshiki, et al.
Published: (2024)
by: Takagi, Yoshiki, et al.
Published: (2024)
LLMs as Policy-Agnostic Teammates: A Case Study in Human Proxy Design for Heterogeneous Agent Teams
by: Justus, Aju Ani, et al.
Published: (2025)
by: Justus, Aju Ani, et al.
Published: (2025)
Learning to Assist Humans without Inferring Rewards
by: Myers, Vivek, et al.
Published: (2024)
by: Myers, Vivek, et al.
Published: (2024)
Why is "Problems" Predictive of Positive Sentiment? A Case Study of Explaining Unintuitive Features in Sentiment Classification
by: Qu, Jiaming, et al.
Published: (2024)
by: Qu, Jiaming, et al.
Published: (2024)
Cost-Effective Online Multi-LLM Selection with Versatile Reward Models
by: Dai, Xiangxiang, et al.
Published: (2024)
by: Dai, Xiangxiang, et al.
Published: (2024)
Explanation through Reward Model Reconciliation using POMDP Tree Search
by: Kraske, Benjamin D., et al.
Published: (2023)
by: Kraske, Benjamin D., et al.
Published: (2023)
Self-Initiated Open World Learning for Autonomous AI Agents
by: Liu, Bing, et al.
Published: (2021)
by: Liu, Bing, et al.
Published: (2021)
Consistency Based Weakly Self-Supervised Learning for Human Activity Recognition with Wearables
by: Sheng, Taoran, et al.
Published: (2024)
by: Sheng, Taoran, et al.
Published: (2024)
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024)
by: Zhou, Xinglin, et al.
Published: (2024)
Reinforcement Learning Driven Generalizable Feature Representation for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2025)
by: Ye, Xiaozhou, et al.
Published: (2025)
AURA: A Reinforcement Learning Framework for AI-Driven Adaptive Conversational Surveys
by: Tang, Jinwen, et al.
Published: (2025)
by: Tang, Jinwen, et al.
Published: (2025)
LLMs for XAI: Future Directions for Explaining Explanations
by: Zytek, Alexandra, et al.
Published: (2024)
by: Zytek, Alexandra, et al.
Published: (2024)
Coprocessor Actor Critic: A Model-Based Reinforcement Learning Approach For Adaptive Brain Stimulation
by: Pan, Michelle, et al.
Published: (2024)
by: Pan, Michelle, et al.
Published: (2024)
Classroom Simulacra: Building Contextual Student Generative Agents in Online Education for Learning Behavioral Simulation
by: Xu, Songlin, et al.
Published: (2025)
by: Xu, Songlin, et al.
Published: (2025)
HARP: Human-Assisted Regrouping with Permutation Invariant Critic for Multi-Agent Reinforcement Learning
by: Hu, Huawen, et al.
Published: (2024)
by: Hu, Huawen, et al.
Published: (2024)
Learning Personalized Decision Support Policies
by: Bhatt, Umang, et al.
Published: (2023)
by: Bhatt, Umang, et al.
Published: (2023)
ClickAgent: Enhancing UI Location Capabilities of Autonomous Agents
by: Hoscilowicz, Jakub, et al.
Published: (2024)
by: Hoscilowicz, Jakub, et al.
Published: (2024)
Reinforcement Learning for Personalized Dialogue Management
by: Hengst, Floris den, et al.
Published: (2019)
by: Hengst, Floris den, et al.
Published: (2019)
Off-Policy Selection for Initiating Human-Centric Experimental Design
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
Interaction Dynamics as a Reward Signal for LLMs
by: Gooding, Sian, et al.
Published: (2025)
by: Gooding, Sian, et al.
Published: (2025)
iScore: Visual Analytics for Interpreting How Language Models Automatically Score Summaries
by: Coscia, Adam, et al.
Published: (2024)
by: Coscia, Adam, et al.
Published: (2024)
End-to-end Optimization of Belief and Policy Learning in Shared Autonomy Paradigms
by: Farhadi, MH, et al.
Published: (2026)
by: Farhadi, MH, et al.
Published: (2026)
Agent AI: Surveying the Horizons of Multimodal Interaction
by: Durante, Zane, et al.
Published: (2024)
by: Durante, Zane, et al.
Published: (2024)
Quantifying the Effect of Feedback Frequency in Interactive Reinforcement Learning for Robotic Tasks
by: Harnack, Daniel, et al.
Published: (2022)
by: Harnack, Daniel, et al.
Published: (2022)
AI Agents for Inventory Control: Human-LLM-OR Complementarity
by: Baek, Jackie, et al.
Published: (2026)
by: Baek, Jackie, et al.
Published: (2026)
More Than Irrational: Modeling Belief-Biased Agents
by: Zhu, Yifan, et al.
Published: (2025)
by: Zhu, Yifan, et al.
Published: (2025)
Robots That Know What to Ask: Recovering Misaligned Rewards through Targeted Explanations
by: Merker, Helena, et al.
Published: (2026)
by: Merker, Helena, et al.
Published: (2026)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
by: Dennler, Nathaniel, et al.
Published: (2024)
by: Dennler, Nathaniel, et al.
Published: (2024)
Reinforcement Learning-Enhanced Procedural Generation for Dynamic Narrative-Driven AR Experiences
by: Joshi, Aniruddha Srinivas
Published: (2025)
by: Joshi, Aniruddha Srinivas
Published: (2025)
Similar Items
-
"Trust me on this" Explaining Agent Behavior to a Human Terminator
by: Menkes, Uri, et al.
Published: (2025) -
A Model for Intelligible Interaction Between Agents That Predict and Explain
by: Baskar, A., et al.
Published: (2023) -
Hindsight PRIORs for Reward Learning from Human Preferences
by: Verma, Mudit, et al.
Published: (2024) -
Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
by: Zhang, Chen, et al.
Published: (2024) -
The AffectToolbox: Affect Analysis for Everyone
by: Mertes, Silvan, et al.
Published: (2024)