Learning from Ambiguous Demonstrations with Self-Explanation Guided Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zha, Yantian, Guan, Lin, Kambhampati, Subbarao |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
by: Bhambri, Siddhant, et al.
Published: (2024)
by: Bhambri, Siddhant, et al.
Published: (2024)
Can Large Language Models Reason and Plan?
by: Kambhampati, Subbarao
Published: (2024)
by: Kambhampati, Subbarao
Published: (2024)
Mind The Gap: Quantifying Mechanistic Gaps in Algorithmic Reasoning via Neural Compilation
by: Saldyt, Lucas, et al.
Published: (2025)
by: Saldyt, Lucas, et al.
Published: (2025)
NatSGD: A Dataset with Speech, Gestures, and Demonstrations for Robot Learning in Natural Human-Robot Interaction
by: Shrestha, Snehesh, et al.
Published: (2024)
by: Shrestha, Snehesh, et al.
Published: (2024)
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
by: Guan, Lin, et al.
Published: (2024)
by: Guan, Lin, et al.
Published: (2024)
SORREL: Suboptimal-Demonstration-Guided Reinforcement Learning for Learning to Branch
by: Feng, Shengyu, et al.
Published: (2024)
by: Feng, Shengyu, et al.
Published: (2024)
Demonstration-Guided Continual Reinforcement Learning in Dynamic Environments
by: Yang, Xue, et al.
Published: (2025)
by: Yang, Xue, et al.
Published: (2025)
Demonstration Guided Multi-Objective Reinforcement Learning
by: Lu, Junlin, et al.
Published: (2024)
by: Lu, Junlin, et al.
Published: (2024)
ExPO: Unlocking Hard Reasoning with Self-Explanation-Guided Reinforcement Learning
by: Zhou, Ruiyang, et al.
Published: (2025)
by: Zhou, Ruiyang, et al.
Published: (2025)
CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries
by: Mu, Ni, et al.
Published: (2025)
by: Mu, Ni, et al.
Published: (2025)
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
by: Valmeekam, Karthik, et al.
Published: (2025)
by: Valmeekam, Karthik, et al.
Published: (2025)
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
by: Kambhampati, Subbarao, et al.
Published: (2024)
by: Kambhampati, Subbarao, et al.
Published: (2024)
Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains
by: Samineni, Soumya Rani, et al.
Published: (2025)
by: Samineni, Soumya Rani, et al.
Published: (2025)
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
by: Samineni, Soumya Rani, et al.
Published: (2025)
by: Samineni, Soumya Rani, et al.
Published: (2025)
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
by: Bhambri, Siddhant, et al.
Published: (2023)
by: Bhambri, Siddhant, et al.
Published: (2023)
Learning from Ambiguous Data with Hard Labels
by: Xie, Zeke, et al.
Published: (2025)
by: Xie, Zeke, et al.
Published: (2025)
Can large language models reason and plan?
by: Subbarao Kambhampati
Published: (2024)
by: Subbarao Kambhampati
Published: (2024)
Approximating Shapley Explanations in Reinforcement Learning
by: Beechey, Daniel, et al.
Published: (2025)
by: Beechey, Daniel, et al.
Published: (2025)
Evaluating the False Trust Engendered by LLM Explanations
by: Palod, Vardhan, et al.
Published: (2026)
by: Palod, Vardhan, et al.
Published: (2026)
On Learning for Ambiguous Chance Constrained Problems
by: Madhusudanarao, A Ch, et al.
Published: (2023)
by: Madhusudanarao, A Ch, et al.
Published: (2023)
Ambiguous Online Learning
by: Kosoy, Vanessa
Published: (2025)
by: Kosoy, Vanessa
Published: (2025)
XStacking: Explanation-Guided Stacked Ensemble Learning
by: Garouani, Moncef, et al.
Published: (2025)
by: Garouani, Moncef, et al.
Published: (2025)
Towards Robust Incremental Learning under Ambiguous Supervision
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
Fast Lifelong Adaptive Inverse Reinforcement Learning from Demonstrations
by: Chen, Letian, et al.
Published: (2022)
by: Chen, Letian, et al.
Published: (2022)
Inverse Reinforcement Learning by Estimating Expertise of Demonstrators
by: Beliaev, Mark, et al.
Published: (2024)
by: Beliaev, Mark, et al.
Published: (2024)
Counterfactual Explanations for Continuous Action Reinforcement Learning
by: Dong, Shuyang, et al.
Published: (2025)
by: Dong, Shuyang, et al.
Published: (2025)
Skill-Enhanced Reinforcement Learning Acceleration from Heterogeneous Demonstrations
by: Zhang, Hanping, et al.
Published: (2024)
by: Zhang, Hanping, et al.
Published: (2024)
A Tutorial: An Intuitive Explanation of Offline Reinforcement Learning Theory
by: Che, Fengdi
Published: (2025)
by: Che, Fengdi
Published: (2025)
Investigating ECG Diagnosis with Ambiguous Labels using Partial Label Learning
by: Rahmani, Sana, et al.
Published: (2025)
by: Rahmani, Sana, et al.
Published: (2025)
Towards Generalizable Reinforcement Learning via Causality-Guided Self-Adaptive Representations
by: Yang, Yupei, et al.
Published: (2024)
by: Yang, Yupei, et al.
Published: (2024)
Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning
by: Trinh, Tu, et al.
Published: (2022)
by: Trinh, Tu, et al.
Published: (2022)
LLM-Driven Stationarity-Aware Expert Demonstrations for Multi-Agent Reinforcement Learning in Mobile Systems
by: Duan, Tianyang, et al.
Published: (2025)
by: Duan, Tianyang, et al.
Published: (2025)
Why the Agent Made that Decision: Contrastive Explanation Learning for Reinforcement Learning
by: Zuo, Rui, et al.
Published: (2024)
by: Zuo, Rui, et al.
Published: (2024)
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026)
by: Rajapakse, Dilina, et al.
Published: (2026)
Machine Learning from Explanations
by: Tao, Jiashu, et al.
Published: (2025)
by: Tao, Jiashu, et al.
Published: (2025)
MEGL: Multimodal Explanation-Guided Learning
by: Zhang, Yifei, et al.
Published: (2024)
by: Zhang, Yifei, et al.
Published: (2024)
Rainbow-DemoRL: Combining Improvements in Demonstration-Augmented Reinforcement Learning
by: Bhatt, Dwait, et al.
Published: (2026)
by: Bhatt, Dwait, et al.
Published: (2026)
Imitation Learning from Purified Demonstrations
by: Wang, Yunke, et al.
Published: (2023)
by: Wang, Yunke, et al.
Published: (2023)
Reverse Forward Curriculum Learning for Extreme Sample and Demonstration Efficiency in Reinforcement Learning
by: Tao, Stone, et al.
Published: (2024)
by: Tao, Stone, et al.
Published: (2024)
Learning from Random Demonstrations: Offline Reinforcement Learning with Importance-Sampled Diffusion Models
by: Fang, Zeyu, et al.
Published: (2024)
by: Fang, Zeyu, et al.
Published: (2024)
Similar Items
-
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
by: Bhambri, Siddhant, et al.
Published: (2024) -
Can Large Language Models Reason and Plan?
by: Kambhampati, Subbarao
Published: (2024) -
Mind The Gap: Quantifying Mechanistic Gaps in Algorithmic Reasoning via Neural Compilation
by: Saldyt, Lucas, et al.
Published: (2025) -
NatSGD: A Dataset with Speech, Gestures, and Demonstrations for Robot Learning in Natural Human-Robot Interaction
by: Shrestha, Snehesh, et al.
Published: (2024) -
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
by: Guan, Lin, et al.
Published: (2024)