A Bayesian Approach to Robust Inverse Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Ran, Zeng, Siliang, Li, Chenliang, Garcia, Alfredo, McDonald, Anthony, Hong, Mingyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Demonstrations Meet Generative World Models: A Maximum Likelihood Framework for Offline Inverse Reinforcement Learning
von: Zeng, Siliang, et al.
Veröffentlicht: (2023)
von: Zeng, Siliang, et al.
Veröffentlicht: (2023)
Understanding Inverse Reinforcement Learning under Overparameterization: Non-Asymptotic Analysis and Global Optimality
von: Zhang, Ruijia, et al.
Veröffentlicht: (2025)
von: Zhang, Ruijia, et al.
Veröffentlicht: (2025)
Aligning Frozen LLMs by Reinforcement Learning: An Iterative Reweight-then-Optimize Approach
von: Zhang, Xinnan, et al.
Veröffentlicht: (2025)
von: Zhang, Xinnan, et al.
Veröffentlicht: (2025)
A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
von: Wei, Ran, et al.
Veröffentlicht: (2023)
von: Wei, Ran, et al.
Veröffentlicht: (2023)
Structural Estimation of Markov Decision Processes in High-Dimensional State Space with Finite-Time Guarantees
von: Zeng, Siliang, et al.
Veröffentlicht: (2022)
von: Zeng, Siliang, et al.
Veröffentlicht: (2022)
Reinforcing Multi-Turn Reasoning in LLM Agents via Turn-Level Reward Design
von: Wei, Quan, et al.
Veröffentlicht: (2025)
von: Wei, Quan, et al.
Veröffentlicht: (2025)
Learning An Active Inference Model of Driver Perception and Control: Application to Vehicle Car-Following
von: Wei, Ran, et al.
Veröffentlicht: (2023)
von: Wei, Ran, et al.
Veröffentlicht: (2023)
Stabilizing Off-Policy Training for Long-Horizon LLM Agent via Turn-Level Importance Sampling and Clipping-Triggered Normalization
von: Li, Chenliang, et al.
Veröffentlicht: (2025)
von: Li, Chenliang, et al.
Veröffentlicht: (2025)
Deep Latent Force Models: ODE-based Process Convolutions for Bayesian Deep Learning
von: Baldwin-McDonald, Thomas, et al.
Veröffentlicht: (2023)
von: Baldwin-McDonald, Thomas, et al.
Veröffentlicht: (2023)
Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
von: Li, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Li, Jiaxiang, et al.
Veröffentlicht: (2024)
From Demonstrations to Rewards: Alignment Without Explicit Human Preferences
von: Zeng, Siliang, et al.
Veröffentlicht: (2025)
von: Zeng, Siliang, et al.
Veröffentlicht: (2025)
Walking the Values in Bayesian Inverse Reinforcement Learning
von: Bajgar, Ondrej, et al.
Veröffentlicht: (2024)
von: Bajgar, Ondrej, et al.
Veröffentlicht: (2024)
Kernel Density Bayesian Inverse Reinforcement Learning
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
ADARL: Adaptive Low-Rank Structures for Robust Policy Learning under Uncertainty
von: Li, Chenliang, et al.
Veröffentlicht: (2025)
von: Li, Chenliang, et al.
Veröffentlicht: (2025)
Bayesian Inverse Reinforcement Learning for Non-Markovian Rewards
von: Topper, Noah, et al.
Veröffentlicht: (2024)
von: Topper, Noah, et al.
Veröffentlicht: (2024)
PAC Apprenticeship Learning with Bayesian Active Inverse Reinforcement Learning
von: Bajgar, Ondrej, et al.
Veröffentlicht: (2025)
von: Bajgar, Ondrej, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
von: Han, Minghao, et al.
Veröffentlicht: (2026)
von: Han, Minghao, et al.
Veröffentlicht: (2026)
Learning Reward and Policy Jointly from Demonstration and Preference Improves Alignment
von: Li, Chenliang, et al.
Veröffentlicht: (2024)
von: Li, Chenliang, et al.
Veröffentlicht: (2024)
HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents
von: Peng, Jiangweizhi, et al.
Veröffentlicht: (2026)
von: Peng, Jiangweizhi, et al.
Veröffentlicht: (2026)
Autonomous Assessment of Demonstration Sufficiency via Bayesian Inverse Reinforcement Learning
von: Trinh, Tu, et al.
Veröffentlicht: (2022)
von: Trinh, Tu, et al.
Veröffentlicht: (2022)
Eliciting Risk Aversion with Inverse Reinforcement Learning via Interactive Questioning
von: Cheng, Ziteng, et al.
Veröffentlicht: (2023)
von: Cheng, Ziteng, et al.
Veröffentlicht: (2023)
A Model Selection Approach for Corruption Robust Reinforcement Learning
von: Wei, Chen-Yu, et al.
Veröffentlicht: (2021)
von: Wei, Chen-Yu, et al.
Veröffentlicht: (2021)
Learning Explainable Dense Reward Shapes via Bayesian Optimization
von: Koo, Ryan, et al.
Veröffentlicht: (2025)
von: Koo, Ryan, et al.
Veröffentlicht: (2025)
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment
von: Cai, Yuang, et al.
Veröffentlicht: (2024)
von: Cai, Yuang, et al.
Veröffentlicht: (2024)
A Curriculum Learning Approach to Reinforcement Learning: Leveraging RAG for Multimodal Question Answering
von: Zhang, Chenliang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenliang, et al.
Veröffentlicht: (2025)
Get RICH or Die Scaling: Profitably Trading Inference Compute for Robustness
von: McDonald, Tavish, et al.
Veröffentlicht: (2025)
von: McDonald, Tavish, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning without Reinforcement Learning
von: Swamy, Gokul, et al.
Veröffentlicht: (2023)
von: Swamy, Gokul, et al.
Veröffentlicht: (2023)
Reward Transfer from Inverse Reinforcement Learning: A Coupled Minimax Approach
von: Hao, Guang-Yuan, et al.
Veröffentlicht: (2026)
von: Hao, Guang-Yuan, et al.
Veröffentlicht: (2026)
Real-Time Stress Monitoring, Detection, and Management in College Students: A Wearable Technology and Machine-Learning Approach
von: Ta, Alan, et al.
Veröffentlicht: (2025)
von: Ta, Alan, et al.
Veröffentlicht: (2025)
Distributional Inverse Reinforcement Learning
von: Wu, Feiyang, et al.
Veröffentlicht: (2025)
von: Wu, Feiyang, et al.
Veröffentlicht: (2025)
Log-Concave Coupling for Sampling Neural Net Posteriors
von: McDonald, Curtis, et al.
Veröffentlicht: (2024)
von: McDonald, Curtis, et al.
Veröffentlicht: (2024)
Robot Arm Control via Cognitive Map Learners
von: McDonald, Nathan, et al.
Veröffentlicht: (2026)
von: McDonald, Nathan, et al.
Veröffentlicht: (2026)
Towards Generalized Inverse Reinforcement Learning
von: Dong, Chaosheng, et al.
Veröffentlicht: (2024)
von: Dong, Chaosheng, et al.
Veröffentlicht: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
von: Wu, David, et al.
Veröffentlicht: (2024)
von: Wu, David, et al.
Veröffentlicht: (2024)
Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates
von: Glentis, Athanasios, et al.
Veröffentlicht: (2026)
von: Glentis, Athanasios, et al.
Veröffentlicht: (2026)
Robust Reinforcement Learning from Corrupted Human Feedback
von: Bukharin, Alexander, et al.
Veröffentlicht: (2024)
von: Bukharin, Alexander, et al.
Veröffentlicht: (2024)
Hybrid Inverse Reinforcement Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
Zero-shot and Few-shot Generation Strategies for Artificial Clinical Records
von: Frayling, Erlend, et al.
Veröffentlicht: (2024)
von: Frayling, Erlend, et al.
Veröffentlicht: (2024)
Towards an Adaptable and Generalizable Optimization Engine in Decision and Control: A Meta Reinforcement Learning Approach
von: Yang, Sungwook, et al.
Veröffentlicht: (2024)
von: Yang, Sungwook, et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
von: Kim, Kihyun, et al.
Veröffentlicht: (2026)
von: Kim, Kihyun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
When Demonstrations Meet Generative World Models: A Maximum Likelihood Framework for Offline Inverse Reinforcement Learning
von: Zeng, Siliang, et al.
Veröffentlicht: (2023) -
Understanding Inverse Reinforcement Learning under Overparameterization: Non-Asymptotic Analysis and Global Optimality
von: Zhang, Ruijia, et al.
Veröffentlicht: (2025) -
Aligning Frozen LLMs by Reinforcement Learning: An Iterative Reweight-then-Optimize Approach
von: Zhang, Xinnan, et al.
Veröffentlicht: (2025) -
A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
von: Wei, Ran, et al.
Veröffentlicht: (2023) -
Structural Estimation of Markov Decision Processes in High-Dimensional State Space with Finite-Time Guarantees
von: Zeng, Siliang, et al.
Veröffentlicht: (2022)