Partial Identifiability and Misspecification in Inverse Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Skalse, Joar, Abate, Alessandro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Partial Identifiability in Inverse Reinforcement Learning For Agents With Non-Exponential Discounting
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
Quantifying the Sensitivity of Inverse Reinforcement Learning to Misspecification
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
von: Skalse, Joar, et al.
Veröffentlicht: (2024)
The Perils of Optimizing Learned Reward Functions: Low Training Error Does Not Guarantee Low Regret
von: Fluri, Lukas, et al.
Veröffentlicht: (2024)
von: Fluri, Lukas, et al.
Veröffentlicht: (2024)
STARC: A General Framework For Quantifying Differences Between Reward Functions
von: Skalse, Joar, et al.
Veröffentlicht: (2023)
von: Skalse, Joar, et al.
Veröffentlicht: (2023)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2026)
DeepLTL: Learning to Efficiently Satisfy Complex LTL Specifications for Multi-Task RL
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2024)
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2024)
Efficient Solution and Learning of Robust Factored MDPs
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
von: Schnitzer, Yannik, et al.
Veröffentlicht: (2025)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
von: Sestini, Alessandro, et al.
Veröffentlicht: (2025)
Semantically Labelled Automata for Multi-Task Reinforcement Learning with LTL Instructions
von: Abate, Alessandro, et al.
Veröffentlicht: (2026)
von: Abate, Alessandro, et al.
Veröffentlicht: (2026)
Temporal-Difference Variational Continual Learning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Efficient Imitation under Misspecification
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning
von: van der Vaart, Pascal R., et al.
Veröffentlicht: (2025)
von: van der Vaart, Pascal R., et al.
Veröffentlicht: (2025)
On The Expressivity of Objective-Specification Formalisms in Reinforcement Learning
von: Subramani, Rohan, et al.
Veröffentlicht: (2023)
von: Subramani, Rohan, et al.
Veröffentlicht: (2023)
Hybrid Inverse Reinforcement Learning
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
von: Ren, Juntao, et al.
Veröffentlicht: (2024)
Zero-Shot Instruction Following in RL via Structured LTL Representations
von: Giuri, Mattia, et al.
Veröffentlicht: (2025)
von: Giuri, Mattia, et al.
Veröffentlicht: (2025)
Robust Shielding for Safe Reinforcement Learning
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
von: Court, Edwin Hamel-De le, et al.
Veröffentlicht: (2026)
Fast Rates for Inverse Reinforcement Learning
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2026)
von: Schlaginhaufen, Andreas, et al.
Veröffentlicht: (2026)
On the Effective Horizon of Inverse Reinforcement Learning
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
Environment Design for Inverse Reinforcement Learning
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2022)
Recursive Deep Inverse Reinforcement Learning
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
von: Ghanem, Paul, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning With Constraint Recovery
von: Das, Nirjhar, et al.
Veröffentlicht: (2023)
von: Das, Nirjhar, et al.
Veröffentlicht: (2023)
Neural Proofs for Sound Verification and Control of Complex Systems
von: Abate, Alessandro
Veröffentlicht: (2025)
von: Abate, Alessandro
Veröffentlicht: (2025)
Uncertainty Representations in State-Space Layers for Deep Reinforcement Learning under Partial Observability
von: Luis, Carlos E., et al.
Veröffentlicht: (2024)
von: Luis, Carlos E., et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning with Sub-optimal Experts
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Inverse Reinforcement Learning by Estimating Expertise of Demonstrators
von: Beliaev, Mark, et al.
Veröffentlicht: (2024)
von: Beliaev, Mark, et al.
Veröffentlicht: (2024)
Is Optimal Transport Necessary for Inverse Reinforcement Learning?
von: Dong, Zixuan, et al.
Veröffentlicht: (2025)
von: Dong, Zixuan, et al.
Veröffentlicht: (2025)
Kernel Density Bayesian Inverse Reinforcement Learning
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
von: Mandyam, Aishwarya, et al.
Veröffentlicht: (2023)
Towards Interpretable Deep Reinforcement Learning Models via Inverse Reinforcement Learning
von: Xie, Sean, et al.
Veröffentlicht: (2022)
von: Xie, Sean, et al.
Veröffentlicht: (2022)
Near-Optimal Partially Observable Reinforcement Learning with Partial Online State Information
von: Shi, Ming, et al.
Veröffentlicht: (2023)
von: Shi, Ming, et al.
Veröffentlicht: (2023)
Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
von: Mechergui, Malek, et al.
Veröffentlicht: (2024)
Is Inverse Reinforcement Learning Harder than Standard Reinforcement Learning? A Theoretical Perspective
von: Zhao, Lei, et al.
Veröffentlicht: (2023)
von: Zhao, Lei, et al.
Veröffentlicht: (2023)
Inverse Reinforcement Learning from Non-Stationary Learning Agents
von: Sivakumar, Kavinayan P., et al.
Veröffentlicht: (2024)
von: Sivakumar, Kavinayan P., et al.
Veröffentlicht: (2024)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
von: Yue, Bo, et al.
Veröffentlicht: (2024)
von: Yue, Bo, et al.
Veröffentlicht: (2024)
Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols
von: Griffin, Charlie, et al.
Veröffentlicht: (2024)
von: Griffin, Charlie, et al.
Veröffentlicht: (2024)
Zero-Shot Instruction Following in RL via Structured LTL Representations
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2026)
von: Jackermeier, Mathias, et al.
Veröffentlicht: (2026)
Zero-Shot Reinforcement Learning Under Partial Observability
von: Jeen, Scott, et al.
Veröffentlicht: (2025)
von: Jeen, Scott, et al.
Veröffentlicht: (2025)
Inverse Delayed Reinforcement Learning
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
von: Zhan, Simon Sinong, et al.
Veröffentlicht: (2024)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Partial Identifiability in Inverse Reinforcement Learning For Agents With Non-Exponential Discounting
von: Skalse, Joar, et al.
Veröffentlicht: (2024) -
Quantifying the Sensitivity of Inverse Reinforcement Learning to Misspecification
von: Skalse, Joar, et al.
Veröffentlicht: (2024) -
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks
von: Skalse, Joar, et al.
Veröffentlicht: (2024) -
The Perils of Optimizing Learned Reward Functions: Low Training Error Does Not Guarantee Low Regret
von: Fluri, Lukas, et al.
Veröffentlicht: (2024) -
STARC: A General Framework For Quantifying Differences Between Reward Functions
von: Skalse, Joar, et al.
Veröffentlicht: (2023)