Inverse Reinforcement Learning without Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Swamy, Gokul, Choudhury, Sanjiban, Bagnell, J. Andrew, Wu, Zhiwei Steven |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
Hybrid Inverse Reinforcement Learning
by: Ren, Juntao, et al.
Published: (2024)
by: Ren, Juntao, et al.
Published: (2024)
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
by: Swamy, Gokul, et al.
Published: (2025)
by: Swamy, Gokul, et al.
Published: (2025)
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
A Minimaximalist Approach to Reinforcement Learning from Human Feedback
by: Swamy, Gokul, et al.
Published: (2024)
by: Swamy, Gokul, et al.
Published: (2024)
Efficient Imitation under Misspecification
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
by: Espinosa-Dice, Nicolas, et al.
Published: (2025)
Hybrid Reinforcement Learning from Offline Observation Alone
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
Multi-Agent Imitation Learning: Value is Easy, Regret is Hard
by: Tang, Jingwu, et al.
Published: (2024)
by: Tang, Jingwu, et al.
Published: (2024)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
by: Jain, Arnav Kumar, et al.
Published: (2024)
by: Jain, Arnav Kumar, et al.
Published: (2024)
REBEL: Reinforcement Learning via Regressing Relative Rewards
by: Gao, Zhaolin, et al.
Published: (2024)
by: Gao, Zhaolin, et al.
Published: (2024)
The Importance of Online Data: Understanding Preference Fine-tuning via Coverage
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
To Distill or Decide? Understanding the Algorithmic Trade-off in Partially Observable Reinforcement Learning
by: Song, Yuda, et al.
Published: (2025)
by: Song, Yuda, et al.
Published: (2025)
A Smooth Sea Never Made a Skilled SAILOR: Robust Imitation via Learning to Search
by: Jain, Arnav Kumar, et al.
Published: (2025)
by: Jain, Arnav Kumar, et al.
Published: (2025)
Process Reward Models for LLM Agents: Practical Framework and Directions
by: Choudhury, Sanjiban
Published: (2025)
by: Choudhury, Sanjiban
Published: (2025)
Aligning LLMs with Domain Invariant Reward Models
by: Wu, David, et al.
Published: (2025)
by: Wu, David, et al.
Published: (2025)
Gained in Translation: Privileged Pairwise Judges Enhance Multilingual Reasoning
by: Sutawika, Lintang, et al.
Published: (2026)
by: Sutawika, Lintang, et al.
Published: (2026)
Expanding the Capabilities of Reinforcement Learning via Text Feedback
by: Song, Yuda, et al.
Published: (2026)
by: Song, Yuda, et al.
Published: (2026)
Distributional Inverse Reinforcement Learning
by: Wu, Feiyang, et al.
Published: (2025)
by: Wu, Feiyang, et al.
Published: (2025)
Better than Your Teacher: LLM Agents that learn from Privileged AI Feedback
by: Choudhury, Sanjiban, et al.
Published: (2024)
by: Choudhury, Sanjiban, et al.
Published: (2024)
Imitation Learning from a Single Temporally Misaligned Video
by: Huey, William, et al.
Published: (2025)
by: Huey, William, et al.
Published: (2025)
Back to Blackwell: Closing the Loop on Intransitivity in Multi-Objective Preference Fine-Tuning
by: Zhang, Jiahao, et al.
Published: (2026)
by: Zhang, Jiahao, et al.
Published: (2026)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
by: Wu, Runzhe, et al.
Published: (2024)
by: Wu, Runzhe, et al.
Published: (2024)
Kernel Density Bayesian Inverse Reinforcement Learning
by: Mandyam, Aishwarya, et al.
Published: (2023)
by: Mandyam, Aishwarya, et al.
Published: (2023)
Vulnerability Analysis of Safe Reinforcement Learning via Inverse Constrained Reinforcement Learning
by: Fan, Jialiang, et al.
Published: (2026)
by: Fan, Jialiang, et al.
Published: (2026)
Inverse Delayed Reinforcement Learning
by: Zhan, Simon Sinong, et al.
Published: (2024)
by: Zhan, Simon Sinong, et al.
Published: (2024)
Towards Generalized Inverse Reinforcement Learning
by: Dong, Chaosheng, et al.
Published: (2024)
by: Dong, Chaosheng, et al.
Published: (2024)
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
by: Zhong, Huiying, et al.
Published: (2024)
by: Zhong, Huiying, et al.
Published: (2024)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
by: Sestini, Alessandro, et al.
Published: (2025)
by: Sestini, Alessandro, et al.
Published: (2025)
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set Approach
by: Kim, Kihyun, et al.
Published: (2026)
by: Kim, Kihyun, et al.
Published: (2026)
Quantifying the Sensitivity of Inverse Reinforcement Learning to Misspecification
by: Skalse, Joar, et al.
Published: (2024)
by: Skalse, Joar, et al.
Published: (2024)
Inverse Reinforcement Learning with Multiple Planning Horizons
by: Yao, Jiayu, et al.
Published: (2024)
by: Yao, Jiayu, et al.
Published: (2024)
Walking the Values in Bayesian Inverse Reinforcement Learning
by: Bajgar, Ondrej, et al.
Published: (2024)
by: Bajgar, Ondrej, et al.
Published: (2024)
Confidence Aware Inverse Constrained Reinforcement Learning
by: Subramanian, Sriram Ganapathi, et al.
Published: (2024)
by: Subramanian, Sriram Ganapathi, et al.
Published: (2024)
Towards Interpretable Deep Reinforcement Learning Models via Inverse Reinforcement Learning
by: Xie, Sean, et al.
Published: (2022)
by: Xie, Sean, et al.
Published: (2022)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Inverse Reinforcement Learning With Constraint Recovery
by: Das, Nirjhar, et al.
Published: (2023)
by: Das, Nirjhar, et al.
Published: (2023)
Fast Rates for Inverse Reinforcement Learning
by: Schlaginhaufen, Andreas, et al.
Published: (2026)
by: Schlaginhaufen, Andreas, et al.
Published: (2026)
Environment Design for Inverse Reinforcement Learning
by: Buening, Thomas Kleine, et al.
Published: (2022)
by: Buening, Thomas Kleine, et al.
Published: (2022)
Recursive Deep Inverse Reinforcement Learning
by: Ghanem, Paul, et al.
Published: (2025)
by: Ghanem, Paul, et al.
Published: (2025)
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
Similar Items
-
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024) -
Hybrid Inverse Reinforcement Learning
by: Ren, Juntao, et al.
Published: (2024) -
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
by: Swamy, Gokul, et al.
Published: (2025) -
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
by: Wu, David, et al.
Published: (2024) -
A Minimaximalist Approach to Reinforcement Learning from Human Feedback
by: Swamy, Gokul, et al.
Published: (2024)