Adversarial Imitation Learning via Boosting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chang, Jonathan D., Sreenivas, Dhruv, Huang, Yingbing, Brantley, Kianté, Sun, Wen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Expressive Value Learning for Scalable Offline Reinforcement Learning
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
von: Oertell, Owen, et al.
Veröffentlicht: (2024)
A Surprising Failure? Multimodal LLMs and the NLVR Challenge
von: Wu, Anne, et al.
Veröffentlicht: (2024)
von: Wu, Anne, et al.
Veröffentlicht: (2024)
Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF
von: Gao, Zhaolin, et al.
Veröffentlicht: (2024)
von: Gao, Zhaolin, et al.
Veröffentlicht: (2024)
Dataset Reset Policy Optimization for RLHF
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
Policy-Gradient Training of Language Models for Ranking
von: Gao, Ge, et al.
Veröffentlicht: (2023)
von: Gao, Ge, et al.
Veröffentlicht: (2023)
Value-Guided Search for Efficient Chain-of-Thought Reasoning
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wang, Kaiwen, et al.
Veröffentlicht: (2025)
LLMs Are In-Context Bandit Reinforcement Learners
von: Monea, Giovanni, et al.
Veröffentlicht: (2024)
von: Monea, Giovanni, et al.
Veröffentlicht: (2024)
Accelerating RL for LLM Reasoning with Optimal Advantage Regression
von: Brantley, Kianté, et al.
Veröffentlicht: (2025)
von: Brantley, Kianté, et al.
Veröffentlicht: (2025)
Diffusing States and Matching Scores: A New Framework for Imitation Learning
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
von: Wu, Runzhe, et al.
Veröffentlicht: (2024)
$Q\sharp$: Provably Optimal Distributional RL for LLM Post-Training
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2025)
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2025)
Scaling Offline RL via Efficient and Expressive Shortcut Models
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
Diffusion-Reward Adversarial Imitation Learning
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
von: Lai, Chun-Mao, et al.
Veröffentlicht: (2024)
LLMs Can Learn to Reason Via Off-Policy RL
von: Ritter, Daniel, et al.
Veröffentlicht: (2026)
von: Ritter, Daniel, et al.
Veröffentlicht: (2026)
Sample-efficient Adversarial Imitation Learning
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
Scaling Laws for Imitation Learning in Single-Agent Games
von: Tuyls, Jens, et al.
Veröffentlicht: (2023)
von: Tuyls, Jens, et al.
Veröffentlicht: (2023)
Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
Efficient Imitation under Misspecification
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization
von: Ye, Yuyang, et al.
Veröffentlicht: (2024)
von: Ye, Yuyang, et al.
Veröffentlicht: (2024)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
Imitation Learning via Focused Satisficing
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
Boolean Satisfiability via Imitation Learning
von: Zhang, Zewei, et al.
Veröffentlicht: (2025)
von: Zhang, Zewei, et al.
Veröffentlicht: (2025)
Denoising-based Contractive Imitation Learning
von: Shen, Macheng, et al.
Veröffentlicht: (2025)
von: Shen, Macheng, et al.
Veröffentlicht: (2025)
DIDA: Denoised Imitation Learning based on Domain Adaptation
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
Reinforcement Learning via Implicit Imitation Guidance
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
Mitigating Adversarial Perturbations for Deep Reinforcement Learning via Vector Quantization
von: Luu, Tung M., et al.
Veröffentlicht: (2024)
von: Luu, Tung M., et al.
Veröffentlicht: (2024)
Confounded Causal Imitation Learning with Instrumental Variables
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
DecompGAIL: Learning Realistic Traffic Behaviors with Decomposed Multi-Agent Generative Adversarial Imitation Learning
von: Guo, Ke, et al.
Veröffentlicht: (2025)
von: Guo, Ke, et al.
Veröffentlicht: (2025)
Imitation Learning from Observation through Optimal Transport
von: Chang, Wei-Di, et al.
Veröffentlicht: (2023)
von: Chang, Wei-Di, et al.
Veröffentlicht: (2023)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
von: Fan, Jiangdong, et al.
Veröffentlicht: (2024)
Imitation Bootstrapped Reinforcement Learning
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
Quantifying Generalisation in Imitation Learning
von: Gavenski, Nathan, et al.
Veröffentlicht: (2025)
von: Gavenski, Nathan, et al.
Veröffentlicht: (2025)
Ctx2TrajGen: Traffic Context-Aware Microscale Vehicle Trajectories using Generative Adversarial Imitation Learning
von: Jin, Joobin, et al.
Veröffentlicht: (2025)
von: Jin, Joobin, et al.
Veröffentlicht: (2025)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
von: Zhao, Eric, et al.
Veröffentlicht: (2025)
Offline Imitation of Badminton Player Behavior via Experiential Contexts and Brownian Motion
von: Wang, Kuang-Da, et al.
Veröffentlicht: (2024)
von: Wang, Kuang-Da, et al.
Veröffentlicht: (2024)
Align Your Intents: Offline Imitation Learning via Optimal Transport
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
von: Bobrin, Maksim, et al.
Veröffentlicht: (2024)
Cross-Domain Imitation Learning via Optimal Transport
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2021)
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2021)
SENSOR: Imitate Third-Person Expert's Behaviors via Active Sensoring
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
von: Huang, Kaichen, et al.
Veröffentlicht: (2024)
Hierarchical Generative Adversarial Imitation Learning with Mid-level Input Generation for Autonomous Driving on Urban Environments
von: Couto, Gustavo Claudio Karl, et al.
Veröffentlicht: (2023)
von: Couto, Gustavo Claudio Karl, et al.
Veröffentlicht: (2023)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
von: Huang, Kevin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Expressive Value Learning for Scalable Offline Reinforcement Learning
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025) -
RL for Consistency Models: Faster Reward Guided Text-to-Image Generation
von: Oertell, Owen, et al.
Veröffentlicht: (2024) -
A Surprising Failure? Multimodal LLMs and the NLVR Challenge
von: Wu, Anne, et al.
Veröffentlicht: (2024) -
Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF
von: Gao, Zhaolin, et al.
Veröffentlicht: (2024) -
Dataset Reset Policy Optimization for RLHF
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)