Understanding Adversarial Imitation Learning in Small Sample Regime: A Stage-coupled Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Tian, Li, Ziniu, Yu, Yang, Luo, Zhi-Quan |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Non-Adversarial Imitation Learning Provably Free of Compounding Errors: The Role of Bellman Constraints
by: Xu, Tian, et al.
Published: (2026)
by: Xu, Tian, et al.
Published: (2026)
ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
by: Li, Ziniu, et al.
Published: (2023)
by: Li, Ziniu, et al.
Published: (2023)
Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms
by: Xu, Tian, et al.
Published: (2026)
by: Xu, Tian, et al.
Published: (2026)
Provably and Practically Efficient Adversarial Imitation Learning with General Function Approximation
by: Xu, Tian, et al.
Published: (2024)
by: Xu, Tian, et al.
Published: (2024)
Policy Optimization in RLHF: The Impact of Out-of-preference Data
by: Li, Ziniu, et al.
Published: (2023)
by: Li, Ziniu, et al.
Published: (2023)
Sample-efficient Adversarial Imitation Learning
by: Jung, Dahuin, et al.
Published: (2023)
by: Jung, Dahuin, et al.
Published: (2023)
A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms
by: Li, Yingru, et al.
Published: (2025)
by: Li, Yingru, et al.
Published: (2025)
Latent Wasserstein Adversarial Imitation Learning
by: Yang, Siqi, et al.
Published: (2026)
by: Yang, Siqi, et al.
Published: (2026)
Why Transformers Need Adam: A Hessian Perspective
by: Zhang, Yushun, et al.
Published: (2024)
by: Zhang, Yushun, et al.
Published: (2024)
Auto-Encoding Adversarial Imitation Learning
by: Zhang, Kaifeng, et al.
Published: (2022)
by: Zhang, Kaifeng, et al.
Published: (2022)
Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning
by: Li, Shangzhe, et al.
Published: (2026)
by: Li, Shangzhe, et al.
Published: (2026)
Preserving Diversity in Supervised Fine-Tuning of Large Language Models
by: Li, Ziniu, et al.
Published: (2024)
by: Li, Ziniu, et al.
Published: (2024)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
Uniformly Stable Algorithms for Adversarial Training and Beyond
by: Xiao, Jiancong, et al.
Published: (2024)
by: Xiao, Jiancong, et al.
Published: (2024)
Adam-mini: Use Fewer Learning Rates To Gain More
by: Zhang, Yushun, et al.
Published: (2024)
by: Zhang, Yushun, et al.
Published: (2024)
Adversarial Imitation Learning via Boosting
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation
by: Li, Ziniu, et al.
Published: (2025)
by: Li, Ziniu, et al.
Published: (2025)
C-GAIL: Stabilizing Generative Adversarial Imitation Learning with Control Theory
by: Luo, Tianjiao, et al.
Published: (2024)
by: Luo, Tianjiao, et al.
Published: (2024)
Q-Star Meets Scalable Posterior Sampling: Bridging Theory and Practice via HyperAgent
by: Li, Yingru, et al.
Published: (2024)
by: Li, Yingru, et al.
Published: (2024)
Adversarial Robustness in One-Stage Learning-to-Defer
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
Near-Optimal Second-Order Guarantees for Model-Based Adversarial Imitation Learning
by: Li, Shangzhe, et al.
Published: (2025)
by: Li, Shangzhe, et al.
Published: (2025)
SD2AIL: Adversarial Imitation Learning from Synthetic Demonstrations via Diffusion Models
by: Li, Pengcheng, et al.
Published: (2025)
by: Li, Pengcheng, et al.
Published: (2025)
On Agnostic PAC Learning in the Small Error Regime
by: Asilis, Julian, et al.
Published: (2025)
by: Asilis, Julian, et al.
Published: (2025)
PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization
by: Ye, Yuyang, et al.
Published: (2024)
by: Ye, Yuyang, et al.
Published: (2024)
Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees
by: Chen, Yilei, et al.
Published: (2024)
by: Chen, Yilei, et al.
Published: (2024)
On the Benefits of Inducing Local Lipschitzness for Robust Generative Adversarial Imitation Learning
by: Memarian, Farzan, et al.
Published: (2021)
by: Memarian, Farzan, et al.
Published: (2021)
Interpretable Imitation Learning via Generative Adversarial STL Inference and Control
by: Liu, Wenliang, et al.
Published: (2024)
by: Liu, Wenliang, et al.
Published: (2024)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
by: Li, Yingru, et al.
Published: (2024)
by: Li, Yingru, et al.
Published: (2024)
Bridging Distributional and Risk-sensitive Reinforcement Learning with Provable Regret Bounds
by: Liang, Hao, et al.
Published: (2022)
by: Liang, Hao, et al.
Published: (2022)
Tackling Small Sample Survival Analysis via Transfer Learning: A Study of Colorectal Cancer Prognosis
by: Zhao, Yonghao, et al.
Published: (2025)
by: Zhao, Yonghao, et al.
Published: (2025)
PAGAR: Taming Reward Misalignment in Inverse Reinforcement Learning-Based Imitation Learning with Protagonist Antagonist Guided Adversarial Reward
by: Zhou, Weichao, et al.
Published: (2023)
by: Zhou, Weichao, et al.
Published: (2023)
Ranking-based Client Selection with Imitation Learning for Efficient Federated Learning
by: Tian, Chunlin, et al.
Published: (2024)
by: Tian, Chunlin, et al.
Published: (2024)
Rethinking Adversarial Inverse Reinforcement Learning: Policy Imitation, Transferable Reward Recovery and Algebraic Equilibrium Proof
by: Zhang, Yangchun, et al.
Published: (2024)
by: Zhang, Yangchun, et al.
Published: (2024)
FinFlowRL: An Imitation-Reinforcement Learning Framework for Adaptive Stochastic Control in Finance
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning
by: Morin, Sacha, et al.
Published: (2026)
by: Morin, Sacha, et al.
Published: (2026)
Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression
by: Zhang, Ziqi, et al.
Published: (2024)
by: Zhang, Ziqi, et al.
Published: (2024)
Adversarial Imitation Learning from Visual Observations using Latent Information
by: Giammarino, Vittorio, et al.
Published: (2023)
by: Giammarino, Vittorio, et al.
Published: (2023)
Bridging Formal Language with Chain-of-Thought Reasoning to Geometry Problem Solving
by: Yang, Tianyun, et al.
Published: (2025)
by: Yang, Tianyun, et al.
Published: (2025)
Learning to Importance Sample in Primary Sample Space
by: Zheng, Quan, et al.
Published: (2018)
by: Zheng, Quan, et al.
Published: (2018)
Boolean Satisfiability via Imitation Learning
by: Zhang, Zewei, et al.
Published: (2025)
by: Zhang, Zewei, et al.
Published: (2025)
Similar Items
-
Non-Adversarial Imitation Learning Provably Free of Compounding Errors: The Role of Bellman Constraints
by: Xu, Tian, et al.
Published: (2026) -
ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
by: Li, Ziniu, et al.
Published: (2023) -
Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms
by: Xu, Tian, et al.
Published: (2026) -
Provably and Practically Efficient Adversarial Imitation Learning with General Function Approximation
by: Xu, Tian, et al.
Published: (2024) -
Policy Optimization in RLHF: The Impact of Out-of-preference Data
by: Li, Ziniu, et al.
Published: (2023)