Near-Optimal Second-Order Guarantees for Model-Based Adversarial Imitation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Shangzhe, Zhou, Dongruo, Zhang, Weitong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning
by: Li, Shangzhe, et al.
Published: (2026)
by: Li, Shangzhe, et al.
Published: (2026)
Imitation from Observations with Trajectory-Level Generative Embeddings
by: Qu, Yongtao, et al.
Published: (2026)
by: Qu, Yongtao, et al.
Published: (2026)
Provably Efficient Offline-to-Online Value Adaptation with General Function Approximation
by: Li, Shangzhe, et al.
Published: (2026)
by: Li, Shangzhe, et al.
Published: (2026)
Reward-free World Models for Online Imitation Learning
by: Li, Shangzhe, et al.
Published: (2024)
by: Li, Shangzhe, et al.
Published: (2024)
Nearly Minimax Optimal Regret for Learning Linear Mixture Stochastic Shortest Path
by: Di, Qiwei, et al.
Published: (2024)
by: Di, Qiwei, et al.
Published: (2024)
Model-based RL as a Minimalist Approach to Horizon-Free and Second-Order Bounds
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
by: Zhang, Junkai, et al.
Published: (2024)
by: Zhang, Junkai, et al.
Published: (2024)
Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning
by: Li, Shangzhe, et al.
Published: (2025)
by: Li, Shangzhe, et al.
Published: (2025)
Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
by: Zhang, Weitong, et al.
Published: (2021)
by: Zhang, Weitong, et al.
Published: (2021)
Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees
by: Chen, Yilei, et al.
Published: (2024)
by: Chen, Yilei, et al.
Published: (2024)
Near-Optimal Distributed Minimax Optimization under the Second-Order Similarity
by: Zhou, Qihao, et al.
Published: (2024)
by: Zhou, Qihao, et al.
Published: (2024)
Augmenting Offline Reinforcement Learning with State-only Interactions
by: Li, Shangzhe, et al.
Published: (2024)
by: Li, Shangzhe, et al.
Published: (2024)
Auto-Encoding Adversarial Imitation Learning
by: Zhang, Kaifeng, et al.
Published: (2022)
by: Zhang, Kaifeng, et al.
Published: (2022)
Breaking the $\log(1/Δ_2)$ Barrier: Better Batched Best Arm Identification with Adaptive Grids
by: Jin, Tianyuan, et al.
Published: (2025)
by: Jin, Tianyuan, et al.
Published: (2025)
PAGAR: Taming Reward Misalignment in Inverse Reinforcement Learning-Based Imitation Learning with Protagonist Antagonist Guided Adversarial Reward
by: Zhou, Weichao, et al.
Published: (2023)
by: Zhou, Weichao, et al.
Published: (2023)
Near-Optimal Dynamic Regret for Adversarial Linear Mixture MDPs
by: Li, Long-Fei, et al.
Published: (2024)
by: Li, Long-Fei, et al.
Published: (2024)
Personalized Federated Training of Diffusion Models with Privacy Guarantees
by: Patel, Kumar Kshitij, et al.
Published: (2025)
by: Patel, Kumar Kshitij, et al.
Published: (2025)
Near-Optimal Regret in Adversarial Kernel Bandits
by: Zhang, Yu-Jie, et al.
Published: (2026)
by: Zhang, Yu-Jie, et al.
Published: (2026)
Latent Wasserstein Adversarial Imitation Learning
by: Yang, Siqi, et al.
Published: (2026)
by: Yang, Siqi, et al.
Published: (2026)
Uncovering Capabilities of Model Pruning in Graph Contrastive Learning
by: Wu, Junran, et al.
Published: (2024)
by: Wu, Junran, et al.
Published: (2024)
Rethinking Adversarial Inverse Reinforcement Learning: Policy Imitation, Transferable Reward Recovery and Algebraic Equilibrium Proof
by: Zhang, Yangchun, et al.
Published: (2024)
by: Zhang, Yangchun, et al.
Published: (2024)
Nearly-Optimal Algorithm for Adversarial Kernelized Bandits
by: Iwazaki, Shogo
Published: (2026)
by: Iwazaki, Shogo
Published: (2026)
Adversarial Imitation Learning via Boosting
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
Sample-efficient Adversarial Imitation Learning
by: Jung, Dahuin, et al.
Published: (2023)
by: Jung, Dahuin, et al.
Published: (2023)
Is Your Imitation Learning Policy Better than Mine? Policy Comparison with Near-Optimal Stopping
by: Snyder, David, et al.
Published: (2025)
by: Snyder, David, et al.
Published: (2025)
Near Optimal Decision Trees in a SPLIT Second
by: Babbar, Varun, et al.
Published: (2025)
by: Babbar, Varun, et al.
Published: (2025)
Imitation Learning of MPC with Neural Networks: Error Guarantees and Sparsification
by: Alsmeier, Hendrik, et al.
Published: (2025)
by: Alsmeier, Hendrik, et al.
Published: (2025)
Instance-Dependent Continuous-Time Reinforcement Learning via Maximum Likelihood Estimation
by: Zhao, Runze, et al.
Published: (2025)
by: Zhao, Runze, et al.
Published: (2025)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Provable Zero-Shot Generalization in Offline Reinforcement Learning
by: Wang, Zhiyong, et al.
Published: (2025)
by: Wang, Zhiyong, et al.
Published: (2025)
Provably and Practically Efficient Adversarial Imitation Learning with General Function Approximation
by: Xu, Tian, et al.
Published: (2024)
by: Xu, Tian, et al.
Published: (2024)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
On the Limits of Test-Time Compute: Sequential Reward Filtering for Better Inference
by: Yu, Yue, et al.
Published: (2025)
by: Yu, Yue, et al.
Published: (2025)
S2O: Enhancing Adversarial Training with Second-Order Statistics of Weights
by: Jin, Gaojie, et al.
Published: (2026)
by: Jin, Gaojie, et al.
Published: (2026)
Quantile Q-Learning: Revisiting Offline Extreme Q-Learning with Quantile Regression
by: Gao, Xinming, et al.
Published: (2025)
by: Gao, Xinming, et al.
Published: (2025)
Model-Based Learning of Near-Optimal Finite-Window Policies in POMDPs
by: Jordan, Philip, et al.
Published: (2026)
by: Jordan, Philip, et al.
Published: (2026)
Near-Optimal Regret for Distributed Adversarial Bandits: A Black-Box Approach
by: Qiu, Hao, et al.
Published: (2026)
by: Qiu, Hao, et al.
Published: (2026)
Adversarial Imitation Learning with General Function Approximation: Theoretical Analysis and Practical Algorithms
by: Xu, Tian, et al.
Published: (2026)
by: Xu, Tian, et al.
Published: (2026)
SD2AIL: Adversarial Imitation Learning from Synthetic Demonstrations via Diffusion Models
by: Li, Pengcheng, et al.
Published: (2025)
by: Li, Pengcheng, et al.
Published: (2025)
Similar Items
-
Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning
by: Li, Shangzhe, et al.
Published: (2026) -
Imitation from Observations with Trajectory-Level Generative Embeddings
by: Qu, Yongtao, et al.
Published: (2026) -
Provably Efficient Offline-to-Online Value Adaptation with General Function Approximation
by: Li, Shangzhe, et al.
Published: (2026) -
Reward-free World Models for Online Imitation Learning
by: Li, Shangzhe, et al.
Published: (2024) -
Nearly Minimax Optimal Regret for Learning Linear Mixture Stochastic Shortest Path
by: Di, Qiwei, et al.
Published: (2024)