Adaptive Inverse Reinforcement Learning with Online Off-Policy Data Collection
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yibei, Cao, Yuexin, Liu, Zhixin, Xie, Lihua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Inverse Continuous-Time Linear Quadratic Regulator: From Control Cost Matrix to Entire Cost Reconstruction
by: Cao, Yuexin, et al.
Published: (2025)
by: Cao, Yuexin, et al.
Published: (2025)
Pattern formation using an intrinsic optimal control approach
by: Li, Tianhao, et al.
Published: (2025)
by: Li, Tianhao, et al.
Published: (2025)
A Differential Dynamic Programming Framework for Inverse Reinforcement Learning
by: Cao, Kun, et al.
Published: (2024)
by: Cao, Kun, et al.
Published: (2024)
Adaptive Mitigation of Insider Threats via Off-Policy Learning
by: Xu, Gehui, et al.
Published: (2026)
by: Xu, Gehui, et al.
Published: (2026)
Element-based Formation Control: a Unified Perspective from Continuum Mechanics
by: Cao, Kun, et al.
Published: (2026)
by: Cao, Kun, et al.
Published: (2026)
Adaptive Optimal Control of Linear Periodic Systems: An Off-Policy Value Iteration Approach
by: Pang, Bo, et al.
Published: (2019)
by: Pang, Bo, et al.
Published: (2019)
TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization
by: Sorokin, D., et al.
Published: (2023)
by: Sorokin, D., et al.
Published: (2023)
Data-Enabled Policy Optimization for Direct Adaptive Learning of the LQR
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Reinforcement Learning for Inverse Linear-quadratic Dynamic Non-cooperative Games
by: Martirosyan, Emin, et al.
Published: (2023)
by: Martirosyan, Emin, et al.
Published: (2023)
A differential game approach to intrinsic encirclement control
by: Zhou, Panpan, et al.
Published: (2025)
by: Zhou, Panpan, et al.
Published: (2025)
Reinforcement Learning for optimal dividend problem under diffusion model
by: Bai, Lihua, et al.
Published: (2023)
by: Bai, Lihua, et al.
Published: (2023)
Reinforcement Learning for Inverse Non-Cooperative Linear-Quadratic Output-feedback Differential Games
by: Martirosyan, Emin, et al.
Published: (2024)
by: Martirosyan, Emin, et al.
Published: (2024)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
by: Li, Wenye, et al.
Published: (2025)
by: Li, Wenye, et al.
Published: (2025)
Online Policy Optimization in Unknown Nonlinear Systems
by: Lin, Yiheng, et al.
Published: (2024)
by: Lin, Yiheng, et al.
Published: (2024)
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025)
by: Dimanidis, Ioannis, et al.
Published: (2025)
Learning Optimal and Fair Policies for Online Allocation of Scarce Societal Resources from Data Collected in Deployment
by: Tang, Bill, et al.
Published: (2023)
by: Tang, Bill, et al.
Published: (2023)
Offline Hierarchical Reinforcement Learning via Inverse Optimization
by: Schmidt, Carolin, et al.
Published: (2024)
by: Schmidt, Carolin, et al.
Published: (2024)
Adaptive Federated Learning to Optimize Integrated Flows in Cyber-Physical Data Centers
by: Liu, Junhong, et al.
Published: (2025)
by: Liu, Junhong, et al.
Published: (2025)
Data-Driven Stochastic Distribution System Hardening Based on Bayesian Online Learning
by: Shi, Wenlong, et al.
Published: (2025)
by: Shi, Wenlong, et al.
Published: (2025)
Policy Iteration Reinforcement Learning Method for Continuous-Time Linear-Quadratic Mean-Field Control Problems
by: Li, Na, et al.
Published: (2023)
by: Li, Na, et al.
Published: (2023)
Distributed adaptive estimation for stochastic large regression models
by: Gan, Die, et al.
Published: (2026)
by: Gan, Die, et al.
Published: (2026)
Online Policy Selection for Inventory Problems
by: Hihat, Massil, et al.
Published: (2024)
by: Hihat, Massil, et al.
Published: (2024)
Fine-tuning for Data-enabled Predictive Control of Noisy Systems by Reinforcement Learning
by: Wang, Jinbao, et al.
Published: (2025)
by: Wang, Jinbao, et al.
Published: (2025)
An Adaptive Data-Enabled Policy Optimization Approach for Autonomous Bicycle Control
by: Persson, Niklas, et al.
Published: (2025)
by: Persson, Niklas, et al.
Published: (2025)
Long-Run Conditional Value-at-Risk Reinforcement Learning
by: Wang, Qixin, et al.
Published: (2026)
by: Wang, Qixin, et al.
Published: (2026)
On some perturbation properties of nonsmooth optimization on Riemannian manifolds with applications
by: Zhou, Yuexin, et al.
Published: (2023)
by: Zhou, Yuexin, et al.
Published: (2023)
On the robust isolated calmness of a class of nonsmooth optimizations on Riemannian manifolds and its applications
by: Bao, Chenglong, et al.
Published: (2022)
by: Bao, Chenglong, et al.
Published: (2022)
Distributed Optimal Control and Application to Consensus of Multi-Agent Systems
by: Zhang, Liping, et al.
Published: (2023)
by: Zhang, Liping, et al.
Published: (2023)
Double Duality: Variational Primal-Dual Policy Optimization for Constrained Reinforcement Learning
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Fast Sampling for Linear Inverse Problems of Vectors and Tensors using Multilinear Extensions
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
Refined Bounds on Near Optimality Finite Window Policies in POMDPs and Their Reinforcement Learning
by: Demirci, Yunus Emre, et al.
Published: (2024)
by: Demirci, Yunus Emre, et al.
Published: (2024)
Machine Learning for Inverse Problems and Data Assimilation
by: Bach, Eviatar, et al.
Published: (2024)
by: Bach, Eviatar, et al.
Published: (2024)
Performance guaranteed MPC Policy Approximation via Cost Guided Learning
by: Zhou, Chenchen, et al.
Published: (2026)
by: Zhou, Chenchen, et al.
Published: (2026)
Solving the Pod Repositioning Problem with Deep Reinforced Adaptive Large Neighborhood Search
by: Xie, Lin, et al.
Published: (2025)
by: Xie, Lin, et al.
Published: (2025)
Adaptive Online Optimization for Microgrids with Renewable Energy Sources
by: van Weerelt, Wouter J. A., et al.
Published: (2025)
by: van Weerelt, Wouter J. A., et al.
Published: (2025)
Linear Dynamics meets Linear MDPs: Closed-Form Optimal Policies via Reinforcement Learning
by: Makdah, Abed AlRahman Al, et al.
Published: (2025)
by: Makdah, Abed AlRahman Al, et al.
Published: (2025)
Stochastic MPC with Online-optimized Policies and Closed-loop Guarantees
by: Bartos, Marcell, et al.
Published: (2025)
by: Bartos, Marcell, et al.
Published: (2025)
Offline and Online Nonlinear Inverse Differential Games with Known and Approximated Cost and Value Function Structures
by: Karg, Philipp, et al.
Published: (2024)
by: Karg, Philipp, et al.
Published: (2024)
Direct Adaptive Control of Grid-Connected Power Converters via Output-Feedback Data-Enabled Policy Optimization
by: Zhao, Feiran, et al.
Published: (2024)
by: Zhao, Feiran, et al.
Published: (2024)
Data-Enabled Policy and Value Iteration for Continuous-Time Linear Quadratic Output Feedback Control
by: Xie, Jun, et al.
Published: (2026)
by: Xie, Jun, et al.
Published: (2026)
Similar Items
-
Inverse Continuous-Time Linear Quadratic Regulator: From Control Cost Matrix to Entire Cost Reconstruction
by: Cao, Yuexin, et al.
Published: (2025) -
Pattern formation using an intrinsic optimal control approach
by: Li, Tianhao, et al.
Published: (2025) -
A Differential Dynamic Programming Framework for Inverse Reinforcement Learning
by: Cao, Kun, et al.
Published: (2024) -
Adaptive Mitigation of Insider Threats via Off-Policy Learning
by: Xu, Gehui, et al.
Published: (2026) -
Element-based Formation Control: a Unified Perspective from Continuum Mechanics
by: Cao, Kun, et al.
Published: (2026)