Randomized algorithms and PAC bounds for inverse reinforcement learning in continuous spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Kamoutsi, Angeliki, Schmitt-Förster, Peter, Sutter, Tobias, Cevher, Volkan, Lygeros, John |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024)
by: Schmitt-Förster, Peter, et al.
Published: (2024)
Stable Nonconvex-Nonconcave Training via Linear Interpolation
by: Pethick, Thomas, et al.
Published: (2023)
by: Pethick, Thomas, et al.
Published: (2023)
Efficient Continual Finite-Sum Minimization
by: Mavrothalassitis, Ioannis, et al.
Published: (2024)
by: Mavrothalassitis, Ioannis, et al.
Published: (2024)
Adversarial Training Should Be Cast as a Non-Zero-Sum Game
by: Robey, Alexander, et al.
Published: (2023)
by: Robey, Alexander, et al.
Published: (2023)
Taming "data-hungry" reinforcement learning? Stability in continuous state-action spaces
by: Duan, Yaqi, et al.
Published: (2024)
by: Duan, Yaqi, et al.
Published: (2024)
Krylov Cubic Regularized Newton: A Subspace Second-Order Method with Dimension-Free Convergence Rate
by: Jiang, Ruichen, et al.
Published: (2024)
by: Jiang, Ruichen, et al.
Published: (2024)
A Two-Timescale Primal-Dual Framework for Reinforcement Learning via Online Dual Variable Guidance
by: Wolter, Axel Friedrich, et al.
Published: (2025)
by: Wolter, Axel Friedrich, et al.
Published: (2025)
On the Role of Batch Size in Stochastic Conditional Gradient Methods
by: Islamov, Rustem, et al.
Published: (2026)
by: Islamov, Rustem, et al.
Published: (2026)
Constrained Stochastic Spectral Preconditioning Converges for Nonconvex Objectives
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
by: Oikonomidis, Konstantinos, et al.
Published: (2026)
Training Deep Learning Models with Norm-Constrained LMOs
by: Pethick, Thomas, et al.
Published: (2025)
by: Pethick, Thomas, et al.
Published: (2025)
Stochastic Halpern iteration in normed spaces and applications to reinforcement learning
by: Bravo, Mario, et al.
Published: (2024)
by: Bravo, Mario, et al.
Published: (2024)
Distributional Adversarial Attacks and Training in Deep Hedging
by: He, Guangyi, et al.
Published: (2025)
by: He, Guangyi, et al.
Published: (2025)
MADA: Meta-Adaptive Optimizers through hyper-gradient Descent
by: Ozkara, Kaan, et al.
Published: (2024)
by: Ozkara, Kaan, et al.
Published: (2024)
Joint Chance Constrained Optimal Control via Linear Programming
by: Schmid, Niklas, et al.
Published: (2024)
by: Schmid, Niklas, et al.
Published: (2024)
Layer-wise Quantization for Quantized Optimistic Dual Averaging
by: Nguyen, Anh Duc, et al.
Published: (2025)
by: Nguyen, Anh Duc, et al.
Published: (2025)
Finite Sample Frequency Domain Identification
by: Tsiamis, Anastasios, et al.
Published: (2024)
by: Tsiamis, Anastasios, et al.
Published: (2024)
Learning to Remove Cuts in Integer Linear Programming
by: Puigdemont, Pol, et al.
Published: (2024)
by: Puigdemont, Pol, et al.
Published: (2024)
Learning-to-Optimize with PAC-Bayesian Guarantees: Theoretical Considerations and Practical Implementation
by: Sucker, Michael, et al.
Published: (2024)
by: Sucker, Michael, et al.
Published: (2024)
What is the objective of reasoning with reinforcement learning?
by: Davis, Damek, et al.
Published: (2025)
by: Davis, Damek, et al.
Published: (2025)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Safe Time-Varying Optimization based on Gaussian Processes with Spatio-Temporal Kernel
by: Li, Jialin, et al.
Published: (2024)
by: Li, Jialin, et al.
Published: (2024)
Contractivity and linear convergence in bilinear saddle-point problems: An operator-theoretic approach
by: Dirren, Colin, et al.
Published: (2024)
by: Dirren, Colin, et al.
Published: (2024)
Online Residual Learning from Offline Experts for Pedestrian Tracking
by: Vlachos, Anastasios, et al.
Published: (2024)
by: Vlachos, Anastasios, et al.
Published: (2024)
Predictive Linear Online Tracking for Unknown Targets
by: Tsiamis, Anastasios, et al.
Published: (2024)
by: Tsiamis, Anastasios, et al.
Published: (2024)
PAC-Bayes Meets Online Contextual Optimization
by: Xie, Zhuojun, et al.
Published: (2025)
by: Xie, Zhuojun, et al.
Published: (2025)
Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise
by: Frikha, Noufel, et al.
Published: (2024)
by: Frikha, Noufel, et al.
Published: (2024)
Towards a Systems Theory of Algorithms
by: Dörfler, Florian, et al.
Published: (2024)
by: Dörfler, Florian, et al.
Published: (2024)
Towards Optimal Offline Reinforcement Learning
by: Li, Mengmeng, et al.
Published: (2025)
by: Li, Mengmeng, et al.
Published: (2025)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
by: Gao, Xuefeng, et al.
Published: (2022)
by: Gao, Xuefeng, et al.
Published: (2022)
Flowsheet synthesis through hierarchical reinforcement learning and graph neural networks
by: Stops, Laura, et al.
Published: (2022)
by: Stops, Laura, et al.
Published: (2022)
Apprenticeship learning with prior beliefs using inverse optimization
by: Junca, Mauricio, et al.
Published: (2025)
by: Junca, Mauricio, et al.
Published: (2025)
Online estimation of the inverse of the Hessian for stochastic optimization with application to universal stochastic Newton algorithms
by: Godichon-Baggioni, Antoine, et al.
Published: (2024)
by: Godichon-Baggioni, Antoine, et al.
Published: (2024)
Online reinforcement learning via sparse Gaussian mixture model Q-functions
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
A reinforcement learning agent for maintenance of deteriorating systems with increasingly imperfect repairs
by: Marugán, Alberto Pliego, et al.
Published: (2025)
by: Marugán, Alberto Pliego, et al.
Published: (2025)
PAC Learnability of Scenario Decision-Making Algorithms: Necessary Conditions and Sufficient Conditions
by: Berger, Guillaume O., et al.
Published: (2025)
by: Berger, Guillaume O., et al.
Published: (2025)
Computing Optimal Joint Chance Constrained Control Policies
by: Schmid, Niklas, et al.
Published: (2023)
by: Schmid, Niklas, et al.
Published: (2023)
CLASP: An online learning algorithm for Convex Losses And Squared Penalties
by: Ferreira, Ricardo N., et al.
Published: (2026)
by: Ferreira, Ricardo N., et al.
Published: (2026)
A PAC-Bayes Approach for Controlling Unknown Linear Discrete-time Systems
by: Luo, Yujia, et al.
Published: (2026)
by: Luo, Yujia, et al.
Published: (2026)
Leveraging Hamilton-Jacobi PDEs with time-dependent Hamiltonians for continual scientific machine learning
by: Chen, Paula, et al.
Published: (2023)
by: Chen, Paula, et al.
Published: (2023)
Independent policy gradient-based reinforcement learning for economic and reliable energy management of multi-microgrid systems
by: Hu, Junkai, et al.
Published: (2025)
by: Hu, Junkai, et al.
Published: (2025)
Similar Items
-
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024) -
Stable Nonconvex-Nonconcave Training via Linear Interpolation
by: Pethick, Thomas, et al.
Published: (2023) -
Efficient Continual Finite-Sum Minimization
by: Mavrothalassitis, Ioannis, et al.
Published: (2024) -
Adversarial Training Should Be Cast as a Non-Zero-Sum Game
by: Robey, Alexander, et al.
Published: (2023) -
Taming "data-hungry" reinforcement learning? Stability in continuous state-action spaces
by: Duan, Yaqi, et al.
Published: (2024)