Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
Fuente:
arXiv
Saved in:
| Main Authors: | Boveiri, Mohammad, Esfahani, Peyman Mohajerin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accelerating Adaptive Systems via Normalized Parameter Estimation Laws
by: Boveiri, Mohammad, et al.
Published: (2025)
by: Boveiri, Mohammad, et al.
Published: (2025)
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025)
by: Dimanidis, Ioannis, et al.
Published: (2025)
Scalable Kernel Inverse Optimization
by: Long, Youyuan, et al.
Published: (2024)
by: Long, Youyuan, et al.
Published: (2024)
Distributionally Robust Model Predictive Control: Closed-loop Guarantees and Scalable Algorithms
by: McAllister, Robert D., et al.
Published: (2023)
by: McAllister, Robert D., et al.
Published: (2023)
Learning in Inverse Optimization: Incenter Cost, Augmented Suboptimality Loss, and Algorithms
by: Scroccaro, Pedro Zattoni, et al.
Published: (2023)
by: Scroccaro, Pedro Zattoni, et al.
Published: (2023)
Nonlinear Distributionally Robust Optimization
by: Sheriff, Mohammed Rayyan, et al.
Published: (2023)
by: Sheriff, Mohammed Rayyan, et al.
Published: (2023)
Fast Algorithm for Constrained Linear Inverse Problems
by: Sheriff, Mohammed Rayyan, et al.
Published: (2022)
by: Sheriff, Mohammed Rayyan, et al.
Published: (2022)
A Saddle Point Algorithm for Robust Data-Driven Factor Model Problems
by: Khodakaramzadeh, Shabnam, et al.
Published: (2025)
by: Khodakaramzadeh, Shabnam, et al.
Published: (2025)
Efficient Uncertainty Propagation with Guarantees in Wasserstein Distance
by: Figueiredo, Eduardo, et al.
Published: (2025)
by: Figueiredo, Eduardo, et al.
Published: (2025)
Robust Multivariate Detection and Estimation with Fault Frequency Content Information
by: Dong, Jingwei, et al.
Published: (2023)
by: Dong, Jingwei, et al.
Published: (2023)
Tight Generalization Bounds for Noiseless Inverse Optimization
by: Fatemi, Pouria, et al.
Published: (2026)
by: Fatemi, Pouria, et al.
Published: (2026)
A Variance-Reduced Stochastic Gradient Tracking Algorithm for Decentralized Optimization with Orthogonality Constraints
by: Wang, Lei, et al.
Published: (2022)
by: Wang, Lei, et al.
Published: (2022)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Inverse Optimization for Routing Problems
by: Scroccaro, Pedro Zattoni, et al.
Published: (2023)
by: Scroccaro, Pedro Zattoni, et al.
Published: (2023)
Wasserstein Distributionally Robust Optimization: Theory and Applications in Machine Learning
by: Kuhn, Daniel, et al.
Published: (2019)
by: Kuhn, Daniel, et al.
Published: (2019)
Rank-One Modified Value Iteration
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
Optimal Bayesian Affine Estimator and Active Learning for the Wiener Model
by: Vakili, Sasan, et al.
Published: (2025)
by: Vakili, Sasan, et al.
Published: (2025)
A Hybrid Algorithm for Monotone Variational Inequalities
by: Baghbadorani, Reza Rahimi, et al.
Published: (2026)
by: Baghbadorani, Reza Rahimi, et al.
Published: (2026)
User-centric Vehicle-to-Grid Optimization with an Input Convex Neural Network-based Battery Degradation Model
by: Mallick, Arghya, et al.
Published: (2025)
by: Mallick, Arghya, et al.
Published: (2025)
A Frank-Wolfe Algorithm for Strongly Monotone Variational Inequalities
by: Baghbadorani, Reza Rahimi, et al.
Published: (2025)
by: Baghbadorani, Reza Rahimi, et al.
Published: (2025)
Convergence and stability of Q-learning in Hierarchical Reinforcement Learning
by: Manenti, Massimiliano, et al.
Published: (2025)
by: Manenti, Massimiliano, et al.
Published: (2025)
Inverse Optimization via Learning Feasible Regions
by: Ren, Ke, et al.
Published: (2025)
by: Ren, Ke, et al.
Published: (2025)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
Gradient Estimation and Variance Reduction in Stochastic and Deterministic Models
by: Keane, Ronan
Published: (2024)
by: Keane, Ronan
Published: (2024)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
Estimation Sample Complexity of a Class of Nonlinear Continuous-time Systems
by: Kuang, Simon, et al.
Published: (2023)
by: Kuang, Simon, et al.
Published: (2023)
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
by: Muehlebach, Michael, et al.
Published: (2025)
by: Muehlebach, Michael, et al.
Published: (2025)
Identification and Adaptive Control of Markov Jump Systems: Sample Complexity and Regret Bounds
by: Sattar, Yahya, et al.
Published: (2021)
by: Sattar, Yahya, et al.
Published: (2021)
Information Theoretically Optimal Sample Complexity of Learning Dynamical Directed Acyclic Graphs
by: Veedu, Mishfad Shaikh, et al.
Published: (2023)
by: Veedu, Mishfad Shaikh, et al.
Published: (2023)
Intersection of Reinforcement Learning and Bayesian Optimization for Intelligent Control of Industrial Processes: A Safe MPC-based DPG using Multi-Objective BO
by: Esfahani, Hossein Nejatbakhsh, et al.
Published: (2025)
by: Esfahani, Hossein Nejatbakhsh, et al.
Published: (2025)
A New Lineserach for Accelerated Composite Minimization
by: Baghbadorani, Reza Rahimi, et al.
Published: (2024)
by: Baghbadorani, Reza Rahimi, et al.
Published: (2024)
Locally Linear Convergence for Nonsmooth Convex Optimization via Coupled Smoothing and Momentum
by: Baghbadorani, Reza Rahimi, et al.
Published: (2025)
by: Baghbadorani, Reza Rahimi, et al.
Published: (2025)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Control and Reinforcement Learning through the Lens of Optimization: An Algorithmic Perspective
by: Ok, Tolga, et al.
Published: (2025)
by: Ok, Tolga, et al.
Published: (2025)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
by: Maity, Sreejeet, et al.
Published: (2025)
by: Maity, Sreejeet, et al.
Published: (2025)
InterQ: A DQN Framework for Optimal Intermittent Control
by: Aggarwal, Shubham, et al.
Published: (2025)
by: Aggarwal, Shubham, et al.
Published: (2025)
Optimal Transportation by Orthogonal Coupling Dynamics
by: Sadr, Mohsen, et al.
Published: (2024)
by: Sadr, Mohsen, et al.
Published: (2024)
Finite Sample Frequency Domain Identification
by: Tsiamis, Anastasios, et al.
Published: (2024)
by: Tsiamis, Anastasios, et al.
Published: (2024)
Bridging Autoencoders and Dynamic Mode Decomposition for Reduced-order Modeling and Control of PDEs
by: Saha, Priyabrata, et al.
Published: (2024)
by: Saha, Priyabrata, et al.
Published: (2024)
Similar Items
-
Accelerating Adaptive Systems via Normalized Parameter Estimation Laws
by: Boveiri, Mohammad, et al.
Published: (2025) -
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023) -
Offline Reinforcement Learning via Inverse Optimization
by: Dimanidis, Ioannis, et al.
Published: (2025) -
Scalable Kernel Inverse Optimization
by: Long, Youyuan, et al.
Published: (2024) -
Distributionally Robust Model Predictive Control: Closed-loop Guarantees and Scalable Algorithms
by: McAllister, Robert D., et al.
Published: (2023)