Rank-One Modified Value Iteration
Fuente:
arXiv
Guardado en:
| Autores principales: | Kolarijani, Arman Sharifi, Ok, Tolga, Esfahani, Peyman Mohajerin, Kolarijani, Mohamad Amin Sharif |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Control and Reinforcement Learning through the Lens of Optimization: An Algorithmic Perspective
por: Ok, Tolga, et al.
Publicado: (2025)
por: Ok, Tolga, et al.
Publicado: (2025)
From Optimization to Control: Quasi Policy Iteration
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023)
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023)
Offline Reinforcement Learning via Inverse Optimization
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
por: Dimanidis, Ioannis, et al.
Publicado: (2025)
Fitted Q-Iteration via Max-Plus-Linear Approximation
por: Liu, Y., et al.
Publicado: (2024)
por: Liu, Y., et al.
Publicado: (2024)
Scalable Kernel Inverse Optimization
por: Long, Youyuan, et al.
Publicado: (2024)
por: Long, Youyuan, et al.
Publicado: (2024)
Variance-Reduced Cascade Q-learning: Algorithms and Sample Complexity
por: Boveiri, Mohammad, et al.
Publicado: (2024)
por: Boveiri, Mohammad, et al.
Publicado: (2024)
Nonlinear Distributionally Robust Optimization
por: Sheriff, Mohammed Rayyan, et al.
Publicado: (2023)
por: Sheriff, Mohammed Rayyan, et al.
Publicado: (2023)
Learning in Inverse Optimization: Incenter Cost, Augmented Suboptimality Loss, and Algorithms
por: Scroccaro, Pedro Zattoni, et al.
Publicado: (2023)
por: Scroccaro, Pedro Zattoni, et al.
Publicado: (2023)
Tight Generalization Bounds for Noiseless Inverse Optimization
por: Fatemi, Pouria, et al.
Publicado: (2026)
por: Fatemi, Pouria, et al.
Publicado: (2026)
Fast Algorithm for Constrained Linear Inverse Problems
por: Sheriff, Mohammed Rayyan, et al.
Publicado: (2022)
por: Sheriff, Mohammed Rayyan, et al.
Publicado: (2022)
Inverse Optimization for Routing Problems
por: Scroccaro, Pedro Zattoni, et al.
Publicado: (2023)
por: Scroccaro, Pedro Zattoni, et al.
Publicado: (2023)
Wasserstein Distributionally Robust Optimization: Theory and Applications in Machine Learning
por: Kuhn, Daniel, et al.
Publicado: (2019)
por: Kuhn, Daniel, et al.
Publicado: (2019)
Accelerating Adaptive Systems via Normalized Parameter Estimation Laws
por: Boveiri, Mohammad, et al.
Publicado: (2025)
por: Boveiri, Mohammad, et al.
Publicado: (2025)
Distributionally Robust Model Predictive Control: Closed-loop Guarantees and Scalable Algorithms
por: McAllister, Robert D., et al.
Publicado: (2023)
por: McAllister, Robert D., et al.
Publicado: (2023)
Inverse Optimization via Learning Feasible Regions
por: Ren, Ke, et al.
Publicado: (2025)
por: Ren, Ke, et al.
Publicado: (2025)
Deflated Dynamics Value Iteration
por: Lee, Jongmin, et al.
Publicado: (2024)
por: Lee, Jongmin, et al.
Publicado: (2024)
A Frank-Wolfe Algorithm for Strongly Monotone Variational Inequalities
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2025)
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2025)
Locally Linear Convergence for Nonsmooth Convex Optimization via Coupled Smoothing and Momentum
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2025)
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2025)
A New Lineserach for Accelerated Composite Minimization
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2024)
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2024)
A Hybrid Algorithm for Monotone Variational Inequalities
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2026)
por: Baghbadorani, Reza Rahimi, et al.
Publicado: (2026)
Optimal Transportation by Orthogonal Coupling Dynamics
por: Sadr, Mohsen, et al.
Publicado: (2024)
por: Sadr, Mohsen, et al.
Publicado: (2024)
Efficient Uncertainty Propagation with Guarantees in Wasserstein Distance
por: Figueiredo, Eduardo, et al.
Publicado: (2025)
por: Figueiredo, Eduardo, et al.
Publicado: (2025)
Robust Multivariate Detection and Estimation with Fault Frequency Content Information
por: Dong, Jingwei, et al.
Publicado: (2023)
por: Dong, Jingwei, et al.
Publicado: (2023)
A Saddle Point Algorithm for Robust Data-Driven Factor Model Problems
por: Khodakaramzadeh, Shabnam, et al.
Publicado: (2025)
por: Khodakaramzadeh, Shabnam, et al.
Publicado: (2025)
Optimal Bayesian Affine Estimator and Active Learning for the Wiener Model
por: Vakili, Sasan, et al.
Publicado: (2025)
por: Vakili, Sasan, et al.
Publicado: (2025)
Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement Learning
por: Di, Qiwei, et al.
Publicado: (2023)
por: Di, Qiwei, et al.
Publicado: (2023)
Convex Formulations for Training Two-Layer ReLU Neural Networks
por: Prakhya, Karthik, et al.
Publicado: (2024)
por: Prakhya, Karthik, et al.
Publicado: (2024)
User-centric Vehicle-to-Grid Optimization with an Input Convex Neural Network-based Battery Degradation Model
por: Mallick, Arghya, et al.
Publicado: (2025)
por: Mallick, Arghya, et al.
Publicado: (2025)
Sequential QCQP for Bilevel Optimization with Line Search
por: Sharifi, Sina, et al.
Publicado: (2025)
por: Sharifi, Sina, et al.
Publicado: (2025)
Truncated Variance Reduced Value Iteration
por: Jin, Yujia, et al.
Publicado: (2024)
por: Jin, Yujia, et al.
Publicado: (2024)
Intersection of Reinforcement Learning and Bayesian Optimization for Intelligent Control of Industrial Processes: A Safe MPC-based DPG using Multi-Objective BO
por: Esfahani, Hossein Nejatbakhsh, et al.
Publicado: (2025)
por: Esfahani, Hossein Nejatbakhsh, et al.
Publicado: (2025)
Single Point-Based Distributed Zeroth-Order Optimization with a Non-Convex Stochastic Objective Function
por: Mhanna, Elissa, et al.
Publicado: (2024)
por: Mhanna, Elissa, et al.
Publicado: (2024)
Safe Gradient Flow for Bilevel Optimization
por: Sharifi, Sina, et al.
Publicado: (2025)
por: Sharifi, Sina, et al.
Publicado: (2025)
Low-Rank Extragradient Method for Nonsmooth and Low-Rank Matrix Optimization Problems
por: Garber, Dan, et al.
Publicado: (2022)
por: Garber, Dan, et al.
Publicado: (2022)
Low-Rank Mirror-Prox for Nonsmooth and Low-Rank Matrix Optimization Problems
por: Garber, Dan, et al.
Publicado: (2022)
por: Garber, Dan, et al.
Publicado: (2022)
One Rank at a Time: Cascading Error Dynamics in Sequential Learning
por: Vandchali, Mahtab Alizadeh, et al.
Publicado: (2025)
por: Vandchali, Mahtab Alizadeh, et al.
Publicado: (2025)
Probabilistic Iterative Hard Thresholding for Sparse Learning
por: Bergamaschi, Matteo, et al.
Publicado: (2024)
por: Bergamaschi, Matteo, et al.
Publicado: (2024)
On the Last-Iterate Convergence of Shuffling Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2024)
por: Liu, Zijian, et al.
Publicado: (2024)
Iterative Minimax Games with Coupled Linear Constraints
por: Zhang, Huiling, et al.
Publicado: (2022)
por: Zhang, Huiling, et al.
Publicado: (2022)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
por: Tang, Xun, et al.
Publicado: (2024)
por: Tang, Xun, et al.
Publicado: (2024)
Ejemplares similares
-
Control and Reinforcement Learning through the Lens of Optimization: An Algorithmic Perspective
por: Ok, Tolga, et al.
Publicado: (2025) -
From Optimization to Control: Quasi Policy Iteration
por: Kolarijani, Mohammad Amin Sharifi, et al.
Publicado: (2023) -
Offline Reinforcement Learning via Inverse Optimization
por: Dimanidis, Ioannis, et al.
Publicado: (2025) -
Fitted Q-Iteration via Max-Plus-Linear Approximation
por: Liu, Y., et al.
Publicado: (2024) -
Scalable Kernel Inverse Optimization
por: Long, Youyuan, et al.
Publicado: (2024)