Saved in:
| Main Authors: | Bruns-Smith, David, Zhou, Angela |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2302.00662 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024)
by: Liu, Y., et al.
Published: (2024)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Structured Difference-of-Q via Orthogonal Learning
by: Cao, Defu, et al.
Published: (2024)
by: Cao, Defu, et al.
Published: (2024)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
by: Wan, Jia, et al.
Published: (2024)
by: Wan, Jia, et al.
Published: (2024)
Robust Sublinear Convergence Rates for Iterative Bregman Projections
by: Peyré, Gabriel
Published: (2026)
by: Peyré, Gabriel
Published: (2026)
Regularized Q-learning through Robust Averaging
by: Schmitt-Förster, Peter, et al.
Published: (2024)
by: Schmitt-Förster, Peter, et al.
Published: (2024)
Distributionally Robust Optimization via Iterative Algorithms in Continuous Probability Spaces
by: Zhu, Linglingzhi, et al.
Published: (2024)
by: Zhu, Linglingzhi, et al.
Published: (2024)
Global Convergence of Iteratively Reweighted Least Squares for Robust Subspace Recovery
by: Lerman, Gilad, et al.
Published: (2025)
by: Lerman, Gilad, et al.
Published: (2025)
On the Last-Iterate Convergence of Shuffling Gradient Methods
by: Liu, Zijian, et al.
Published: (2024)
by: Liu, Zijian, et al.
Published: (2024)
Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods
by: Liu, Zijian, et al.
Published: (2023)
by: Liu, Zijian, et al.
Published: (2023)
Reward-Relevance-Filtered Linear Offline Reinforcement Learning
by: Zhou, Angela
Published: (2024)
by: Zhou, Angela
Published: (2024)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
by: Xu, Mingyuan, et al.
Published: (2026)
by: Xu, Mingyuan, et al.
Published: (2026)
Robustness of Iteratively Pre-Conditioned Gradient-Descent Method: The Case of Distributed Linear Regression Problem
by: Chakrabarti, Kushal, et al.
Published: (2021)
by: Chakrabarti, Kushal, et al.
Published: (2021)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022)
by: Neufeld, Ariel, et al.
Published: (2022)
Distributionally Robust Deep Q-Learning
by: Lu, Chung I, et al.
Published: (2025)
by: Lu, Chung I, et al.
Published: (2025)
Improved Last-Iterate Convergence of Shuffling Gradient Methods for Nonsmooth Convex Optimization
by: Liu, Zijian, et al.
Published: (2025)
by: Liu, Zijian, et al.
Published: (2025)
Stochastic Hessian Fittings with Lie Groups
by: Li, Xi-Lin
Published: (2024)
by: Li, Xi-Lin
Published: (2024)
Inferring Dynamic Networks from Marginals with Iterative Proportional Fitting
by: Chang, Serina, et al.
Published: (2024)
by: Chang, Serina, et al.
Published: (2024)
Robust Losses for Decision-Focused Learning
by: Schutte, Noah, et al.
Published: (2023)
by: Schutte, Noah, et al.
Published: (2023)
SHANG++: Robust Stochastic Acceleration under Multiplicative Noise
by: Yu, Yaxin, et al.
Published: (2026)
by: Yu, Yaxin, et al.
Published: (2026)
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024)
by: Lee, Jongmin, et al.
Published: (2024)
A Primal-Dual Online Learning Approach for Dynamic Pricing of Sequentially Displayed Complementary Items under Sale Constraints
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
A Sequential Quadratic Programming Method with High Probability Complexity Bounds for Nonlinear Equality Constrained Stochastic Optimization
by: Berahas, Albert S., et al.
Published: (2023)
by: Berahas, Albert S., et al.
Published: (2023)
Adaptively Robust LLM Inference Optimization under Prediction Uncertainty
by: Chen, Zixi, et al.
Published: (2025)
by: Chen, Zixi, et al.
Published: (2025)
Robust Neural IDA-PBC: passivity-based stabilization under approximations
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
by: Sanchez-Escalonilla, Santiago, et al.
Published: (2024)
Rank-One Modified Value Iteration
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
by: Buyuktahtakin, I. Esra
Published: (2026)
by: Buyuktahtakin, I. Esra
Published: (2026)
From Optimization to Control: Quasi Policy Iteration
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
by: Kolarijani, Mohammad Amin Sharifi, et al.
Published: (2023)
From Inexact Gradients to Byzantine Robustness: Acceleration and Optimization under Similarity
by: Gaucher, Renaud, et al.
Published: (2026)
by: Gaucher, Renaud, et al.
Published: (2026)
Sequential QCQP for Bilevel Optimization with Line Search
by: Sharifi, Sina, et al.
Published: (2025)
by: Sharifi, Sina, et al.
Published: (2025)
Probabilistic Iterative Hard Thresholding for Sparse Learning
by: Bergamaschi, Matteo, et al.
Published: (2024)
by: Bergamaschi, Matteo, et al.
Published: (2024)
Iterative Minimax Games with Coupled Linear Constraints
by: Zhang, Huiling, et al.
Published: (2022)
by: Zhang, Huiling, et al.
Published: (2022)
Accelerating Sinkhorn Algorithm with Sparse Newton Iterations
by: Tang, Xun, et al.
Published: (2024)
by: Tang, Xun, et al.
Published: (2024)
Q-Learning under Finite Model Uncertainty
by: Sester, Julian, et al.
Published: (2024)
by: Sester, Julian, et al.
Published: (2024)
On the Foundation of Distributionally Robust Reinforcement Learning
by: Wang, Shengbo, et al.
Published: (2023)
by: Wang, Shengbo, et al.
Published: (2023)
Self-Supervised Learning of Iterative Solvers for Constrained Optimization
by: Lüken, Lukas, et al.
Published: (2024)
by: Lüken, Lukas, et al.
Published: (2024)
Gradient Descent's Last Iterate is Often (slightly) Suboptimal
by: Kornowski, Guy, et al.
Published: (2026)
by: Kornowski, Guy, et al.
Published: (2026)
Convergence Rate of the Last Iterate of Stochastic Proximal Algorithms
by: Vaidyan, Kevin Kurian Thomas, et al.
Published: (2026)
by: Vaidyan, Kevin Kurian Thomas, et al.
Published: (2026)
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
Similar Items
-
Fitted Q-Iteration via Max-Plus-Linear Approximation
by: Liu, Y., et al.
Published: (2024) -
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024) -
Structured Difference-of-Q via Orthogonal Learning
by: Cao, Defu, et al.
Published: (2024) -
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
by: Wang, Shengbo, et al.
Published: (2023) -
Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
by: Wan, Jia, et al.
Published: (2024)