Analysis of Value Iteration Through Absolute Probability Sequences
Fuente:
arXiv
Saved in:
| Main Authors: | Mustafin, Arsenii, Colla, Sebastien, Olshevsky, Alex, Paschalidis, Ioannis Ch. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Closing the gap between SVRG and TD-SVRG with Gradient Splitting
by: Mustafin, Arsenii, et al.
Published: (2022)
by: Mustafin, Arsenii, et al.
Published: (2022)
Geometric Re-Analysis of Classical MDP Solving Algorithms
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
MDP Geometry, Normalization and Reward Balancing Solvers
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Revisiting Value Iteration: Unified Analysis of Discounted and Average-Reward Cases
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
One-Shot Averaging for Distributed TD($λ$) Under Markov Sampling
by: Tian, Haoxing, et al.
Published: (2024)
by: Tian, Haoxing, et al.
Published: (2024)
Bridging the Gap Between Average and Discounted TD Learning
by: Tian, Haoxing, et al.
Published: (2026)
by: Tian, Haoxing, et al.
Published: (2026)
Distributionally Robust Learning in Survival Analysis
by: Jin, Yeping, et al.
Published: (2025)
by: Jin, Yeping, et al.
Published: (2025)
Multiple-policy Evaluation via Density Estimation
by: Chen, Yilei, et al.
Published: (2024)
by: Chen, Yilei, et al.
Published: (2024)
Distributionally Robust Token Optimization in RLHF
by: Jin, Yeping, et al.
Published: (2026)
by: Jin, Yeping, et al.
Published: (2026)
Adversarial Imitation Learning from Visual Observations using Latent Information
by: Giammarino, Vittorio, et al.
Published: (2023)
by: Giammarino, Vittorio, et al.
Published: (2023)
DRO-Augment Framework: Robustness by Synergizing Wasserstein Distributionally Robust Optimization and Data Augmentation
by: Hu, Jiaming, et al.
Published: (2025)
by: Hu, Jiaming, et al.
Published: (2025)
Visually Robust Adversarial Imitation Learning from Videos with Contrastive Learning
by: Giammarino, Vittorio, et al.
Published: (2024)
by: Giammarino, Vittorio, et al.
Published: (2024)
Generalized Policy Improvement Algorithms with Theoretically Supported Sample Reuse
by: Queeney, James, et al.
Published: (2022)
by: Queeney, James, et al.
Published: (2022)
Provably Efficient Off-Policy Adversarial Imitation Learning with Convergence Guarantees
by: Chen, Yilei, et al.
Published: (2024)
by: Chen, Yilei, et al.
Published: (2024)
Improving Adaptive Online Learning Using Refined Discretization
by: Zhang, Zhiyu, et al.
Published: (2023)
by: Zhang, Zhiyu, et al.
Published: (2023)
A Model-Based Approach for Improving Reinforcement Learning Efficiency Leveraging Expert Observations
by: Ozcan, Erhan Can, et al.
Published: (2024)
by: Ozcan, Erhan Can, et al.
Published: (2024)
Convex SGD: Generalization Without Early Stopping
by: Hendrickx, Julien, et al.
Published: (2024)
by: Hendrickx, Julien, et al.
Published: (2024)
Network Epidemic Control via Model Predictive Control: Extended Version
by: Talaei, Mahtab, et al.
Published: (2026)
by: Talaei, Mahtab, et al.
Published: (2026)
Smooth Ranking SVM via Cutting-Plane Method
by: Ozcan, Erhan Can, et al.
Published: (2024)
by: Ozcan, Erhan Can, et al.
Published: (2024)
Network-Based Epidemic Control Through Optimal Travel and Quarantine Management
by: Talaei, Mahtab, et al.
Published: (2024)
by: Talaei, Mahtab, et al.
Published: (2024)
Optimal Transport Perturbations for Safe Reinforcement Learning with Robustness Guarantees
by: Queeney, James, et al.
Published: (2023)
by: Queeney, James, et al.
Published: (2023)
Towards General Preference Alignment: Diffusion Models at Nash Equilibrium
by: Hu, Jiaming, et al.
Published: (2026)
by: Hu, Jiaming, et al.
Published: (2026)
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Data Deletion Can Help in Adaptive RL
by: Budhraja, Param, et al.
Published: (2026)
by: Budhraja, Param, et al.
Published: (2026)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
Cooperative Multi-Agent Graph Bandits: UCB Algorithm and Regret Analysis
by: Paschalidis, Phevos, et al.
Published: (2024)
by: Paschalidis, Phevos, et al.
Published: (2024)
Adaptive Iterative Soft-Thresholding Algorithm with the Median Absolute Deviation
by: Feng, Yining, et al.
Published: (2025)
by: Feng, Yining, et al.
Published: (2025)
Towards Stable Machine Learning Model Retraining via Slowly Varying Sequences
by: Bertsimas, Dimitris, et al.
Published: (2024)
by: Bertsimas, Dimitris, et al.
Published: (2024)
Absolute State-wise Constrained Policy Optimization: High-Probability State-wise Constraints Satisfaction
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
AbbIE: Autoregressive Block-Based Iterative Encoder for Efficient Sequence Modeling
by: Aleksandrov, Preslav, et al.
Published: (2025)
by: Aleksandrov, Preslav, et al.
Published: (2025)
PRISM: Parallel Residual Iterative Sequence Model
by: Jiang, Jie, et al.
Published: (2026)
by: Jiang, Jie, et al.
Published: (2026)
Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
Deflated Dynamics Value Iteration
by: Lee, Jongmin, et al.
Published: (2024)
by: Lee, Jongmin, et al.
Published: (2024)
Highway Value Iteration Networks
by: Wang, Yuhui, et al.
Published: (2024)
by: Wang, Yuhui, et al.
Published: (2024)
Upper Entropy for 2-Monotone Lower Probabilities
by: Vu, Tuan-Anh, et al.
Published: (2026)
by: Vu, Tuan-Anh, et al.
Published: (2026)
Guided Sequence-Structure Generative Modeling for Iterative Antibody Optimization
by: Raghu, Aniruddh, et al.
Published: (2025)
by: Raghu, Aniruddh, et al.
Published: (2025)
Rank-One Modified Value Iteration
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
by: Kolarijani, Arman Sharifi, et al.
Published: (2025)
Finite-Time Analysis of Q-Value Iteration for General-Sum Stackelberg Games
by: Jeong, Narim, et al.
Published: (2026)
by: Jeong, Narim, et al.
Published: (2026)
Similar Items
-
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024) -
Closing the gap between SVRG and TD-SVRG with Gradient Splitting
by: Mustafin, Arsenii, et al.
Published: (2022) -
Geometric Re-Analysis of Classical MDP Solving Algorithms
by: Mustafin, Arsenii, et al.
Published: (2025) -
MDP Geometry, Normalization and Reward Balancing Solvers
by: Mustafin, Arsenii, et al.
Published: (2024) -
Revisiting Value Iteration: Unified Analysis of Discounted and Average-Reward Cases
by: Mustafin, Arsenii, et al.
Published: (2025)