MDP Geometry, Normalization and Reward Balancing Solvers
Fuente:
arXiv
Saved in:
| Main Authors: | Mustafin, Arsenii, Pakharev, Aleksei, Olshevsky, Alex, Paschalidis, Ioannis Ch. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometric Re-Analysis of Classical MDP Solving Algorithms
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Closing the gap between SVRG and TD-SVRG with Gradient Splitting
by: Mustafin, Arsenii, et al.
Published: (2022)
by: Mustafin, Arsenii, et al.
Published: (2022)
Analysis of Value Iteration Through Absolute Probability Sequences
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
Distributionally Robust Learning in Survival Analysis
by: Jin, Yeping, et al.
Published: (2025)
by: Jin, Yeping, et al.
Published: (2025)
Network Epidemic Control via Model Predictive Control: Extended Version
by: Talaei, Mahtab, et al.
Published: (2026)
by: Talaei, Mahtab, et al.
Published: (2026)
Sample Complexity of the Linear Quadratic Regulator: A Reinforcement Learning Lens
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2024)
Sample Complexity of Linear Quadratic Regulator Without Initial Stability
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
by: Moghaddam, Amirreza Neshaei, et al.
Published: (2025)
Data Deletion Can Help in Adaptive RL
by: Budhraja, Param, et al.
Published: (2026)
by: Budhraja, Param, et al.
Published: (2026)
A Stackelberg Game Approach to Control the Overall Load Consumption of a Residential Neighborhood
by: Ozcan, Erhan Can, et al.
Published: (2023)
by: Ozcan, Erhan Can, et al.
Published: (2023)
One-Shot Averaging for Distributed TD($λ$) Under Markov Sampling
by: Tian, Haoxing, et al.
Published: (2024)
by: Tian, Haoxing, et al.
Published: (2024)
Tsallis Entropy Regularization for Linearly Solvable MDP and Linear Quadratic Regulator
by: Hashizume, Yota, et al.
Published: (2024)
by: Hashizume, Yota, et al.
Published: (2024)
A Minimax-MDP Framework with Future-imposed Conditions for Learning-augmented Problems
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Bridging the Gap Between Average and Discounted TD Learning
by: Tian, Haoxing, et al.
Published: (2026)
by: Tian, Haoxing, et al.
Published: (2026)
Locally Interdependent Multi-Agent MDP: Theoretical Framework for Decentralized Agents with Dynamic Dependencies
by: DeWeese, Alex, et al.
Published: (2024)
by: DeWeese, Alex, et al.
Published: (2024)
A Distributed Optimization Framework to Regulate the Electricity Consumption of a Residential Neighborhood with Renewables
by: Ozcan, Erhan Can, et al.
Published: (2023)
by: Ozcan, Erhan Can, et al.
Published: (2023)
Ito Diffusion Approximation of Universal Ito Chains for Sampling, Optimization and Boosting
by: Ustimenko, Aleksei, et al.
Published: (2023)
by: Ustimenko, Aleksei, et al.
Published: (2023)
Adversarial Imitation Learning from Visual Observations using Latent Information
by: Giammarino, Vittorio, et al.
Published: (2023)
by: Giammarino, Vittorio, et al.
Published: (2023)
Differentiation Through Black-Box Quadratic Programming Solvers
by: Magoon, Connor W., et al.
Published: (2024)
by: Magoon, Connor W., et al.
Published: (2024)
Learning to Configure Mathematical Programming Solvers by Mathematical Programming
by: Iommazzo, Gabriele, et al.
Published: (2024)
by: Iommazzo, Gabriele, et al.
Published: (2024)
Self-Supervised Learning of Iterative Solvers for Constrained Optimization
by: Lüken, Lukas, et al.
Published: (2024)
by: Lüken, Lukas, et al.
Published: (2024)
On Learning for Ambiguous Chance Constrained Problems
by: Madhusudanarao, A Ch, et al.
Published: (2023)
by: Madhusudanarao, A Ch, et al.
Published: (2023)
Towards Stable Machine Learning Model Retraining via Slowly Varying Sequences
by: Bertsimas, Dimitris, et al.
Published: (2024)
by: Bertsimas, Dimitris, et al.
Published: (2024)
Speeding Up Mixed-Integer Programming Solvers with Sparse Learning for Branching
by: Bayramoğlu, Selin, et al.
Published: (2026)
by: Bayramoğlu, Selin, et al.
Published: (2026)
A Penalty Approach for Differentiation Through Black-Box Quadratic Programming Solvers
by: Linghu, Yuxuan, et al.
Published: (2026)
by: Linghu, Yuxuan, et al.
Published: (2026)
Revisiting Value Iteration: Unified Analysis of Discounted and Average-Reward Cases
by: Mustafin, Arsenii, et al.
Published: (2025)
by: Mustafin, Arsenii, et al.
Published: (2025)
The Limit Points of (Optimistic) Gradient Descent in Min-Max Optimization
by: Daskalakis, Constantinos, et al.
Published: (2018)
by: Daskalakis, Constantinos, et al.
Published: (2018)
Parameter-Efficient Distributional RL via Normalizing Flows and a Geometry-Aware Cramér Surrogate
by: C., Simo Alami, et al.
Published: (2025)
by: C., Simo Alami, et al.
Published: (2025)
Machine Learning Algorithms for Improving Exact Classical Solvers in Mixed Integer Continuous Optimization
by: Kimiaei, Morteza, et al.
Published: (2025)
by: Kimiaei, Morteza, et al.
Published: (2025)
Robust Implicit Regularization via Weight Normalization
by: Chou, Hung-Hsu, et al.
Published: (2023)
by: Chou, Hung-Hsu, et al.
Published: (2023)
Gradient-Normalized Smoothness for Optimization with Approximate Hessians
by: Semenov, Andrei, et al.
Published: (2025)
by: Semenov, Andrei, et al.
Published: (2025)
Robust Q-Learning under Corrupted Rewards
by: Maity, Sreejeet, et al.
Published: (2024)
by: Maity, Sreejeet, et al.
Published: (2024)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
Reward-Relevance-Filtered Linear Offline Reinforcement Learning
by: Zhou, Angela
Published: (2024)
by: Zhou, Angela
Published: (2024)
Learning Fair And Effective Points-Based Rewards Programs
by: Hssaine, Chamsi, et al.
Published: (2025)
by: Hssaine, Chamsi, et al.
Published: (2025)
MDP modeling for multi-stage stochastic programs
by: Morton, David P., et al.
Published: (2025)
by: Morton, David P., et al.
Published: (2025)
Regularized Adaptive Momentum Dual Averaging with an Efficient Inexact Subproblem Solver for Training Structured Neural Network
by: Huang, Zih-Syuan, et al.
Published: (2024)
by: Huang, Zih-Syuan, et al.
Published: (2024)
$ϕ$-Balancing for Mixture-of-Experts Training
by: Chen, Lizhang, et al.
Published: (2026)
by: Chen, Lizhang, et al.
Published: (2026)
From Gradient Clipping to Normalization for Heavy Tailed SGD
by: Hübler, Florian, et al.
Published: (2024)
by: Hübler, Florian, et al.
Published: (2024)
Absolute Ranking: An Essential Normalization for Benchmarking Optimization Algorithms
by: Jinng, Yunpeng, et al.
Published: (2024)
by: Jinng, Yunpeng, et al.
Published: (2024)
Similar Items
-
Geometric Re-Analysis of Classical MDP Solving Algorithms
by: Mustafin, Arsenii, et al.
Published: (2025) -
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024) -
Closing the gap between SVRG and TD-SVRG with Gradient Splitting
by: Mustafin, Arsenii, et al.
Published: (2022) -
Analysis of Value Iteration Through Absolute Probability Sequences
by: Mustafin, Arsenii, et al.
Published: (2025) -
Distributionally Robust Learning in Survival Analysis
by: Jin, Yeping, et al.
Published: (2025)