Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation
Fuente:
arXiv
Saved in:
| Main Authors: | Rozada, Sergio, Orejuela, Jose Luis, Marques, Antonio G. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tensor Low-rank Approximation of Finite-horizon Value Functions
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2022)
by: Rozada, Sergio, et al.
Published: (2022)
Multilinear Tensor Low-Rank Approximation for Policy-Gradient Methods in Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2025)
by: Rozada, Sergio, et al.
Published: (2025)
A Tensor Low-Rank Approximation for Value Functions in Multi-Task Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2025)
by: Rozada, Sergio, et al.
Published: (2025)
Matrix Low-Rank Approximation For Policy Gradient Methods
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
A Multi-resolution Low-rank Tensor Decomposition
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Low-Rank Tensors for Multi-Dimensional Markov Models
by: Navarro, Madeline, et al.
Published: (2024)
by: Navarro, Madeline, et al.
Published: (2024)
Matrix Low-Rank Trust Region Policy Optimization
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
by: Hong, Kihyuk, et al.
Published: (2024)
by: Hong, Kihyuk, et al.
Published: (2024)
Landscape of Policy Optimization for Finite Horizon MDPs with General State and Action
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
Low-Rank MDPs with Continuous Action Spaces
by: Bennett, Andrew, et al.
Published: (2023)
by: Bennett, Andrew, et al.
Published: (2023)
Low-Rank Tensor Approximation of Weights in Large Language Models via Cosine Lanczos Bidiagonalization
by: Ichi, A. El, et al.
Published: (2026)
by: Ichi, A. El, et al.
Published: (2026)
Efficient Model-Free Exploration in Low-Rank MDPs
by: Mhammedi, Zakaria, et al.
Published: (2023)
by: Mhammedi, Zakaria, et al.
Published: (2023)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
by: Zhou, Runlin, et al.
Published: (2025)
by: Zhou, Runlin, et al.
Published: (2025)
A Random Matrix Approach to Low-Multilinear-Rank Tensor Approximation
by: Lebeau, Hugo, et al.
Published: (2024)
by: Lebeau, Hugo, et al.
Published: (2024)
Beating Adversarial Low-Rank MDPs with Unknown Transition and Bandit Feedback
by: Liu, Haolin, et al.
Published: (2024)
by: Liu, Haolin, et al.
Published: (2024)
Breaking the Computational Barrier: Provably Efficient Actor-Critic for Low-Rank MDPs
by: Huang, Ruiquan, et al.
Published: (2026)
by: Huang, Ruiquan, et al.
Published: (2026)
Offline Bayesian Aleatoric and Epistemic Uncertainty Quantification and Posterior Value Optimisation in Finite-State MDPs
by: Valdettaro, Filippo, et al.
Published: (2024)
by: Valdettaro, Filippo, et al.
Published: (2024)
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Low Rank Tensor Completion via Adaptive ADMM
by: Führling, Niclas, et al.
Published: (2026)
by: Führling, Niclas, et al.
Published: (2026)
Performance of Rank-One Tensor Approximation on Incomplete Data
by: Lebeau, Hugo
Published: (2025)
by: Lebeau, Hugo
Published: (2025)
Deterministic Policy Gradient Primal-Dual Methods for Continuous-Space Constrained MDPs
by: Rozada, Sergio, et al.
Published: (2024)
by: Rozada, Sergio, et al.
Published: (2024)
Learning Policy Representations for Steerable Behavior Synthesis
by: Li, Beiming, et al.
Published: (2026)
by: Li, Beiming, et al.
Published: (2026)
Graph-Aware Diffusion for Signal Generation
by: Rozada, Sergio, et al.
Published: (2025)
by: Rozada, Sergio, et al.
Published: (2025)
Quantizer Design for Finite Model Approximations, Model Learning, and Quantized Q-Learning for MDPs with Unbounded Spaces
by: Bicer, Osman, et al.
Published: (2025)
by: Bicer, Osman, et al.
Published: (2025)
Score-Based Model for Low-Rank Tensor Recovery
by: Cheng, Zhengyun, et al.
Published: (2025)
by: Cheng, Zhengyun, et al.
Published: (2025)
Efficient Generalized Low-Rank Tensor Contextual Bandits
by: Yi, Qianxin, et al.
Published: (2023)
by: Yi, Qianxin, et al.
Published: (2023)
Low-Rank Tensor Learning by Generalized Nonconvex Regularization
by: Xia, Sijia, et al.
Published: (2024)
by: Xia, Sijia, et al.
Published: (2024)
Low-Rank Key Value Attention
by: O'Neill, James, et al.
Published: (2026)
by: O'Neill, James, et al.
Published: (2026)
Low-Tubal-Rank Tensor Recovery via Factorized Gradient Descent
by: Liu, Zhiyu, et al.
Published: (2024)
by: Liu, Zhiyu, et al.
Published: (2024)
Is Pure Exploitation Sufficient in Exogenous MDPs with Linear Function Approximation?
by: Liang, Hao, et al.
Published: (2026)
by: Liang, Hao, et al.
Published: (2026)
Penalizing Localized Dirichlet Energies in Low Rank Tensor Products
by: Karakasis, Paris A., et al.
Published: (2026)
by: Karakasis, Paris A., et al.
Published: (2026)
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs
by: Zhang, Junkai, et al.
Published: (2023)
by: Zhang, Junkai, et al.
Published: (2023)
Guaranteed Nonconvex Low-Rank Tensor Estimation via Scaled Gradient Descent
by: Wu, Tong
Published: (2025)
by: Wu, Tong
Published: (2025)
Optimal Variance-Dependent Regret Bounds for Infinite-Horizon MDPs
by: Zamir, Guy, et al.
Published: (2026)
by: Zamir, Guy, et al.
Published: (2026)
Low-Rank Tensor Decompositions for the Theory of Neural Networks
by: Borsoi, Ricardo, et al.
Published: (2025)
by: Borsoi, Ricardo, et al.
Published: (2025)
Low Tensor-Rank Adaptation of Kolmogorov--Arnold Networks
by: Gao, Yihang, et al.
Published: (2025)
by: Gao, Yihang, et al.
Published: (2025)
Two-Timescale Critic-Actor for Average Reward MDPs with Function Approximation
by: Panda, Prashansa, et al.
Published: (2024)
by: Panda, Prashansa, et al.
Published: (2024)
Operator SVD with Neural Networks via Nested Low-Rank Approximation
by: Ryu, J. Jon, et al.
Published: (2024)
by: Ryu, J. Jon, et al.
Published: (2024)
Order-Optimal Regret with Novel Policy Gradient Approaches in Infinite-Horizon Average Reward MDPs
by: Ganesh, Swetha, et al.
Published: (2024)
by: Ganesh, Swetha, et al.
Published: (2024)
Similar Items
-
Tensor Low-rank Approximation of Finite-horizon Value Functions
by: Rozada, Sergio, et al.
Published: (2024) -
Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2022) -
Multilinear Tensor Low-Rank Approximation for Policy-Gradient Methods in Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2025) -
A Tensor Low-Rank Approximation for Value Functions in Multi-Task Reinforcement Learning
by: Rozada, Sergio, et al.
Published: (2025) -
Matrix Low-Rank Approximation For Policy Gradient Methods
by: Rozada, Sergio, et al.
Published: (2024)