Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rozada, Sergio, Paternain, Santiago, Marques, Antonio G. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tensor Low-rank Approximation of Finite-horizon Value Functions
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
A Tensor Low-Rank Approximation for Value Functions in Multi-Task Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Matrix Low-Rank Approximation For Policy Gradient Methods
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
Matrix Low-Rank Trust Region Policy Optimization
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
Multilinear Tensor Low-Rank Approximation for Policy-Gradient Methods in Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Low-Rank Tensors for Multi-Dimensional Markov Models
von: Navarro, Madeline, et al.
Veröffentlicht: (2024)
von: Navarro, Madeline, et al.
Veröffentlicht: (2024)
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
A Multi-resolution Low-rank Tensor Decomposition
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
Learning Policy Representations for Steerable Behavior Synthesis
von: Li, Beiming, et al.
Veröffentlicht: (2026)
von: Li, Beiming, et al.
Veröffentlicht: (2026)
Graph-Aware Diffusion for Signal Generation
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
von: Cho, Taehyun, et al.
Veröffentlicht: (2024)
von: Cho, Taehyun, et al.
Veröffentlicht: (2024)
Low-Rank Tensor Decompositions for the Theory of Neural Networks
von: Borsoi, Ricardo, et al.
Veröffentlicht: (2025)
von: Borsoi, Ricardo, et al.
Veröffentlicht: (2025)
Low Tensor-Rank Adaptation of Kolmogorov--Arnold Networks
von: Gao, Yihang, et al.
Veröffentlicht: (2025)
von: Gao, Yihang, et al.
Veröffentlicht: (2025)
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
von: Zhuang, Yuan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yuan, et al.
Veröffentlicht: (2026)
Low-Rank GEMM: Efficient Matrix Multiplication via Low-Rank Approximation with FP8 Acceleration
von: Metere, Alfredo
Veröffentlicht: (2025)
von: Metere, Alfredo
Veröffentlicht: (2025)
Joint Tensor-Train Parameterization for Efficient and Expressive Low-Rank Adaptation
von: Qi, Jun, et al.
Veröffentlicht: (2025)
von: Qi, Jun, et al.
Veröffentlicht: (2025)
Rank-1 Approximation of Inverse Fisher for Natural Policy Gradients in Deep Reinforcement Learning
von: Huo, Yingxiao, et al.
Veröffentlicht: (2026)
von: Huo, Yingxiao, et al.
Veröffentlicht: (2026)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Optimal Policy Sparsification and Low Rank Decomposition for Deep Reinforcement Learning
von: Goddla, Vikram
Veröffentlicht: (2024)
von: Goddla, Vikram
Veröffentlicht: (2024)
Towards Differentially Private Reinforcement Learning with General Function Approximation
von: He, Yi, et al.
Veröffentlicht: (2026)
von: He, Yi, et al.
Veröffentlicht: (2026)
MetaLoRA: Tensor-Enhanced Adaptive Low-Rank Fine-tuning
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
von: Wang, Maolin, et al.
Veröffentlicht: (2025)
Ranking-aware Reinforcement Learning for Ordinal Ranking
von: Hao, Aiming, et al.
Veröffentlicht: (2026)
von: Hao, Aiming, et al.
Veröffentlicht: (2026)
VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
LoTR: Low Tensor Rank Weight Adaptation
von: Bershatsky, Daniel, et al.
Veröffentlicht: (2024)
von: Bershatsky, Daniel, et al.
Veröffentlicht: (2024)
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
von: Huang, Jiawei, et al.
Veröffentlicht: (2023)
von: Huang, Jiawei, et al.
Veröffentlicht: (2023)
AutoLoRA: Automatically Tuning Matrix Ranks in Low-Rank Adaptation Based on Meta Learning
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
Is Value Functions Estimation with Classification Plug-and-play for Offline Reinforcement Learning?
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
von: Tarasov, Denis, et al.
Veröffentlicht: (2024)
TLoRA: Tri-Matrix Low-Rank Adaptation of Large Language Models
von: Islam, Tanvir
Veröffentlicht: (2025)
von: Islam, Tanvir
Veröffentlicht: (2025)
Is there Value in Reinforcement Learning?
von: Fox, Lior, et al.
Veröffentlicht: (2025)
von: Fox, Lior, et al.
Veröffentlicht: (2025)
Understanding the Learning Dynamics of LoRA: A Gradient Flow Perspective on Low-Rank Adaptation in Matrix Factorization
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
von: Xu, Ziqing, et al.
Veröffentlicht: (2025)
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation
von: Zhao, Runze, et al.
Veröffentlicht: (2025)
von: Zhao, Runze, et al.
Veröffentlicht: (2025)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
Sample and Oracle Efficient Reinforcement Learning for MDPs with Linearly-Realizable Value Functions
von: Mhammedi, Zakaria
Veröffentlicht: (2024)
von: Mhammedi, Zakaria
Veröffentlicht: (2024)
Value-Distributional Model-Based Reinforcement Learning
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
von: Luis, Carlos E., et al.
Veröffentlicht: (2023)
RL for Mitigating Cascading Failures: Targeted Exploration via Sensitivity Factors
von: Dwivedi, Anmol, et al.
Veröffentlicht: (2024)
von: Dwivedi, Anmol, et al.
Veröffentlicht: (2024)
Tensor Train Low-rank Approximation (TT-LoRA): Democratizing AI with Accelerated LLMs
von: Anjum, Afia, et al.
Veröffentlicht: (2024)
von: Anjum, Afia, et al.
Veröffentlicht: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
Dynamic Low-rank Approximation of Full-Matrix Preconditioner for Training Generalized Linear Models
von: Matveeva, Tatyana, et al.
Veröffentlicht: (2025)
von: Matveeva, Tatyana, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Tensor Low-rank Approximation of Finite-horizon Value Functions
von: Rozada, Sergio, et al.
Veröffentlicht: (2024) -
A Tensor Low-Rank Approximation for Value Functions in Multi-Task Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025) -
Matrix Low-Rank Approximation For Policy Gradient Methods
von: Rozada, Sergio, et al.
Veröffentlicht: (2024) -
Matrix Low-Rank Trust Region Policy Optimization
von: Rozada, Sergio, et al.
Veröffentlicht: (2024) -
Multilinear Tensor Low-Rank Approximation for Policy-Gradient Methods in Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)