Accelerating Multi-Task Temporal Difference Learning under Low-Rank Representation
Fuente:
arXiv
Salvato in:
| Autori principali: | Bai, Yitao, Zeng, Sihan, Romberg, Justin, Doan, Thinh T. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
di: Bai, Yitao, et al.
Pubblicazione: (2026)
di: Bai, Yitao, et al.
Pubblicazione: (2026)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2021)
di: Zeng, Sihan, et al.
Pubblicazione: (2021)
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
di: Zeng, Sihan, et al.
Pubblicazione: (2021)
di: Zeng, Sihan, et al.
Pubblicazione: (2021)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
Bayesian meta learning for trustworthy uncertainty quantification
di: Yuan, Zhenyuan, et al.
Pubblicazione: (2024)
di: Yuan, Zhenyuan, et al.
Pubblicazione: (2024)
Accelerated Multi-Time-Scale Stochastic Approximation: Optimal Complexity and Applications in Reinforcement Learning and Multi-Agent Games
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
di: Zeng, Sihan, et al.
Pubblicazione: (2024)
PID Accelerated Temporal Difference Algorithms
di: Bedaywi, Mark, et al.
Pubblicazione: (2024)
di: Bedaywi, Mark, et al.
Pubblicazione: (2024)
Fast Nonlinear Two-Time-Scale Stochastic Approximation: Achieving $O(1/k)$ Finite-Sample Complexity
di: Doan, Thinh T.
Pubblicazione: (2024)
di: Doan, Thinh T.
Pubblicazione: (2024)
Low-Rank Tensors for Multi-Dimensional Markov Models
di: Navarro, Madeline, et al.
Pubblicazione: (2024)
di: Navarro, Madeline, et al.
Pubblicazione: (2024)
Generating Causal Temporal Interaction Graphs for Counterfactual Validation of Temporal Link Prediction
di: Rahman, Aniq Ur, et al.
Pubblicazione: (2026)
di: Rahman, Aniq Ur, et al.
Pubblicazione: (2026)
Finite-Time Accuracy of Temporal-Difference Learning Under Schur-Stable Recursions
di: Lee, Donghwan, et al.
Pubblicazione: (2022)
di: Lee, Donghwan, et al.
Pubblicazione: (2022)
Towards Fast Rates for Federated and Multi-Task Reinforcement Learning
di: Zhu, Feng, et al.
Pubblicazione: (2024)
di: Zhu, Feng, et al.
Pubblicazione: (2024)
Leveraging Multi-Task Learning for Multi-Label Power System Security Assessment
di: Za'ter, Muhy Eddin, et al.
Pubblicazione: (2025)
di: Za'ter, Muhy Eddin, et al.
Pubblicazione: (2025)
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning
di: Mitra, Aritra, et al.
Pubblicazione: (2023)
di: Mitra, Aritra, et al.
Pubblicazione: (2023)
Structured Cooperative Multi-Agent Reinforcement Learning: a Bayesian Network Perspective
di: Syed, Shahbaz P Qadri, et al.
Pubblicazione: (2025)
di: Syed, Shahbaz P Qadri, et al.
Pubblicazione: (2025)
Training Task Reasoning LLM Agents for Multi-turn Task Planning via Single-turn Reinforcement Learning
di: Hu, Hanjiang, et al.
Pubblicazione: (2025)
di: Hu, Hanjiang, et al.
Pubblicazione: (2025)
Internal State-Based Policy Gradient Methods for Partially Observable Markov Potential Games
di: Yang, Wonseok, et al.
Pubblicazione: (2026)
di: Yang, Wonseok, et al.
Pubblicazione: (2026)
Learning AC Power Flow Solutions using a Data-Dependent Variational Quantum Circuit
di: Le, Thinh Viet, et al.
Pubblicazione: (2025)
di: Le, Thinh Viet, et al.
Pubblicazione: (2025)
Regret Analysis of Multi-task Representation Learning for Linear-Quadratic Adaptive Control
di: Lee, Bruce D., et al.
Pubblicazione: (2024)
di: Lee, Bruce D., et al.
Pubblicazione: (2024)
Semi-Supervised Multi-Task Learning Based Framework for Power System Security Assessment
di: Za'ter, Muhy Eddin, et al.
Pubblicazione: (2024)
di: Za'ter, Muhy Eddin, et al.
Pubblicazione: (2024)
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning
di: Xu, Duo, et al.
Pubblicazione: (2024)
di: Xu, Duo, et al.
Pubblicazione: (2024)
A Hierarchical Surrogate Model for Efficient Multi-Task Parameter Learning in Closed-Loop Control
di: Hirt, Sebastian, et al.
Pubblicazione: (2025)
di: Hirt, Sebastian, et al.
Pubblicazione: (2025)
Towards Safe Multi-Task Bayesian Optimization
di: Lübsen, Jannis O., et al.
Pubblicazione: (2023)
di: Lübsen, Jannis O., et al.
Pubblicazione: (2023)
Neural Operators for Multi-Task Control and Adaptation
di: Sewell, David, et al.
Pubblicazione: (2026)
di: Sewell, David, et al.
Pubblicazione: (2026)
Federated Multi-Agent Deep Reinforcement Learning Approach via Physics-Informed Reward for Multi-Microgrid Energy Management
di: Li, Yuanzheng, et al.
Pubblicazione: (2022)
di: Li, Yuanzheng, et al.
Pubblicazione: (2022)
An Analysis of Safety Guarantees in Multi-Task Bayesian Optimization
di: Luebsen, Jannis O., et al.
Pubblicazione: (2025)
di: Luebsen, Jannis O., et al.
Pubblicazione: (2025)
Accelerating Optimization and Machine Learning through Decentralization
di: Chen, Ziqin, et al.
Pubblicazione: (2026)
di: Chen, Ziqin, et al.
Pubblicazione: (2026)
Adaptive Policy Learning to Additional Tasks
di: Hao, Wenjian, et al.
Pubblicazione: (2023)
di: Hao, Wenjian, et al.
Pubblicazione: (2023)
Online Algorithm for Node Feature Forecasting in Temporal Graphs
di: Rahman, Aniq Ur, et al.
Pubblicazione: (2024)
di: Rahman, Aniq Ur, et al.
Pubblicazione: (2024)
Solving Conic Programs over Sparse Graphs using a Variational Quantum Approach: The Case of the Optimal Power Flow
di: Le, Thinh Viet, et al.
Pubblicazione: (2025)
di: Le, Thinh Viet, et al.
Pubblicazione: (2025)
Multi-Mode Process Control Using Multi-Task Inverse Reinforcement Learning
di: Lin, Runze, et al.
Pubblicazione: (2025)
di: Lin, Runze, et al.
Pubblicazione: (2025)
CBF-based Probabilistic Safe Navigation under Unknown Nonlinear Obstacle Dynamics
di: Lee, Jiwon, et al.
Pubblicazione: (2026)
di: Lee, Jiwon, et al.
Pubblicazione: (2026)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
di: Adams, Katherine B., et al.
Pubblicazione: (2025)
di: Adams, Katherine B., et al.
Pubblicazione: (2025)
Learning Physically Consistent Lagrangian Control Models Without Acceleration Measurements
di: Laiche, Ibrahim, et al.
Pubblicazione: (2025)
di: Laiche, Ibrahim, et al.
Pubblicazione: (2025)
Accelerating Reinforcement Learning for Wind Farm Control via Expert Demonstrations
di: Nilsen, Marcus Binder, et al.
Pubblicazione: (2026)
di: Nilsen, Marcus Binder, et al.
Pubblicazione: (2026)
Identifiable Representation and Model Learning for Latent Dynamic Systems
di: Zhang, Congxi, et al.
Pubblicazione: (2024)
di: Zhang, Congxi, et al.
Pubblicazione: (2024)
Robust Q-Learning under Corrupted Rewards
di: Maity, Sreejeet, et al.
Pubblicazione: (2024)
di: Maity, Sreejeet, et al.
Pubblicazione: (2024)
Scaling Learning based Policy Optimization for Temporal Logic Tasks by Controller Network Dropout
di: Hashemi, Navid, et al.
Pubblicazione: (2024)
di: Hashemi, Navid, et al.
Pubblicazione: (2024)
Cluster-Based Multi-Agent Task Scheduling for Space-Air-Ground Integrated Networks
di: Wang, Zhiying, et al.
Pubblicazione: (2024)
di: Wang, Zhiying, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2024) -
Finite-Time Analysis of Projected Two-Time-Scale Stochastic Approximation
di: Bai, Yitao, et al.
Pubblicazione: (2026) -
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2021) -
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
di: Zeng, Sihan, et al.
Pubblicazione: (2021) -
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
di: Zeng, Sihan, et al.
Pubblicazione: (2024)