Theoretical Analysis of Meta Reinforcement Learning: Generalization Bounds and Convergence Guarantees
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Cangqing, Sui, Mingxiu, Sun, Dan, Zhang, Zecheng, Zhou, Yan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CA-BERT: Leveraging Context Awareness for Enhanced Multi-Turn Chat Interaction
by: Liu, Minghao, et al.
Published: (2024)
by: Liu, Minghao, et al.
Published: (2024)
Revisiting Meta-Learning with Noisy Labels: Reweighting Dynamics and Theoretical Guarantees
by: Zhang, Yiming, et al.
Published: (2025)
by: Zhang, Yiming, et al.
Published: (2025)
Cooperative Backdoor Attack in Decentralized Reinforcement Learning with Theoretical Guarantee
by: Gao, Mengtong, et al.
Published: (2024)
by: Gao, Mengtong, et al.
Published: (2024)
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
A Theoretical Understanding of Gradient Bias in Meta-Reinforcement Learning
by: Feng, Xidong, et al.
Published: (2021)
by: Feng, Xidong, et al.
Published: (2021)
Pre-training with Synthetic Data Helps Offline Reinforcement Learning
by: Wang, Zecheng, et al.
Published: (2023)
by: Wang, Zecheng, et al.
Published: (2023)
Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
by: Han, Minghao, et al.
Published: (2026)
by: Han, Minghao, et al.
Published: (2026)
Internalizing Meta-Experience into Memory for Guided Reinforcement Learning in Large Language Models
by: Huang, Shiting, et al.
Published: (2026)
by: Huang, Shiting, et al.
Published: (2026)
Deep Analysis of Time Series Data for Smart Grid Startup Strategies: A Transformer-LSTM-PSO Model Approach
by: Zhang, Zecheng
Published: (2024)
by: Zhang, Zecheng
Published: (2024)
Theoretical Convergence of SMOTE-Generated Samples
by: Kamalov, Firuz, et al.
Published: (2026)
by: Kamalov, Firuz, et al.
Published: (2026)
Gradient Flow Convergence Guarantee for General Neural Network Architectures
by: Jakhmola, Yash
Published: (2025)
by: Jakhmola, Yash
Published: (2025)
On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks
by: Yan, Mingsong, et al.
Published: (2025)
by: Yan, Mingsong, et al.
Published: (2025)
Adapting LLMs for Efficient Context Processing through Soft Prompt Compression
by: Wang, Cangqing, et al.
Published: (2024)
by: Wang, Cangqing, et al.
Published: (2024)
Learning Theory of the SVRG: Generalization and Convergence Analysis
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Locally Linear Continual Learning for Time Series based on VC-Theoretical Generalization Bounds
by: Ferreira, Yan V. G., et al.
Published: (2026)
by: Ferreira, Yan V. G., et al.
Published: (2026)
Quantifying Multimodal Capabilities: Formal Generalization Guarantees in Pairwise Metric Learning
by: Zhou, Richeng, et al.
Published: (2026)
by: Zhou, Richeng, et al.
Published: (2026)
Research on Key Technologies for Cross-Cloud Federated Training of Large Language Models
by: Yang, Haowei, et al.
Published: (2024)
by: Yang, Haowei, et al.
Published: (2024)
Directed-MAML: Meta Reinforcement Learning Algorithm with Task-directed Approximation
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
A Unified and Stable Risk Minimization Framework for Weakly Supervised Learning with Theoretical Guarantees
by: Zhang, Miao, et al.
Published: (2025)
by: Zhang, Miao, et al.
Published: (2025)
Robust Heterogeneous Analog-Digital Computing for Mixture-of-Experts Models with Theoretical Generalization Guarantees
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
by: Chowdhury, Mohammed Nowaz Rabbani, et al.
Published: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
Is Inverse Reinforcement Learning Harder than Standard Reinforcement Learning? A Theoretical Perspective
by: Zhao, Lei, et al.
Published: (2023)
by: Zhao, Lei, et al.
Published: (2023)
Learning from N-Tuple Data with M Positive Instances: Unbiased Risk Estimation and Theoretical Guarantees
by: Zhang, Miao, et al.
Published: (2025)
by: Zhang, Miao, et al.
Published: (2025)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
Optimal Transport Perturbations for Safe Reinforcement Learning with Robustness Guarantees
by: Queeney, James, et al.
Published: (2023)
by: Queeney, James, et al.
Published: (2023)
A Survey of Constraint Formulations in Safe Reinforcement Learning
by: Wachi, Akifumi, et al.
Published: (2024)
by: Wachi, Akifumi, et al.
Published: (2024)
Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery
by: Cho, Minjae, et al.
Published: (2024)
by: Cho, Minjae, et al.
Published: (2024)
Gradients Must Earn Their Influence: Unifying SFT with Generalized Entropic Objectives
by: Wang, Zecheng, et al.
Published: (2026)
by: Wang, Zecheng, et al.
Published: (2026)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Model-Based Offline Reinforcement Learning with Reliability-Guaranteed Sequence Modeling
by: He, Shenghong
Published: (2025)
by: He, Shenghong
Published: (2025)
FedAdamW: A Communication-Efficient Optimizer with Convergence and Generalization Guarantees for Federated Large Models
by: Liu, Junkang, et al.
Published: (2025)
by: Liu, Junkang, et al.
Published: (2025)
Do Vendi Scores Converge with Finite Samples? Truncated Vendi Score for Finite-Sample Convergence Guarantees
by: Ospanov, Azim, et al.
Published: (2024)
by: Ospanov, Azim, et al.
Published: (2024)
Meta-Cognitive Reinforcement Learning with Self-Doubt and Recovery
by: Zhang, Zhipeng, et al.
Published: (2026)
by: Zhang, Zhipeng, et al.
Published: (2026)
Meta-Learning Reinforcement Learning for Crypto-Return Prediction
by: Wang, Junqiao, et al.
Published: (2025)
by: Wang, Junqiao, et al.
Published: (2025)
Theoretical Barriers in Bellman-Based Reinforcement Learning
by: Pinon, Brieuc, et al.
Published: (2025)
by: Pinon, Brieuc, et al.
Published: (2025)
When Do Skills Help Reinforcement Learning? A Theoretical Analysis of Temporal Abstractions
by: Li, Zhening, et al.
Published: (2024)
by: Li, Zhening, et al.
Published: (2024)
RL-STaR: Theoretical Analysis of Reinforcement Learning Frameworks for Self-Taught Reasoner
by: Chang, Fu-Chieh, et al.
Published: (2024)
by: Chang, Fu-Chieh, et al.
Published: (2024)
Game-Theoretic Robust Reinforcement Learning Handles Temporally-Coupled Perturbations
by: Liang, Yongyuan, et al.
Published: (2023)
by: Liang, Yongyuan, et al.
Published: (2023)
Enhancing Convolutional Neural Networks with Higher-Order Numerical Difference Methods
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
Similar Items
-
CA-BERT: Leveraging Context Awareness for Enhanced Multi-Turn Chat Interaction
by: Liu, Minghao, et al.
Published: (2024) -
Revisiting Meta-Learning with Noisy Labels: Reweighting Dynamics and Theoretical Guarantees
by: Zhang, Yiming, et al.
Published: (2025) -
Cooperative Backdoor Attack in Decentralized Reinforcement Learning with Theoretical Guarantee
by: Gao, Mengtong, et al.
Published: (2024) -
Spectral-Risk Safe Reinforcement Learning with Convergence Guarantees
by: Kim, Dohyeong, et al.
Published: (2024) -
A Theoretical Understanding of Gradient Bias in Meta-Reinforcement Learning
by: Feng, Xidong, et al.
Published: (2021)