Mastering the Game of Go with Self-play Experience Replay
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Jingbin, Wang, Xuechun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VLM-Guided Experience Replay
di: Sharony, Elad, et al.
Pubblicazione: (2026)
di: Sharony, Elad, et al.
Pubblicazione: (2026)
Experience Replay with Random Reshuffling
di: Fujita, Yasuhiro
Pubblicazione: (2025)
di: Fujita, Yasuhiro
Pubblicazione: (2025)
ROER: Regularized Optimal Experience Replay
di: Li, Changling, et al.
Pubblicazione: (2024)
di: Li, Changling, et al.
Pubblicazione: (2024)
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning
di: Zhao, Kaiyan, et al.
Pubblicazione: (2024)
di: Zhao, Kaiyan, et al.
Pubblicazione: (2024)
Contextual Experience Replay for Self-Improvement of Language Agents
di: Liu, Yitao, et al.
Pubblicazione: (2025)
di: Liu, Yitao, et al.
Pubblicazione: (2025)
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
di: Lim, Han-Dong, et al.
Pubblicazione: (2023)
Generalized Back-Stepping Experience Replay in Sparse-Reward Environments
di: Lyu, Guwen, et al.
Pubblicazione: (2024)
di: Lyu, Guwen, et al.
Pubblicazione: (2024)
CIER: A Novel Experience Replay Approach with Causal Inference in Deep Reinforcement Learning
di: Wang, Jingwen, et al.
Pubblicazione: (2024)
di: Wang, Jingwen, et al.
Pubblicazione: (2024)
Enhancing LLM Agents for Code Generation with Possibility and Pass-rate Prioritized Experience Replay
di: Chen, Yuyang, et al.
Pubblicazione: (2024)
di: Chen, Yuyang, et al.
Pubblicazione: (2024)
Enabling Option Learning in Sparse Rewards with Hindsight Experience Replay
di: Romio, Gabriel, et al.
Pubblicazione: (2026)
di: Romio, Gabriel, et al.
Pubblicazione: (2026)
Experience Replay Addresses Loss of Plasticity in Continual Learning
di: Wang, Jiuqi, et al.
Pubblicazione: (2025)
di: Wang, Jiuqi, et al.
Pubblicazione: (2025)
Learning to play: A Multimodal Agent for 3D Game-Play
di: Yue, Yuguang, et al.
Pubblicazione: (2025)
di: Yue, Yuguang, et al.
Pubblicazione: (2025)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
di: Liu, Jinyi, et al.
Pubblicazione: (2023)
di: Liu, Jinyi, et al.
Pubblicazione: (2023)
On the Convergence of Experience Replay in Policy Optimization: Characterizing Bias, Variance, and Finite-Time Convergence
di: Zheng, Hua, et al.
Pubblicazione: (2021)
di: Zheng, Hua, et al.
Pubblicazione: (2021)
Sample Efficient Experience Replay in Non-stationary Environments
di: Duan, Tianyang, et al.
Pubblicazione: (2025)
di: Duan, Tianyang, et al.
Pubblicazione: (2025)
CUER: Corrected Uniform Experience Replay for Off-Policy Continuous Deep Reinforcement Learning Algorithms
di: Yenicesu, Arda Sarp, et al.
Pubblicazione: (2024)
di: Yenicesu, Arda Sarp, et al.
Pubblicazione: (2024)
DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
di: Perkins, Daniel, et al.
Pubblicazione: (2025)
di: Perkins, Daniel, et al.
Pubblicazione: (2025)
Mastering the Game of Guandan with Deep Reinforcement Learning and Behavior Regulating
di: Yanggong, Yifan, et al.
Pubblicazione: (2024)
di: Yanggong, Yifan, et al.
Pubblicazione: (2024)
IDER: IDempotent Experience Replay for Reliable Continual Learning
di: Liu, Zhanwang, et al.
Pubblicazione: (2026)
di: Liu, Zhanwang, et al.
Pubblicazione: (2026)
Self-Attribution Bias: When AI Monitors Go Easy on Themselves
di: Khullar, Dipika, et al.
Pubblicazione: (2026)
di: Khullar, Dipika, et al.
Pubblicazione: (2026)
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
di: Li, Keyu, et al.
Pubblicazione: (2021)
di: Li, Keyu, et al.
Pubblicazione: (2021)
Self-generated Replay Memories for Continual Neural Machine Translation
di: Resta, Michele, et al.
Pubblicazione: (2024)
di: Resta, Michele, et al.
Pubblicazione: (2024)
CodeIt: Self-Improving Language Models with Prioritized Hindsight Replay
di: Butt, Natasha, et al.
Pubblicazione: (2024)
di: Butt, Natasha, et al.
Pubblicazione: (2024)
Investigating the Interplay of Prioritized Replay and Generalization
di: Panahi, Parham Mohammad, et al.
Pubblicazione: (2024)
di: Panahi, Parham Mohammad, et al.
Pubblicazione: (2024)
D-SPEAR: Dual-Stream Prioritized Experience Adaptive Replay for Stable Reinforcement Learning in Robotic Manipulation
di: Zhang, Yu, et al.
Pubblicazione: (2026)
di: Zhang, Yu, et al.
Pubblicazione: (2026)
Continual Offline Reinforcement Learning via Diffusion-based Dual Generative Replay
di: Liu, Jinmei, et al.
Pubblicazione: (2024)
di: Liu, Jinmei, et al.
Pubblicazione: (2024)
IRIS: Interpolative Rényi Iterative Self-play for Large Language Model Fine-Tuning
di: Liao, Wenjie, et al.
Pubblicazione: (2026)
di: Liao, Wenjie, et al.
Pubblicazione: (2026)
Absolute Zero: Reinforced Self-play Reasoning with Zero Data
di: Zhao, Andrew, et al.
Pubblicazione: (2025)
di: Zhao, Andrew, et al.
Pubblicazione: (2025)
MiniZero: Comparative Analysis of AlphaZero and MuZero on Go, Othello, and Atari Games
di: Wu, Ti-Rong, et al.
Pubblicazione: (2023)
di: Wu, Ti-Rong, et al.
Pubblicazione: (2023)
PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning
di: Dong, Daize, et al.
Pubblicazione: (2026)
di: Dong, Daize, et al.
Pubblicazione: (2026)
Catastrophic Forgetting Mitigation via Discrepancy-Weighted Experience Replay
di: Xu, Xinrun, et al.
Pubblicazione: (2025)
di: Xu, Xinrun, et al.
Pubblicazione: (2025)
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
di: Wang, Xi, et al.
Pubblicazione: (2026)
di: Wang, Xi, et al.
Pubblicazione: (2026)
ARROW: Augmented Replay for RObust World models
di: Alyahya, Abdulaziz, et al.
Pubblicazione: (2026)
di: Alyahya, Abdulaziz, et al.
Pubblicazione: (2026)
Better Generative Replay for Continual Federated Learning
di: Qi, Daiqing, et al.
Pubblicazione: (2023)
di: Qi, Daiqing, et al.
Pubblicazione: (2023)
R^3: Replay, Reflection, and Ranking Rewards for LLM Reinforcement Learning
di: Jiang, Zhizheng, et al.
Pubblicazione: (2026)
di: Jiang, Zhizheng, et al.
Pubblicazione: (2026)
EFRame: Deeper Reasoning via Exploration-Filter-Replay Reinforcement Learning Framework
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
di: Song, Chihyeon, et al.
Pubblicazione: (2025)
di: Song, Chihyeon, et al.
Pubblicazione: (2025)
The Effectiveness of Approximate Regularized Replay for Efficient Supervised Fine-Tuning of Large Language Models
di: Riemer, Matthew, et al.
Pubblicazione: (2025)
di: Riemer, Matthew, et al.
Pubblicazione: (2025)
Task-Aware Adaptive Modulation: A Replay-Free and Resource-Efficient Approach For Continual Graph Learning
di: Liu, Jingtao, et al.
Pubblicazione: (2025)
di: Liu, Jingtao, et al.
Pubblicazione: (2025)
Adaptable Hindsight Experience Replay for Search-Based Learning
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
di: Vazaios, Alexandros, et al.
Pubblicazione: (2025)
Documenti analoghi
-
VLM-Guided Experience Replay
di: Sharony, Elad, et al.
Pubblicazione: (2026) -
Experience Replay with Random Reshuffling
di: Fujita, Yasuhiro
Pubblicazione: (2025) -
ROER: Regularized Optimal Experience Replay
di: Li, Changling, et al.
Pubblicazione: (2024) -
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning
di: Zhao, Kaiyan, et al.
Pubblicazione: (2024) -
Contextual Experience Replay for Self-Improvement of Language Agents
di: Liu, Yitao, et al.
Pubblicazione: (2025)