Improvements of Dark Experience Replay and Reservoir Sampling towards Better Balance between Consolidation and Plasticity
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Kobayashi, Taisuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Experience Replayable Conditions
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
von: Kobayashi, Taisuke
Veröffentlicht: (2022)
Flexible Empowerment at Reasoning with Extended Best-of-N Sampling
von: Kobayashi, Taisuke
Veröffentlicht: (2026)
von: Kobayashi, Taisuke
Veröffentlicht: (2026)
Variational Adaptive Noise and Dropout towards Stable Recurrent Neural Networks
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2025)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2025)
Design of Restricted Normalizing Flow towards Arbitrary Stochastic Policy with Computational Efficiency
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)
Pseudo-Quantized Actor-Critic Algorithm for Robustness to Noisy Temporal Difference Error
von: Kobayashi, Taisuke
Veröffentlicht: (2026)
von: Kobayashi, Taisuke
Veröffentlicht: (2026)
DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
von: Kobayashi, Taisuke
Veröffentlicht: (2024)
CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
Intentionally-underestimated Value Function at Terminal State for Temporal-difference Learning with Mis-designed Reward
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
von: Kobayashi, Taisuke
Veröffentlicht: (2023)
Non-Uniform Memory Sampling in Experience Replay
von: Krutsylo, Andrii
Veröffentlicht: (2025)
von: Krutsylo, Andrii
Veröffentlicht: (2025)
Prior-free Balanced Replay: Uncertainty-guided Reservoir Sampling for Long-Tailed Continual Learning
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
Experience Replay Addresses Loss of Plasticity in Continual Learning
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
Contextual Experience Replay for Self-Improvement of Language Agents
von: Liu, Yitao, et al.
Veröffentlicht: (2025)
von: Liu, Yitao, et al.
Veröffentlicht: (2025)
Towards Autonomous Driving of Personal Mobility with Small and Noisy Dataset using Tsallis-statistics-based Behavioral Cloning
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2021)
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2021)
Uncertainty Prioritized Experience Replay
von: Carrasco-Davis, Rodrigo, et al.
Veröffentlicht: (2025)
von: Carrasco-Davis, Rodrigo, et al.
Veröffentlicht: (2025)
Sample Efficient Experience Replay in Non-stationary Environments
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
von: Duan, Tianyang, et al.
Veröffentlicht: (2025)
Reliability-Adjusted Prioritized Experience Replay
von: Pleiss, Leonard S., et al.
Veröffentlicht: (2025)
von: Pleiss, Leonard S., et al.
Veröffentlicht: (2025)
Maximum Entropy Hindsight Experience Replay
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
Better Generative Replay for Continual Federated Learning
von: Qi, Daiqing, et al.
Veröffentlicht: (2023)
von: Qi, Daiqing, et al.
Veröffentlicht: (2023)
Experience Replay with Random Reshuffling
von: Fujita, Yasuhiro
Veröffentlicht: (2025)
von: Fujita, Yasuhiro
Veröffentlicht: (2025)
VLM-Guided Experience Replay
von: Sharony, Elad, et al.
Veröffentlicht: (2026)
von: Sharony, Elad, et al.
Veröffentlicht: (2026)
Variance Reduction via Resampling and Experience Replay
von: Han, Jiale, et al.
Veröffentlicht: (2025)
von: Han, Jiale, et al.
Veröffentlicht: (2025)
Efficient RL Training for LLMs with Experience Replay
von: Arnal, Charles, et al.
Veröffentlicht: (2026)
von: Arnal, Charles, et al.
Veröffentlicht: (2026)
On the Limitation and Experience Replay for GNNs in Continual Learning
von: Su, Junwei, et al.
Veröffentlicht: (2023)
von: Su, Junwei, et al.
Veröffentlicht: (2023)
Hindsight Experience Replay Accelerates Proximal Policy Optimization
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
von: Crowder, Douglas C., et al.
Veröffentlicht: (2024)
Variance Reduction Based Experience Replay for Policy Optimization
von: Zheng, Hua, et al.
Veröffentlicht: (2026)
von: Zheng, Hua, et al.
Veröffentlicht: (2026)
A Tighter Convergence Proof of Reverse Experience Replay
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
ROER: Regularized Optimal Experience Replay
von: Li, Changling, et al.
Veröffentlicht: (2024)
von: Li, Changling, et al.
Veröffentlicht: (2024)
Preliminary Tests of the Anticipatory Classifier System with Hindsight Experience Replay
von: Unold, Olgierd, et al.
Veröffentlicht: (2026)
von: Unold, Olgierd, et al.
Veröffentlicht: (2026)
Balance-aware Sequence Sampling Makes Multi-modal Learning Better
von: Guan, Zhi-Hao
Veröffentlicht: (2025)
von: Guan, Zhi-Hao
Veröffentlicht: (2025)
Neuron-level Balance between Stability and Plasticity in Deep Reinforcement Learning
von: Lan, Jiahua, et al.
Veröffentlicht: (2025)
von: Lan, Jiahua, et al.
Veröffentlicht: (2025)
Reservoir Network with Structural Plasticity for Human Activity Recognition
von: Zyarah, Abdullah M., et al.
Veröffentlicht: (2025)
von: Zyarah, Abdullah M., et al.
Veröffentlicht: (2025)
Sample-efficient LLM Optimization with Reset Replay
von: Liu, Zichuan, et al.
Veröffentlicht: (2025)
von: Liu, Zichuan, et al.
Veröffentlicht: (2025)
Balancing Synthetic Data and Replay for Enhancing Task-Specific Capabilities
von: Spiegelhalter, Urs, et al.
Veröffentlicht: (2025)
von: Spiegelhalter, Urs, et al.
Veröffentlicht: (2025)
Mastering the Game of Go with Self-play Experience Replay
von: Liu, Jingbin, et al.
Veröffentlicht: (2026)
von: Liu, Jingbin, et al.
Veröffentlicht: (2026)
AdaER: An Adaptive Experience Replay Approach for Continual Lifelong Learning
von: Li, Xingyu, et al.
Veröffentlicht: (2023)
von: Li, Xingyu, et al.
Veröffentlicht: (2023)
Enabling On-Device Learning via Experience Replay with Efficient Dataset Condensation
von: Xu, Gelei, et al.
Veröffentlicht: (2024)
von: Xu, Gelei, et al.
Veröffentlicht: (2024)
Sharper Bounds for $\ell_p$ Sensitivity Sampling
von: Woodruff, David P., et al.
Veröffentlicht: (2023)
von: Woodruff, David P., et al.
Veröffentlicht: (2023)
Weber-Fechner Law in Temporal Difference learning derived from Control as Inference
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
von: Takahashi, Keiichiro, et al.
Veröffentlicht: (2024)
Ridge Leverage Score Sampling for $\ell_p$ Subspace Approximation
von: Woodruff, David P., et al.
Veröffentlicht: (2024)
von: Woodruff, David P., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Revisiting Experience Replayable Conditions
von: Kobayashi, Taisuke
Veröffentlicht: (2024) -
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning
von: Kobayashi, Taisuke
Veröffentlicht: (2022) -
Flexible Empowerment at Reasoning with Extended Best-of-N Sampling
von: Kobayashi, Taisuke
Veröffentlicht: (2026) -
Variational Adaptive Noise and Dropout towards Stable Recurrent Neural Networks
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2025) -
Design of Restricted Normalizing Flow towards Arbitrary Stochastic Policy with Computational Efficiency
von: Kobayashi, Taisuke, et al.
Veröffentlicht: (2024)