Neuroplastic Expansion in Deep Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jiashun, Obando-Ceron, Johan, Courville, Aaron, Pan, Ling |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025)
by: Mayor, Walter, et al.
Published: (2025)
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026)
by: Pasand, Ali Saheb, et al.
Published: (2026)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
by: Tang, Hongyao, et al.
Published: (2025)
by: Tang, Hongyao, et al.
Published: (2025)
Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
by: Castanyer, Roger Creus, et al.
Published: (2025)
by: Castanyer, Roger Creus, et al.
Published: (2025)
Adaptive Computation Pruning for the Forgetting Transformer
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
by: Sokar, Ghada, et al.
Published: (2024)
by: Sokar, Ghada, et al.
Published: (2024)
Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents
by: Obando-Ceron, Johan, et al.
Published: (2025)
by: Obando-Ceron, Johan, et al.
Published: (2025)
Learning Intractable Multimodal Policies with Reparameterization and Diversity Regularization
by: Wang, Ziqi, et al.
Published: (2025)
by: Wang, Ziqi, et al.
Published: (2025)
A Mechanistic Analysis of Looped Reasoning Language Models
by: Blayney, Hugh, et al.
Published: (2026)
by: Blayney, Hugh, et al.
Published: (2026)
Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning
by: Liu, Zihe, et al.
Published: (2025)
by: Liu, Zihe, et al.
Published: (2025)
Distributional GFlowNets with Quantile Flows
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
Complementary Reinforcement Learning
by: Muhtar, Dilxat, et al.
Published: (2026)
by: Muhtar, Dilxat, et al.
Published: (2026)
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
by: Shah, Vedant, et al.
Published: (2025)
by: Shah, Vedant, et al.
Published: (2025)
Mixture of Experts in a Mixture of RL settings
by: Willi, Timon, et al.
Published: (2024)
by: Willi, Timon, et al.
Published: (2024)
Versatile Energy-Based Probabilistic Models for High Energy Physics
by: Cheng, Taoli, et al.
Published: (2023)
by: Cheng, Taoli, et al.
Published: (2023)
Value-Based Deep Multi-Agent Reinforcement Learning with Dynamic Sparse Training
by: Hu, Pihe, et al.
Published: (2024)
by: Hu, Pihe, et al.
Published: (2024)
The Rank and Gradient Lost in Non-stationarity: Sample Weight Decay for Mitigating Plasticity Loss in Reinforcement Learning
by: Wu, Zihao, et al.
Published: (2026)
by: Wu, Zihao, et al.
Published: (2026)
Learning Robust Social Strategies with Large Language Models
by: Piche, Dereck, et al.
Published: (2025)
by: Piche, Dereck, et al.
Published: (2025)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
The Curse of Diversity in Ensemble-Based Exploration
by: Lin, Zhixuan, et al.
Published: (2024)
by: Lin, Zhixuan, et al.
Published: (2024)
LOQA: Learning with Opponent Q-Learning Awareness
by: Aghajohari, Milad, et al.
Published: (2024)
by: Aghajohari, Milad, et al.
Published: (2024)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
Recursive Self-Aggregation Unlocks Deep Thinking in Large Language Models
by: Venkatraman, Siddarth, et al.
Published: (2025)
by: Venkatraman, Siddarth, et al.
Published: (2025)
Not All LLM Reasoners Are Created Equal
by: Hosseini, Arian, et al.
Published: (2024)
by: Hosseini, Arian, et al.
Published: (2024)
Neuroplasticity-inspired dynamic ANNs for multi-task demand forecasting
by: Żarski, Mateusz, et al.
Published: (2025)
by: Żarski, Mateusz, et al.
Published: (2025)
Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training
by: Bartoldson, Brian, et al.
Published: (2025)
by: Bartoldson, Brian, et al.
Published: (2025)
MPCS: Neuroplastic Continual Learning via Multi-Component Plasticity and Topology-Aware EWC
by: Hentsch, Joern
Published: (2026)
by: Hentsch, Joern
Published: (2026)
Scattered Mixture-of-Experts Implementation
by: Tan, Shawn, et al.
Published: (2024)
by: Tan, Shawn, et al.
Published: (2024)
Fitting networks with a cancellation trick
by: Jin, Jiashun, et al.
Published: (2025)
by: Jin, Jiashun, et al.
Published: (2025)
Gradient-Based Neuroplastic Adaptation for Concurrent Optimization of Neuro-Fuzzy Networks
by: Hostetter, John Wesley, et al.
Published: (2025)
by: Hostetter, John Wesley, et al.
Published: (2025)
Neuroplasticity and Corruption in Model Mechanisms: A Case Study Of Indirect Object Identification
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
Advancing Out-of-Distribution Detection via Local Neuroplasticity
by: Canevaro, Alessandro, et al.
Published: (2025)
by: Canevaro, Alessandro, et al.
Published: (2025)
Network Topology Optimization via Deep Reinforcement Learning
by: Li, Zhuoran, et al.
Published: (2022)
by: Li, Zhuoran, et al.
Published: (2022)
Strategy and Skill Learning for Physics-based Table Tennis Animation
by: Wang, Jiashun, et al.
Published: (2024)
by: Wang, Jiashun, et al.
Published: (2024)
Forgetting Transformer: Softmax Attention with a Forget Gate
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Similar Items
-
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2025) -
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
by: Liu, Jiashun, et al.
Published: (2025) -
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025) -
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026) -
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)