In value-based deep reinforcement learning, a pruned network is a good network
Fuente:
arXiv
Saved in:
| Main Authors: | Obando-Ceron, Johan, Courville, Aaron, Castro, Pablo Samuel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
by: Sokar, Ghada, et al.
Published: (2024)
by: Sokar, Ghada, et al.
Published: (2024)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025)
by: Mayor, Walter, et al.
Published: (2025)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
by: Tang, Hongyao, et al.
Published: (2025)
by: Tang, Hongyao, et al.
Published: (2025)
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026)
by: Pasand, Ali Saheb, et al.
Published: (2026)
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents
by: Obando-Ceron, Johan, et al.
Published: (2025)
by: Obando-Ceron, Johan, et al.
Published: (2025)
A Mechanistic Analysis of Looped Reasoning Language Models
by: Blayney, Hugh, et al.
Published: (2026)
by: Blayney, Hugh, et al.
Published: (2026)
Adaptive Computation Pruning for the Forgetting Transformer
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Mixture of Experts in a Mixture of RL settings
by: Willi, Timon, et al.
Published: (2024)
by: Willi, Timon, et al.
Published: (2024)
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
by: Shah, Vedant, et al.
Published: (2025)
by: Shah, Vedant, et al.
Published: (2025)
Elimination-compensation pruning for fully-connected neural networks
by: Ballini, Enrico, et al.
Published: (2026)
by: Ballini, Enrico, et al.
Published: (2026)
Active management of battery degradation in wireless sensor network using deep reinforcement learning for group battery replacement
by: Jeong, Jong-Hyun, et al.
Published: (2025)
by: Jeong, Jong-Hyun, et al.
Published: (2025)
An approach of deep reinforcement learning for maximizing the net present value of stochastic projects
by: Xu, Wei, et al.
Published: (2025)
by: Xu, Wei, et al.
Published: (2025)
Neuroplastic Expansion in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2024)
by: Liu, Jiashun, et al.
Published: (2024)
Economic span selection of bridge based on deep reinforcement learning
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
by: Ly, Adrian, et al.
Published: (2025)
by: Ly, Adrian, et al.
Published: (2025)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
APF+: Boosting adaptive-potential function reinforcement learning methods with a W-shaped network for high-dimensional games
by: Chen, Yifei, et al.
Published: (2025)
by: Chen, Yifei, et al.
Published: (2025)
Designing a double deep reinforcement learning selection tool for resilient demand prediction
by: Benziane, Bilel Abderrahmane, et al.
Published: (2026)
by: Benziane, Bilel Abderrahmane, et al.
Published: (2026)
Stacking-based deep neural network for player scouting in football 1
by: Lacan, Simon
Published: (2024)
by: Lacan, Simon
Published: (2024)
An efficient deep reinforcement learning environment for flexible job-shop scheduling
by: Wu, Xinquan, et al.
Published: (2025)
by: Wu, Xinquan, et al.
Published: (2025)
Self-adaptive weights based on balanced residual decay rate for physics-informed neural networks and deep operator networks
by: Chen, Wenqian, et al.
Published: (2024)
by: Chen, Wenqian, et al.
Published: (2024)
Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
by: Castanyer, Roger Creus, et al.
Published: (2025)
by: Castanyer, Roger Creus, et al.
Published: (2025)
On convex decision regions in deep network representations
by: Tětková, Lenka, et al.
Published: (2023)
by: Tětková, Lenka, et al.
Published: (2023)
Extreme value forecasting using relevance-based data augmentation with deep learning models
by: Hua, Junru, et al.
Published: (2025)
by: Hua, Junru, et al.
Published: (2025)
A transformer-based deep reinforcement learning approach to spatial navigation in a partially observable Morris Water Maze
by: Eggen, Marte, et al.
Published: (2024)
by: Eggen, Marte, et al.
Published: (2024)
CST-AFNet: A dual attention-based deep learning framework for intrusion detection in IoT networks
by: Ishtiaq, Waqas, et al.
Published: (2025)
by: Ishtiaq, Waqas, et al.
Published: (2025)
Utilizing Lyapunov Exponents in designing deep neural networks
by: Mittra, Tirthankar
Published: (2024)
by: Mittra, Tirthankar
Published: (2024)
Forgetting Transformer: Softmax Attention with a Forget Gate
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Towards modeling evolving longitudinal health trajectories with a transformer-based deep learning model
by: Moen, Hans, et al.
Published: (2024)
by: Moen, Hans, et al.
Published: (2024)
STARFISH: faST Accuracy Recovery in pruned networks From Internal State Healing
by: Maon, Shir, et al.
Published: (2026)
by: Maon, Shir, et al.
Published: (2026)
Mind the GAP! The Challenges of Scale in Pixel-based Deep Reinforcement Learning
by: Sokar, Ghada, et al.
Published: (2025)
by: Sokar, Ghada, et al.
Published: (2025)
Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization
by: Zhang, Dinghuai, et al.
Published: (2023)
by: Zhang, Dinghuai, et al.
Published: (2023)
RDI: An adversarial robustness evaluation metric for deep neural networks based on model statistical features
by: Song, Jialei, et al.
Published: (2025)
by: Song, Jialei, et al.
Published: (2025)
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
by: Mirbakhsh, Shahin, et al.
Published: (2024)
by: Mirbakhsh, Shahin, et al.
Published: (2024)
Solving the flexible job-shop scheduling problem through an enhanced deep reinforcement learning approach
by: Echeverria, Imanol, et al.
Published: (2023)
by: Echeverria, Imanol, et al.
Published: (2023)
Similar Items
-
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024) -
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
by: Sokar, Ghada, et al.
Published: (2024) -
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025) -
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
by: Tang, Hongyao, et al.
Published: (2025) -
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations
by: Pasand, Ali Saheb, et al.
Published: (2026)