The Curse of Diversity in Ensemble-Based Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Zhixuan, D'Oro, Pierluca, Nikishin, Evgenii, Courville, Aaron |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Forgetting Transformer: Softmax Attention with a Forget Gate
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
by: Dufort-Labbé, Simon, et al.
Published: (2024)
by: Dufort-Labbé, Simon, et al.
Published: (2024)
Mol-MoE: Training Preference-Guided Routers for Molecule Generation
by: Calanzone, Diego, et al.
Published: (2025)
by: Calanzone, Diego, et al.
Published: (2025)
ADEPTS: A Capability Framework for Human-Centered Agent Design
by: D'Oro, Pierluca, et al.
Published: (2025)
by: D'Oro, Pierluca, et al.
Published: (2025)
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)
by: Fujimoto, Scott, et al.
Published: (2025)
Do Transformer World Models Give Better Policy Gradients?
by: Ma, Michel, et al.
Published: (2024)
by: Ma, Michel, et al.
Published: (2024)
Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control
by: Rahn, Nate, et al.
Published: (2023)
by: Rahn, Nate, et al.
Published: (2023)
Hierarchical Behaviour Spaces
by: Matthews, Michael Tryfan, et al.
Published: (2026)
by: Matthews, Michael Tryfan, et al.
Published: (2026)
Controlling Multimodal LLMs via Reward-guided Decoding
by: Mañas, Oscar, et al.
Published: (2025)
by: Mañas, Oscar, et al.
Published: (2025)
Adaptive Computation Pruning for the Forgetting Transformer
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
Versatile Energy-Based Probabilistic Models for High Energy Physics
by: Cheng, Taoli, et al.
Published: (2023)
by: Cheng, Taoli, et al.
Published: (2023)
Controlling Large Language Model Agents with Entropic Activation Steering
by: Rahn, Nate, et al.
Published: (2024)
by: Rahn, Nate, et al.
Published: (2024)
MaestroMotif: Skill Design from Artificial Intelligence Feedback
by: Klissarov, Martin, et al.
Published: (2024)
by: Klissarov, Martin, et al.
Published: (2024)
Neuroplastic Expansion in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2024)
by: Liu, Jiashun, et al.
Published: (2024)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
Not All LLM Reasoners Are Created Equal
by: Hosseini, Arian, et al.
Published: (2024)
by: Hosseini, Arian, et al.
Published: (2024)
SCALEX: Scalable Concept and Latent Exploration for Diffusion Models
by: Zeng, E. Zhixuan, et al.
Published: (2025)
by: Zeng, E. Zhixuan, et al.
Published: (2025)
Understanding the Curse of Unrolling
by: Mehmood, Sheheryar, et al.
Published: (2026)
by: Mehmood, Sheheryar, et al.
Published: (2026)
Scattered Mixture-of-Experts Implementation
by: Tan, Shawn, et al.
Published: (2024)
by: Tan, Shawn, et al.
Published: (2024)
Constrained Ensemble Exploration for Unsupervised Skill Discovery
by: Bai, Chenjia, et al.
Published: (2024)
by: Bai, Chenjia, et al.
Published: (2024)
Bias Analysis in Unconditional Image Generative Models
by: Zhang, Xiaofeng, et al.
Published: (2025)
by: Zhang, Xiaofeng, et al.
Published: (2025)
Is BatchEnsemble a Single Model? On Calibration and Diversity of Efficient Ensembles
by: Zamyatin, Anton, et al.
Published: (2026)
by: Zamyatin, Anton, et al.
Published: (2026)
Pathologies of Predictive Diversity in Deep Ensembles
by: Abe, Taiga, et al.
Published: (2023)
by: Abe, Taiga, et al.
Published: (2023)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025)
by: Mayor, Walter, et al.
Published: (2025)
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
Learning Robust Social Strategies with Large Language Models
by: Piche, Dereck, et al.
Published: (2025)
by: Piche, Dereck, et al.
Published: (2025)
Modeling Caption Diversity in Contrastive Vision-Language Pretraining
by: Lavoie, Samuel, et al.
Published: (2024)
by: Lavoie, Samuel, et al.
Published: (2024)
Dispelling the Curse of Singularities in Neural Network Optimizations
by: Cao, Hengjie, et al.
Published: (2026)
by: Cao, Hengjie, et al.
Published: (2026)
Unleashing Diverse Thinking Modes in LLMs through Multi-Agent Collaboration
by: He, Zhixuan, et al.
Published: (2025)
by: He, Zhixuan, et al.
Published: (2025)
The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse and More
by: Kitouni, Ouail, et al.
Published: (2024)
by: Kitouni, Ouail, et al.
Published: (2024)
World Modelling Improves Language Model Agents
by: Guo, Shangmin, et al.
Published: (2025)
by: Guo, Shangmin, et al.
Published: (2025)
Curse of Dimensionality in Neural Network Optimization
by: Na, Sanghoon, et al.
Published: (2025)
by: Na, Sanghoon, et al.
Published: (2025)
The Blessing and Curse of Dimensionality in Safety Alignment
by: Teo, Rachel S. Y., et al.
Published: (2025)
by: Teo, Rachel S. Y., et al.
Published: (2025)
The Curse of Depth in Large Language Models
by: Sun, Wenfang, et al.
Published: (2025)
by: Sun, Wenfang, et al.
Published: (2025)
Scaling Stick-Breaking Attention: An Efficient Implementation and In-depth Study
by: Tan, Shawn, et al.
Published: (2024)
by: Tan, Shawn, et al.
Published: (2024)
BiXSE: Improving Dense Retrieval via Probabilistic Graded Relevance Distillation
by: Tsirigotis, Christos, et al.
Published: (2025)
by: Tsirigotis, Christos, et al.
Published: (2025)
Tempo: Confidentiality Preservation in Cloud-Based Neural Network Training
by: Xu, Rongwu, et al.
Published: (2024)
by: Xu, Rongwu, et al.
Published: (2024)
Anant-Net: Breaking the Curse of Dimensionality with Scalable and Interpretable Neural Surrogate for High-Dimensional PDEs
by: Menon, Sidharth S., et al.
Published: (2025)
by: Menon, Sidharth S., et al.
Published: (2025)
Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
Similar Items
-
Forgetting Transformer: Softmax Attention with a Forget Gate
by: Lin, Zhixuan, et al.
Published: (2025) -
Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
by: Dufort-Labbé, Simon, et al.
Published: (2024) -
Mol-MoE: Training Preference-Guided Routers for Molecule Generation
by: Calanzone, Diego, et al.
Published: (2025) -
ADEPTS: A Capability Framework for Human-Centered Agent Design
by: D'Oro, Pierluca, et al.
Published: (2025) -
Towards General-Purpose Model-Free Reinforcement Learning
by: Fujimoto, Scott, et al.
Published: (2025)