Maxwell's Demon at Work: Efficient Pruning by Leveraging Saturation of Neurons
Fuente:
arXiv
Salvato in:
| Autori principali: | Dufort-Labbé, Simon, D'Oro, Pierluca, Nikishin, Evgenii, Pascanu, Razvan, Bacon, Pierre-Luc, Baratin, Aristide |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026)
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026)
Navigating Potholes with Geometry-Aware Sharpness Minimization
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026)
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026)
The Curse of Diversity in Ensemble-Based Exploration
di: Lin, Zhixuan, et al.
Pubblicazione: (2024)
di: Lin, Zhixuan, et al.
Pubblicazione: (2024)
Mol-MoE: Training Preference-Guided Routers for Molecule Generation
di: Calanzone, Diego, et al.
Pubblicazione: (2025)
di: Calanzone, Diego, et al.
Pubblicazione: (2025)
Do Transformer World Models Give Better Policy Gradients?
di: Ma, Michel, et al.
Pubblicazione: (2024)
di: Ma, Michel, et al.
Pubblicazione: (2024)
Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control
di: Rahn, Nate, et al.
Pubblicazione: (2023)
di: Rahn, Nate, et al.
Pubblicazione: (2023)
ADEPTS: A Capability Framework for Human-Centered Agent Design
di: D'Oro, Pierluca, et al.
Pubblicazione: (2025)
di: D'Oro, Pierluca, et al.
Pubblicazione: (2025)
Torque-Aware Momentum
di: Malviya, Pranshu, et al.
Pubblicazione: (2024)
di: Malviya, Pranshu, et al.
Pubblicazione: (2024)
Promoting Exploration in Memory-Augmented Adam using Critical Momenta
di: Malviya, Pranshu, et al.
Pubblicazione: (2023)
di: Malviya, Pranshu, et al.
Pubblicazione: (2023)
Towards General-Purpose Model-Free Reinforcement Learning
di: Fujimoto, Scott, et al.
Pubblicazione: (2025)
di: Fujimoto, Scott, et al.
Pubblicazione: (2025)
MaestroMotif: Skill Design from Artificial Intelligence Feedback
di: Klissarov, Martin, et al.
Pubblicazione: (2024)
di: Klissarov, Martin, et al.
Pubblicazione: (2024)
Lattice: Learning to Efficiently Compress the Memory
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
di: Karami, Mahdi, et al.
Pubblicazione: (2025)
Controlling Large Language Model Agents with Entropic Activation Steering
di: Rahn, Nate, et al.
Pubblicazione: (2024)
di: Rahn, Nate, et al.
Pubblicazione: (2024)
Hierarchical Behaviour Spaces
di: Matthews, Michael Tryfan, et al.
Pubblicazione: (2026)
di: Matthews, Michael Tryfan, et al.
Pubblicazione: (2026)
Controlling Multimodal LLMs via Reward-guided Decoding
di: Mañas, Oscar, et al.
Pubblicazione: (2025)
di: Mañas, Oscar, et al.
Pubblicazione: (2025)
Lookbehind-SAM: k steps back, 1 step forward
di: Mordido, Gonçalo, et al.
Pubblicazione: (2023)
di: Mordido, Gonçalo, et al.
Pubblicazione: (2023)
Should I Have Expressed a Different Intent? Counterfactual Generation for LLM-Based Autonomous Control
di: Farzaneh, Amirmohammad, et al.
Pubblicazione: (2026)
di: Farzaneh, Amirmohammad, et al.
Pubblicazione: (2026)
Forgetting Transformer: Softmax Attention with a Forget Gate
di: Lin, Zhixuan, et al.
Pubblicazione: (2025)
di: Lin, Zhixuan, et al.
Pubblicazione: (2025)
Computer Use at the Edge of the Statistical Precipice
di: D'Oro, Pierluca, et al.
Pubblicazione: (2026)
di: D'Oro, Pierluca, et al.
Pubblicazione: (2026)
Meta-learning how to Share Credit among Macro-Actions
di: Hosu, Ionel-Alexandru, et al.
Pubblicazione: (2025)
di: Hosu, Ionel-Alexandru, et al.
Pubblicazione: (2025)
Revisiting Adam for Streaming Reinforcement Learning
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
di: Gogianu, Florin, et al.
Pubblicazione: (2026)
Perplexity Cannot Always Tell Right from Wrong
di: Veličković, Petar, et al.
Pubblicazione: (2026)
di: Veličković, Petar, et al.
Pubblicazione: (2026)
Maxwell's Demon
di: Kastner, R. E.
Pubblicazione: (2026)
di: Kastner, R. E.
Pubblicazione: (2026)
Softmax is not Enough (for Sharp Size Generalisation)
di: Veličković, Petar, et al.
Pubblicazione: (2024)
di: Veličković, Petar, et al.
Pubblicazione: (2024)
Lazy vs hasty: linearization in deep networks impacts learning schedule based on example difficulty
di: George, Thomas, et al.
Pubblicazione: (2022)
di: George, Thomas, et al.
Pubblicazione: (2022)
Fine-Tuned In-Context Learners for Efficient Adaptation
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
di: Bornschein, Jorg, et al.
Pubblicazione: (2025)
Beyond Connectivity: An Open Architecture for AI-RAN Convergence in 6G
di: Polese, Michele, et al.
Pubblicazione: (2025)
di: Polese, Michele, et al.
Pubblicazione: (2025)
Asynchronous Algorithmic Alignment with Cocycles
di: Dudzik, Andrew, et al.
Pubblicazione: (2023)
di: Dudzik, Andrew, et al.
Pubblicazione: (2023)
Hunting for Maxwell's Demon in the Wild
di: Buisson, Johan du, et al.
Pubblicazione: (2025)
di: Buisson, Johan du, et al.
Pubblicazione: (2025)
How connectivity structure shapes rich and lazy learning in neural circuits
di: Liu, Yuhan Helena, et al.
Pubblicazione: (2023)
di: Liu, Yuhan Helena, et al.
Pubblicazione: (2023)
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
di: Schmied, Thomas, et al.
Pubblicazione: (2025)
di: Schmied, Thomas, et al.
Pubblicazione: (2025)
State Soup: In-Context Skill Learning, Retrieval and Mixing
di: Pióro, Maciej, et al.
Pubblicazione: (2024)
di: Pióro, Maciej, et al.
Pubblicazione: (2024)
LibIQ: Toward Real-Time Spectrum Classification in O-RAN dApps
di: Olimpieri, Filippo, et al.
Pubblicazione: (2025)
di: Olimpieri, Filippo, et al.
Pubblicazione: (2025)
On AI Verification in Open RAN
di: Soundrarajan, Rahul, et al.
Pubblicazione: (2025)
di: Soundrarajan, Rahul, et al.
Pubblicazione: (2025)
The Three Regimes of Offline-to-Online Reinforcement Learning
di: Li, Lu, et al.
Pubblicazione: (2025)
di: Li, Lu, et al.
Pubblicazione: (2025)
Mining Generalizable Activation Functions
di: Vitvitskyi, Alex, et al.
Pubblicazione: (2026)
di: Vitvitskyi, Alex, et al.
Pubblicazione: (2026)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
di: Schmied, Thomas, et al.
Pubblicazione: (2024)
di: Schmied, Thomas, et al.
Pubblicazione: (2024)
A Text-Based Recommender System that Leverages Explicit Affective State Preferences
di: Hasan, Tonmoy, et al.
Pubblicazione: (2025)
di: Hasan, Tonmoy, et al.
Pubblicazione: (2025)
AgentRAN: An Agentic AI Architecture for Autonomous Control of Open 6G Networks
di: Elkael, Maxime, et al.
Pubblicazione: (2025)
di: Elkael, Maxime, et al.
Pubblicazione: (2025)
Maxwell's Demon and Irreversible Ideal Gas Expansion?
di: Ruggeri, Francesco R.
Pubblicazione: (2025)
di: Ruggeri, Francesco R.
Pubblicazione: (2025)
Documenti analoghi
-
Layerwise LQR for Geometry-Aware Optimization of Deep Networks
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026) -
Navigating Potholes with Geometry-Aware Sharpness Minimization
di: Dufort-Labbé, Simon, et al.
Pubblicazione: (2026) -
The Curse of Diversity in Ensemble-Based Exploration
di: Lin, Zhixuan, et al.
Pubblicazione: (2024) -
Mol-MoE: Training Preference-Guided Routers for Molecule Generation
di: Calanzone, Diego, et al.
Pubblicazione: (2025) -
Do Transformer World Models Give Better Policy Gradients?
di: Ma, Michel, et al.
Pubblicazione: (2024)