On The Specialization of Neural Modules
Fuente:
arXiv
Saved in:
| Main Authors: | Jarvis, Devon, Klein, Richard, Rosman, Benjamin, Saxe, Andrew M. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Make Haste Slowly: A Theory of Emergent Structured Mixed Selectivity in Feature Learning ReLU Networks
by: Jarvis, Devon, et al.
Published: (2025)
by: Jarvis, Devon, et al.
Published: (2025)
Revisiting the Role of Relearning in Semantic Dementia
by: Jarvis, Devon, et al.
Published: (2025)
by: Jarvis, Devon, et al.
Published: (2025)
Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communities
by: Jarvis, Devon, et al.
Published: (2026)
by: Jarvis, Devon, et al.
Published: (2026)
Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
by: Tasse, Geraud Nangue, et al.
Published: (2025)
by: Tasse, Geraud Nangue, et al.
Published: (2025)
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
by: Tasse, Geraud Nangue, et al.
Published: (2022)
by: Tasse, Geraud Nangue, et al.
Published: (2022)
Distinct Computations Emerge From Compositional Curricula in In-Context Learning
by: Lee, Jin Hwa, et al.
Published: (2025)
by: Lee, Jin Hwa, et al.
Published: (2025)
Generalisable Agents for Neural Network Optimisation
by: Tessera, Kale-ab, et al.
Published: (2023)
by: Tessera, Kale-ab, et al.
Published: (2023)
Get rich quick: exact solutions reveal how unbalanced initializations promote rapid feature learning
by: Kunin, Daniel, et al.
Published: (2024)
by: Kunin, Daniel, et al.
Published: (2024)
On the Strengths and Weaknesses of Data for Open-set Embodied Assistance
by: Tambwekar, Pradyumna, et al.
Published: (2026)
by: Tambwekar, Pradyumna, et al.
Published: (2026)
A Theory of Initialisation's Impact on Specialisation
by: Jarvis, Devon, et al.
Published: (2025)
by: Jarvis, Devon, et al.
Published: (2025)
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
by: Rosen, Simon, et al.
Published: (2026)
by: Rosen, Simon, et al.
Published: (2026)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
by: Hori, Toshiaki, et al.
Published: (2025)
by: Hori, Toshiaki, et al.
Published: (2025)
Completed Hyperparameter Transfer across Modules, Width, Depth, Batch and Duration
by: Mlodozeniec, Bruno, et al.
Published: (2025)
by: Mlodozeniec, Bruno, et al.
Published: (2025)
GRACE: A Language Model Framework for Explainable Inverse Reinforcement Learning
by: Sapora, Silvia, et al.
Published: (2025)
by: Sapora, Silvia, et al.
Published: (2025)
The Interaction Bottleneck of Deep Neural Networks: Discovery, Proof, and Modulation
by: Deng, Huiqi, et al.
Published: (2025)
by: Deng, Huiqi, et al.
Published: (2025)
Sophisticated Learning: A novel algorithm for active learning during model-based planning
by: Hodson, Rowan, et al.
Published: (2023)
by: Hodson, Rowan, et al.
Published: (2023)
Inducing, Detecting and Characterising Neural Modules: A Pipeline for Functional Interpretability in Reinforcement Learning
by: Soligo, Anna, et al.
Published: (2025)
by: Soligo, Anna, et al.
Published: (2025)
Information-Theoretic Foundations for Neural Scaling Laws
by: Jeon, Hong Jun, et al.
Published: (2024)
by: Jeon, Hong Jun, et al.
Published: (2024)
Building Hybrid B-Spline And Neural Network Operators
by: Romagnoli, Raffaele, et al.
Published: (2024)
by: Romagnoli, Raffaele, et al.
Published: (2024)
Improving Reasoning Performance in Large Language Models via Representation Engineering
by: Højer, Bertram, et al.
Published: (2025)
by: Højer, Bertram, et al.
Published: (2025)
On Measuring Long-Range Interactions in Graph Neural Networks
by: Bamberger, Jacob, et al.
Published: (2025)
by: Bamberger, Jacob, et al.
Published: (2025)
Magnitude-Modulated Equivariant Adapter for Parameter-Efficient Fine-Tuning of Equivariant Graph Neural Networks
by: Jin, Dian, et al.
Published: (2025)
by: Jin, Dian, et al.
Published: (2025)
Deep Graph Neural Networks via Posteriori-Sampling-based Node-Adaptive Residual Module
by: Zhou, Jingbo, et al.
Published: (2023)
by: Zhou, Jingbo, et al.
Published: (2023)
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
by: Estermann, Benjamin, et al.
Published: (2024)
by: Estermann, Benjamin, et al.
Published: (2024)
An Operator-Consistent Graph Neural Network for Learning Diffusion Dynamics on Irregular Meshes
by: Li, Yuelian, et al.
Published: (2025)
by: Li, Yuelian, et al.
Published: (2025)
Simple and Effective Specialized Representations for Fair Classifiers
by: Sinigaglia, Alberto, et al.
Published: (2025)
by: Sinigaglia, Alberto, et al.
Published: (2025)
DataS^3: Dataset Subset Selection for Specialization
by: Hulkund, Neha, et al.
Published: (2025)
by: Hulkund, Neha, et al.
Published: (2025)
Layer Specialization Underlying Compositional Reasoning in Transformers
by: Liu, Jing
Published: (2025)
by: Liu, Jing
Published: (2025)
Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks
by: Lee, Su Hyeong, et al.
Published: (2025)
by: Lee, Su Hyeong, et al.
Published: (2025)
Transformers as Neural Operators for Solutions of Differential Equations with Finite Regularity
by: Shih, Benjamin, et al.
Published: (2024)
by: Shih, Benjamin, et al.
Published: (2024)
Prediction Is Not Physics: Learning and Evaluating Conserved Quantities in Neural Simulators
by: Bukowski, Andrew, et al.
Published: (2026)
by: Bukowski, Andrew, et al.
Published: (2026)
Uncertainty-Aware Deep Attention Recurrent Neural Network for Heterogeneous Time Series Imputation
by: Qian, Linglong, et al.
Published: (2024)
by: Qian, Linglong, et al.
Published: (2024)
Learning Dynamic Graph Embeddings with Neural Controlled Differential Equations
by: Qin, Tiexin, et al.
Published: (2023)
by: Qin, Tiexin, et al.
Published: (2023)
Deep Spatio-Temporal Neural Network for Air Quality Reanalysis
by: Kheder, Ammar, et al.
Published: (2025)
by: Kheder, Ammar, et al.
Published: (2025)
Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings
by: Bean, Andrew M., et al.
Published: (2025)
by: Bean, Andrew M., et al.
Published: (2025)
Dynamic Neural Control Flow Execution: An Agent-Based Deep Equilibrium Approach for Binary Vulnerability Detection
by: Li, Litao, et al.
Published: (2024)
by: Li, Litao, et al.
Published: (2024)
Exploring the Landscape for Generative Sequence Models for Specialized Data Synthesis
by: Zbeeb, Mohammad, et al.
Published: (2024)
by: Zbeeb, Mohammad, et al.
Published: (2024)
Beyond Redundancy: Diverse and Specialized Multi-Expert Sparse Autoencoder
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
SD-MoE: Spectral Decomposition for Effective Expert Specialization
by: Huang, Ruijun, et al.
Published: (2026)
by: Huang, Ruijun, et al.
Published: (2026)
Informed Spectral Normalized Gaussian Processes for Trajectory Prediction
by: Schlauch, Christian, et al.
Published: (2024)
by: Schlauch, Christian, et al.
Published: (2024)
Similar Items
-
Make Haste Slowly: A Theory of Emergent Structured Mixed Selectivity in Feature Learning ReLU Networks
by: Jarvis, Devon, et al.
Published: (2025) -
Revisiting the Role of Relearning in Semantic Dementia
by: Jarvis, Devon, et al.
Published: (2025) -
Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communities
by: Jarvis, Devon, et al.
Published: (2026) -
Beyond Sliding Windows: Learning to Manage Memory in Non-Markovian Environments
by: Tasse, Geraud Nangue, et al.
Published: (2025) -
Skill Machines: Temporal Logic Skill Composition in Reinforcement Learning
by: Tasse, Geraud Nangue, et al.
Published: (2022)