On the creation of narrow AI: hierarchy and nonlocality of neural network skills
Fuente:
arXiv
Salvato in:
| Autori principali: | Michaud, Eric J., Parker-Sartori, Asher, Tegmark, Max |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Quantization Model of Neural Scaling
di: Michaud, Eric J., et al.
Pubblicazione: (2023)
di: Michaud, Eric J., et al.
Pubblicazione: (2023)
Not All Language Model Features Are One-Dimensionally Linear
di: Engels, Joshua, et al.
Pubblicazione: (2024)
di: Engels, Joshua, et al.
Pubblicazione: (2024)
Efficient Dictionary Learning with Switch Sparse Autoencoders
di: Mudide, Anish, et al.
Pubblicazione: (2024)
di: Mudide, Anish, et al.
Pubblicazione: (2024)
Survival of the Fittest Representation: A Case Study with Modular Addition
di: Ding, Xiaoman Delores, et al.
Pubblicazione: (2024)
di: Ding, Xiaoman Delores, et al.
Pubblicazione: (2024)
Physics of Skill Learning
di: Liu, Ziming, et al.
Pubblicazione: (2025)
di: Liu, Ziming, et al.
Pubblicazione: (2025)
Do Two AI Scientists Agree?
di: Fu, Xinghong, et al.
Pubblicazione: (2025)
di: Fu, Xinghong, et al.
Pubblicazione: (2025)
The Geometry of Concepts: Sparse Autoencoder Feature Structure
di: Li, Yuxiao, et al.
Pubblicazione: (2024)
di: Li, Yuxiao, et al.
Pubblicazione: (2024)
OptPDE: Discovering Novel Integrable Systems via AI-Human Collaboration
di: Kantamneni, Subhash, et al.
Pubblicazione: (2024)
di: Kantamneni, Subhash, et al.
Pubblicazione: (2024)
Opening the AI black box: program synthesis via mechanistic interpretability
di: Michaud, Eric J., et al.
Pubblicazione: (2024)
di: Michaud, Eric J., et al.
Pubblicazione: (2024)
Towards Understanding Distilled Reasoning Models: A Representational Approach
di: Baek, David D., et al.
Pubblicazione: (2025)
di: Baek, David D., et al.
Pubblicazione: (2025)
Harmonic Loss Trains Interpretable AI Models
di: Baek, David D., et al.
Pubblicazione: (2025)
di: Baek, David D., et al.
Pubblicazione: (2025)
When narrower is better: the narrow width limit of Bayesian parallel branching neural networks
di: Zhang, Zechen, et al.
Pubblicazione: (2024)
di: Zhang, Zechen, et al.
Pubblicazione: (2024)
Language Models Use Trigonometry to Do Addition
di: Kantamneni, Subhash, et al.
Pubblicazione: (2025)
di: Kantamneni, Subhash, et al.
Pubblicazione: (2025)
Language Models Represent Space and Time
di: Gurnee, Wes, et al.
Pubblicazione: (2023)
di: Gurnee, Wes, et al.
Pubblicazione: (2023)
High-dimensional learning of narrow neural networks
di: Cui, Hugo
Pubblicazione: (2024)
di: Cui, Hugo
Pubblicazione: (2024)
Low-Rank Adapting Models for Sparse Autoencoders
di: Chen, Matthew, et al.
Pubblicazione: (2025)
di: Chen, Matthew, et al.
Pubblicazione: (2025)
Decomposing The Dark Matter of Sparse Autoencoders
di: Engels, Joshua, et al.
Pubblicazione: (2024)
di: Engels, Joshua, et al.
Pubblicazione: (2024)
Universal approximation with complex-valued deep narrow neural networks
di: Geuchen, Paul, et al.
Pubblicazione: (2023)
di: Geuchen, Paul, et al.
Pubblicazione: (2023)
A Neural Scaling Law from Lottery Ticket Ensembling
di: Liu, Ziming, et al.
Pubblicazione: (2023)
di: Liu, Ziming, et al.
Pubblicazione: (2023)
Improved weight initialization for deep and narrow feedforward neural network
di: Lee, Hyunwoo, et al.
Pubblicazione: (2023)
di: Lee, Hyunwoo, et al.
Pubblicazione: (2023)
GenEFT: Understanding Statics and Dynamics of Model Generalization via Effective Theory
di: Baek, David D., et al.
Pubblicazione: (2024)
di: Baek, David D., et al.
Pubblicazione: (2024)
Investigating Representation Universality: Case Study on Genealogical Representations
di: Baek, David D., et al.
Pubblicazione: (2024)
di: Baek, David D., et al.
Pubblicazione: (2024)
Generating adversarial inputs for a graph neural network model of AC power flow
di: Parker, Robert
Pubblicazione: (2026)
di: Parker, Robert
Pubblicazione: (2026)
Luck, skill, and depth of competition in games and social hierarchies
di: Jerdee, Maximilian, et al.
Pubblicazione: (2023)
di: Jerdee, Maximilian, et al.
Pubblicazione: (2023)
How Do Transformers "Do" Physics? Investigating the Simple Harmonic Oscillator
di: Kantamneni, Subhash, et al.
Pubblicazione: (2024)
di: Kantamneni, Subhash, et al.
Pubblicazione: (2024)
Association-sensory spatiotemporal hierarchy and functional gradient-regularised recurrent neural network with implications for schizophrenia
di: Abulikemu, Subati, et al.
Pubblicazione: (2025)
di: Abulikemu, Subati, et al.
Pubblicazione: (2025)
An in-depth look at approximation via deep and narrow neural networks
di: Dommel, Joris, et al.
Pubblicazione: (2025)
di: Dommel, Joris, et al.
Pubblicazione: (2025)
A Resource Model For Neural Scaling Law
di: Song, Jinyeop, et al.
Pubblicazione: (2024)
di: Song, Jinyeop, et al.
Pubblicazione: (2024)
Understanding sparse autoencoder scaling in the presence of feature manifolds
di: Michaud, Eric J., et al.
Pubblicazione: (2025)
di: Michaud, Eric J., et al.
Pubblicazione: (2025)
Are Sparse Autoencoders Useful? A Case Study in Sparse Probing
di: Kantamneni, Subhash, et al.
Pubblicazione: (2025)
di: Kantamneni, Subhash, et al.
Pubblicazione: (2025)
Scaling Laws For Scalable Oversight
di: Engels, Joshua, et al.
Pubblicazione: (2025)
di: Engels, Joshua, et al.
Pubblicazione: (2025)
The Remarkable Robustness of LLMs: Stages of Inference?
di: Lad, Vedang, et al.
Pubblicazione: (2024)
di: Lad, Vedang, et al.
Pubblicazione: (2024)
Neural Thermodynamic Laws for Large Language Model Training
di: Liu, Ziming, et al.
Pubblicazione: (2025)
di: Liu, Ziming, et al.
Pubblicazione: (2025)
Formulations and scalability of neural network surrogates in nonlinear optimization problems
di: Parker, Robert B., et al.
Pubblicazione: (2024)
di: Parker, Robert B., et al.
Pubblicazione: (2024)
Scaling transformer neural networks for skillful and reliable medium-range weather forecasting
di: Nguyen, Tung, et al.
Pubblicazione: (2023)
di: Nguyen, Tung, et al.
Pubblicazione: (2023)
A Case for Library-Level k-Means Binning in Histogram Gradient-Boosted Trees
di: Labovich, Asher
Pubblicazione: (2025)
di: Labovich, Asher
Pubblicazione: (2025)
Foundation Models for AI-Enabled Biological Design
di: Moldwin, Asher, et al.
Pubblicazione: (2025)
di: Moldwin, Asher, et al.
Pubblicazione: (2025)
Can neural networks do arithmetic? A survey on the elementary numerical skills of state-of-the-art deep learning models
di: Testolin, Alberto
Pubblicazione: (2023)
di: Testolin, Alberto
Pubblicazione: (2023)
Fixed points of nonnegative neural networks
di: Piotrowski, Tomasz J., et al.
Pubblicazione: (2021)
di: Piotrowski, Tomasz J., et al.
Pubblicazione: (2021)
The Artificial Intelligence Ontology: LLM-assisted construction of AI concept hierarchies
di: Joachimiak, Marcin P., et al.
Pubblicazione: (2024)
di: Joachimiak, Marcin P., et al.
Pubblicazione: (2024)
Documenti analoghi
-
The Quantization Model of Neural Scaling
di: Michaud, Eric J., et al.
Pubblicazione: (2023) -
Not All Language Model Features Are One-Dimensionally Linear
di: Engels, Joshua, et al.
Pubblicazione: (2024) -
Efficient Dictionary Learning with Switch Sparse Autoencoders
di: Mudide, Anish, et al.
Pubblicazione: (2024) -
Survival of the Fittest Representation: A Case Study with Modular Addition
di: Ding, Xiaoman Delores, et al.
Pubblicazione: (2024) -
Physics of Skill Learning
di: Liu, Ziming, et al.
Pubblicazione: (2025)