Scaling can lead to compositional generalization
Fuente:
arXiv
Guardado en:
| Autores principales: | Redhardt, Florian, Akram, Yassir, Schug, Simon |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
When can transformers compositionally generalize in-context?
por: Kobayashi, Seijin, et al.
Publicado: (2024)
por: Kobayashi, Seijin, et al.
Publicado: (2024)
Discovering modular solutions that generalize compositionally
por: Schug, Simon, et al.
Publicado: (2023)
por: Schug, Simon, et al.
Publicado: (2023)
Gated recurrent neural networks discover attention
por: Zucchet, Nicolas, et al.
Publicado: (2023)
por: Zucchet, Nicolas, et al.
Publicado: (2023)
Forward Direct Feedback Alignment for Online Gradient Estimates of Spiking Neural Networks
por: Bacho, Florian, et al.
Publicado: (2024)
por: Bacho, Florian, et al.
Publicado: (2024)
Scale-covariant spiking wavelets
por: Pedersen, Jens Egholm, et al.
Publicado: (2026)
por: Pedersen, Jens Egholm, et al.
Publicado: (2026)
Gradient-based inference of abstract task representations for generalization in neural networks
por: Hummos, Ali, et al.
Publicado: (2024)
por: Hummos, Ali, et al.
Publicado: (2024)
Scaling Equilibrium Propagation to Deeper Neural Network Architectures
por: Elayedam, Sankar Vinayak, et al.
Publicado: (2025)
por: Elayedam, Sankar Vinayak, et al.
Publicado: (2025)
Scaling Down Deep Learning with MNIST-1D
por: Greydanus, Sam, et al.
Publicado: (2020)
por: Greydanus, Sam, et al.
Publicado: (2020)
Effects of Introducing Synaptic Scaling on Spiking Neural Network Learning
por: Touda, Shinnosuke, et al.
Publicado: (2026)
por: Touda, Shinnosuke, et al.
Publicado: (2026)
Hierarchical Kernel Transformer: Multi-Scale Attention with an Information-Theoretic Approximation Analysis
por: Cirrincione, Giansalvo
Publicado: (2026)
por: Cirrincione, Giansalvo
Publicado: (2026)
Auto-Configured Networks for Multi-Scale Multi-Output Time-Series Forecasting
por: Zha, Yumeng, et al.
Publicado: (2026)
por: Zha, Yumeng, et al.
Publicado: (2026)
Universality of Real Minimal Complexity Reservoir
por: Fong, Robert Simon, et al.
Publicado: (2024)
por: Fong, Robert Simon, et al.
Publicado: (2024)
HEATACO: Heatmap-Guided Ant Colony Decoding for Large-Scale Travelling Salesman Problems
por: Lin, Bo-Cheng, et al.
Publicado: (2026)
por: Lin, Bo-Cheng, et al.
Publicado: (2026)
Predictive Modeling in the Reservoir Kernel Motif Space
por: Tino, Peter, et al.
Publicado: (2024)
por: Tino, Peter, et al.
Publicado: (2024)
Reservoir Computing via Multi-Scale Random Fourier Features for Forecasting Fast-Slow Dynamical Systems
por: Laha, S. K.
Publicado: (2025)
por: Laha, S. K.
Publicado: (2025)
GT-SNT: A Linear-Time Transformer for Large-Scale Graphs via Spiking Node Tokenization
por: Zhang, Huizhe, et al.
Publicado: (2025)
por: Zhang, Huizhe, et al.
Publicado: (2025)
Structuring Multiple Simple Cycle Reservoirs with Particle Swarm Optimization
por: Li, Ziqiang, et al.
Publicado: (2025)
por: Li, Ziqiang, et al.
Publicado: (2025)
State-space models can learn in-context by gradient descent
por: Sushma, Neeraj Mohan, et al.
Publicado: (2024)
por: Sushma, Neeraj Mohan, et al.
Publicado: (2024)
NeuroPareto: Calibrated Acquisition for Costly Many-Goal Search in Vast Parameter Spaces
por: Fu, Rong, et al.
Publicado: (2026)
por: Fu, Rong, et al.
Publicado: (2026)
Painful intelligence: What AI can tell us about human suffering
por: Hyvärinen, Aapo
Publicado: (2022)
por: Hyvärinen, Aapo
Publicado: (2022)
Cancer-inspired Genomics Mapper Model for the Generation of Synthetic DNA Sequences with Desired Genomics Signatures
por: Lazebnik, Teddy, et al.
Publicado: (2023)
por: Lazebnik, Teddy, et al.
Publicado: (2023)
Dynamical similarity analysis can identify compositional dynamics developing in RNNs
por: Guilhot, Quentin, et al.
Publicado: (2024)
por: Guilhot, Quentin, et al.
Publicado: (2024)
Why all roads don't lead to Rome: Representation geometry varies across the human visual cortical hierarchy
por: Ghosh, Arna, et al.
Publicado: (2025)
por: Ghosh, Arna, et al.
Publicado: (2025)
Random Features Hopfield Networks generalize retrieval to previously unseen examples
por: Kalaj, Silvio, et al.
Publicado: (2024)
por: Kalaj, Silvio, et al.
Publicado: (2024)
Polyra Swarms: A Shape-Based Approach to Machine Learning
por: Klüttermann, Simon, et al.
Publicado: (2025)
por: Klüttermann, Simon, et al.
Publicado: (2025)
QUIVER: Cost-Aware Adaptive Preference Querying in Surrogate-Assisted Evolutionary Multi-Objective Optimization
por: Burnat, Florian A. D.
Publicado: (2026)
por: Burnat, Florian A. D.
Publicado: (2026)
Dropout Universality: Scaling Laws and Optimal Scheduling at the Edge-of-Chaos
por: Sarmiento, Lucas Fernandez
Publicado: (2026)
por: Sarmiento, Lucas Fernandez
Publicado: (2026)
SpiNNaker2: A Large-Scale Neuromorphic System for Event-Based and Asynchronous Machine Learning
por: Gonzalez, Hector A., et al.
Publicado: (2024)
por: Gonzalez, Hector A., et al.
Publicado: (2024)
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
por: Majumdar, Somshubra, et al.
Publicado: (2024)
por: Majumdar, Somshubra, et al.
Publicado: (2024)
Increasing biases can be more efficient than increasing weights
por: Metta, Carlo, et al.
Publicado: (2023)
por: Metta, Carlo, et al.
Publicado: (2023)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
por: Xing, Xingrun, et al.
Publicado: (2024)
por: Xing, Xingrun, et al.
Publicado: (2024)
Towards Constraint-Based Adaptive Hypergraph Learning for Solving Vehicle Routing: An End-to-End Solution
por: Wang, Zhenwei, et al.
Publicado: (2025)
por: Wang, Zhenwei, et al.
Publicado: (2025)
PlatMetaX: An Integrated MATLAB platform for Meta-Black-Box Optimization
por: Yang, Xu, et al.
Publicado: (2025)
por: Yang, Xu, et al.
Publicado: (2025)
Neuro-Evolutionary Approach to Physics-Aware Symbolic Regression
por: Kubalík, Jiří, et al.
Publicado: (2025)
por: Kubalík, Jiří, et al.
Publicado: (2025)
Topology-Aware Activation Functions in Neural Networks
por: Snopov, Pavel, et al.
Publicado: (2025)
por: Snopov, Pavel, et al.
Publicado: (2025)
Sequential Multi-Agent Dynamic Algorithm Configuration
por: Lu, Chen, et al.
Publicado: (2025)
por: Lu, Chen, et al.
Publicado: (2025)
A Complete Pipeline for deploying SNNs with Synaptic Delays on Loihi 2
por: Mészáros, Balázs, et al.
Publicado: (2025)
por: Mészáros, Balázs, et al.
Publicado: (2025)
Application-oriented automatic hyperparameter optimization for spiking neural network prototyping
por: Fra, Vittorio
Publicado: (2025)
por: Fra, Vittorio
Publicado: (2025)
Spiking Brain Compression: Post-Training Second-order Compression for Spiking Neural Networks
por: Shi, Lianfeng, et al.
Publicado: (2025)
por: Shi, Lianfeng, et al.
Publicado: (2025)
Evolutionary Optimization for the Classification of Small Molecules Regulating the Circadian Rhythm Period: A Reliable Assessment
por: Arauzo-Azofra, Antonio, et al.
Publicado: (2025)
por: Arauzo-Azofra, Antonio, et al.
Publicado: (2025)
Ejemplares similares
-
When can transformers compositionally generalize in-context?
por: Kobayashi, Seijin, et al.
Publicado: (2024) -
Discovering modular solutions that generalize compositionally
por: Schug, Simon, et al.
Publicado: (2023) -
Gated recurrent neural networks discover attention
por: Zucchet, Nicolas, et al.
Publicado: (2023) -
Forward Direct Feedback Alignment for Online Gradient Estimates of Spiking Neural Networks
por: Bacho, Florian, et al.
Publicado: (2024) -
Scale-covariant spiking wavelets
por: Pedersen, Jens Egholm, et al.
Publicado: (2026)