Metalearning Continual Learning Algorithms
Fuente:
arXiv
Saved in:
| Main Authors: | Irie, Kazuki, Csordás, Róbert, Schmidhuber, Jürgen |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Self-Organising Neural Discrete Representation Learning à la Kohonen
by: Irie, Kazuki, et al.
Published: (2023)
by: Irie, Kazuki, et al.
Published: (2023)
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
by: Csordás, Róbert, et al.
Published: (2023)
by: Csordás, Róbert, et al.
Published: (2023)
MoEUT: Mixture-of-Experts Universal Transformers
by: Csordás, Róbert, et al.
Published: (2024)
by: Csordás, Róbert, et al.
Published: (2024)
Exploring the Promise and Limits of Real-Time Recurrent Learning
by: Irie, Kazuki, et al.
Published: (2023)
by: Irie, Kazuki, et al.
Published: (2023)
Mixture of Sparse Attention: Content-Based Learnable Sparse Attention via Expert-Choice Routing
by: Piękos, Piotr, et al.
Published: (2025)
by: Piękos, Piotr, et al.
Published: (2025)
Measuring In-Context Computation Complexity via Hidden State Prediction
by: Herrmann, Vincent, et al.
Published: (2025)
by: Herrmann, Vincent, et al.
Published: (2025)
Decoupling the "What" and "Where" With Polar Coordinate Positional Embeddings
by: Gopalakrishnan, Anand, et al.
Published: (2025)
by: Gopalakrishnan, Anand, et al.
Published: (2025)
Why Are Positional Encodings Nonessential for Deep Autoregressive Transformers? Revisiting a Petroglyph
by: Irie, Kazuki
Published: (2024)
by: Irie, Kazuki
Published: (2024)
Metalearners for Ranking Treatment Effects
by: Vanderschueren, Toon, et al.
Published: (2024)
by: Vanderschueren, Toon, et al.
Published: (2024)
Privacy in Metalearning and Multitask Learning: Modeling and Separations
by: Aliakbarpour, Maryam, et al.
Published: (2024)
by: Aliakbarpour, Maryam, et al.
Published: (2024)
IDSIA/recurrent-fwp: v1.0.0 - Public Code Release
by: Kazuki Irie, et al.
Published: (2025)
by: Kazuki Irie, et al.
Published: (2025)
Learning to Forget: Continual Learning with Adaptive Weight Decay
by: Ramesh, Aditya A., et al.
Published: (2026)
by: Ramesh, Aditya A., et al.
Published: (2026)
On the Necessity of Metalearning: Learning Suitable Parameterizations for Learning Processes
by: Hamidi, Massinissa, et al.
Published: (2023)
by: Hamidi, Massinissa, et al.
Published: (2023)
Navigating High Dimensional Concept Space with Metalearning
by: Gupta, Max
Published: (2025)
by: Gupta, Max
Published: (2025)
Dynamic Design of Machine Learning Pipelines via Metalearning
by: Alcobaça, Edesio, et al.
Published: (2025)
by: Alcobaça, Edesio, et al.
Published: (2025)
Overcoming classic challenges for artificial neural networks by providing incentives and practice
by: Irie, Kazuki, et al.
Published: (2024)
by: Irie, Kazuki, et al.
Published: (2024)
Fast weight programming and linear transformers: from machine learning to neurobiology
by: Irie, Kazuki, et al.
Published: (2025)
by: Irie, Kazuki, et al.
Published: (2025)
Blending Complementary Memory Systems in Hybrid Quadratic-Linear Transformers
by: Irie, Kazuki, et al.
Published: (2025)
by: Irie, Kazuki, et al.
Published: (2025)
Metalearning with Very Few Samples Per Task
by: Aliakbarpour, Maryam, et al.
Published: (2023)
by: Aliakbarpour, Maryam, et al.
Published: (2023)
Effective Regularization Through Loss-Function Metalearning
by: Gonzalez, Santiago, et al.
Published: (2020)
by: Gonzalez, Santiago, et al.
Published: (2020)
Metalearning traffic assignment for network disruptions with graph convolutional neural networks
by: Agriesti, Serio, et al.
Published: (2026)
by: Agriesti, Serio, et al.
Published: (2026)
Metalearning
by: Brazdil, Pavel, et al.
Published: (2022)
by: Brazdil, Pavel, et al.
Published: (2022)
Learning Useful Representations of Recurrent Neural Network Weight Matrices
by: Herrmann, Vincent, et al.
Published: (2024)
by: Herrmann, Vincent, et al.
Published: (2024)
Key-value memory in the brain
by: Gershman, Samuel J., et al.
Published: (2025)
by: Gershman, Samuel J., et al.
Published: (2025)
Interestingness as an Inductive Heuristic for Future Compression Progress
by: Herrmann, Vincent, et al.
Published: (2026)
by: Herrmann, Vincent, et al.
Published: (2026)
Why Depth Matters in Parallelizable Sequence Models: A Lie Algebraic View
by: Heo, Gyuryang, et al.
Published: (2026)
by: Heo, Gyuryang, et al.
Published: (2026)
Who invented deep residual learning?
by: Schmidhuber, Juergen
Published: (2025)
by: Schmidhuber, Juergen
Published: (2025)
Sequence Compression Speeds Up Credit Assignment in Reinforcement Learning
by: Ramesh, Aditya A., et al.
Published: (2024)
by: Ramesh, Aditya A., et al.
Published: (2024)
Recurrent Neural Networks Learn to Store and Generate Sequences using Non-Linear Representations
by: Csordás, Róbert, et al.
Published: (2024)
by: Csordás, Róbert, et al.
Published: (2024)
Do Language Models Use Their Depth Efficiently?
by: Csordás, Róbert, et al.
Published: (2025)
by: Csordás, Róbert, et al.
Published: (2025)
Fairness Overfitting in Machine Learning: An Information-Theoretic Perspective
by: Laakom, Firas, et al.
Published: (2025)
by: Laakom, Firas, et al.
Published: (2025)
PULSE: Practical Evaluation Scenarios for Large Multimodal Model Unlearning
by: Kawakami, Tatsuki, et al.
Published: (2025)
by: Kawakami, Tatsuki, et al.
Published: (2025)
Sequential-Parallel Duality in Prefix Scannable Models
by: Yau, Morris, et al.
Published: (2025)
by: Yau, Morris, et al.
Published: (2025)
Dissecting the Interplay of Attention Paths in a Statistical Mechanics Theory of Transformers
by: Tiberi, Lorenzo, et al.
Published: (2024)
by: Tiberi, Lorenzo, et al.
Published: (2024)
A Causal Analysis of CO2 Reduction Strategies in Electricity Markets Through Machine Learning-Driven Metalearners
by: Naeini, Iman Emtiazi, et al.
Published: (2024)
by: Naeini, Iman Emtiazi, et al.
Published: (2024)
Multiple Token Divergence: Measuring and Steering In-Context Computation Density
by: Herrmann, Vincent, et al.
Published: (2025)
by: Herrmann, Vincent, et al.
Published: (2025)
Recurrent Complex-Weighted Autoencoders for Unsupervised Object Discovery
by: Gopalakrishnan, Anand, et al.
Published: (2024)
by: Gopalakrishnan, Anand, et al.
Published: (2024)
Bayesian Optimization for Simultaneous Selection of Machine Learning Algorithms and Hyperparameters on Shared Latent Space
by: Ishikawa, Kazuki, et al.
Published: (2025)
by: Ishikawa, Kazuki, et al.
Published: (2025)
Thoughtbubbles: an Unsupervised Method for Parallel Thinking in Latent Space
by: Liu, Houjun, et al.
Published: (2025)
by: Liu, Houjun, et al.
Published: (2025)
FACTS: A Factored State-Space Framework For World Modelling
by: Nanbo, Li, et al.
Published: (2024)
by: Nanbo, Li, et al.
Published: (2024)
Similar Items
-
Self-Organising Neural Discrete Representation Learning à la Kohonen
by: Irie, Kazuki, et al.
Published: (2023) -
SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention
by: Csordás, Róbert, et al.
Published: (2023) -
MoEUT: Mixture-of-Experts Universal Transformers
by: Csordás, Róbert, et al.
Published: (2024) -
Exploring the Promise and Limits of Real-Time Recurrent Learning
by: Irie, Kazuki, et al.
Published: (2023) -
Mixture of Sparse Attention: Content-Based Learnable Sparse Attention via Expert-Choice Routing
by: Piękos, Piotr, et al.
Published: (2025)