Enabling Robust In-Context Memory and Rapid Task Adaptation in Transformers with Hebbian and Gradient-Based Plasticity
Fuente:
arXiv
Saved in:
| Main Author: | Chaudhary, Siddharth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SNAP: Stopping Catastrophic Forgetting in Hebbian Learning with Sigmoidal Neuronal Adaptive Plasticity
by: Xu, Tianyi, et al.
Published: (2024)
by: Xu, Tianyi, et al.
Published: (2024)
Learning Successor Features with Distributed Hebbian Temporal Memory
by: Dzhivelikian, Evgenii, et al.
Published: (2023)
by: Dzhivelikian, Evgenii, et al.
Published: (2023)
Neuron-centric Hebbian Learning
by: Ferigo, Andrea, et al.
Published: (2024)
by: Ferigo, Andrea, et al.
Published: (2024)
A Truly Sparse and General Implementation of Gradient-Based Synaptic Plasticity
by: Lohoff, Jamie, et al.
Published: (2025)
by: Lohoff, Jamie, et al.
Published: (2025)
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
Hebbian Learning based Orthogonal Projection for Continual Learning of Spiking Neural Networks
by: Xiao, Mingqing, et al.
Published: (2024)
by: Xiao, Mingqing, et al.
Published: (2024)
Understanding Transformer Optimization via Gradient Heterogeneity
by: Tomihari, Akiyoshi, et al.
Published: (2025)
by: Tomihari, Akiyoshi, et al.
Published: (2025)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
by: Bossens, David M.
Published: (2023)
by: Bossens, David M.
Published: (2023)
General-Purpose In-Context Learning by Meta-Learning Transformers
by: Kirsch, Louis, et al.
Published: (2022)
by: Kirsch, Louis, et al.
Published: (2022)
Evaluating Open-Source Sparse Autoencoders on Disentangling Factual Knowledge in GPT-2 Small
by: Chaudhary, Maheep, et al.
Published: (2024)
by: Chaudhary, Maheep, et al.
Published: (2024)
Provably Optimal Memory Capacity for Modern Hopfield Models: Transformer-Compatible Dense Associative Memories as Spherical Codes
by: Hu, Jerry Yao-Chieh, et al.
Published: (2024)
by: Hu, Jerry Yao-Chieh, et al.
Published: (2024)
Ken Utilization Layer: Hebbian Replay Within a Student's Ken for Adaptive Exercise Recommendation
by: Kuling, Grey, et al.
Published: (2025)
by: Kuling, Grey, et al.
Published: (2025)
Learning in Spiking Neural Networks with a Calcium-based Hebbian Rule for Spike-timing-dependent Plasticity
by: Girão, Willian Soares, et al.
Published: (2025)
by: Girão, Willian Soares, et al.
Published: (2025)
MTSpark: Enabling Multi-Task Learning with Spiking Neural Networks for Generalist Agents
by: Devkota, Avaneesh, et al.
Published: (2024)
by: Devkota, Avaneesh, et al.
Published: (2024)
Experience Replay Addresses Loss of Plasticity in Continual Learning
by: Wang, Jiuqi, et al.
Published: (2025)
by: Wang, Jiuqi, et al.
Published: (2025)
RMAAT: Astrocyte-Inspired Memory Compression and Replay for Efficient Long-Context Transformers
by: Mia, Md Zesun Ahmed, et al.
Published: (2026)
by: Mia, Md Zesun Ahmed, et al.
Published: (2026)
The Alpha-Alternator: Dynamic Adaptation To Varying Noise Levels In Sequences Using The Vendi Score For Improved Robustness and Performance
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
Predicting Deterioration in Mild Cognitive Impairment with Survival Transformers, Extreme Gradient Boosting and Cox Proportional Hazard Modelling
by: Musto, Henry, et al.
Published: (2024)
by: Musto, Henry, et al.
Published: (2024)
Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning
by: Lillo, Lute, et al.
Published: (2026)
by: Lillo, Lute, et al.
Published: (2026)
Memory Mosaics
by: Zhang, Jianyu, et al.
Published: (2024)
by: Zhang, Jianyu, et al.
Published: (2024)
Multi-Task Optimization over Networks of Tasks
by: Hatzky, Julian, et al.
Published: (2026)
by: Hatzky, Julian, et al.
Published: (2026)
Hebbian-Oscillatory Co-Learning
by: Hays, Hasi
Published: (2026)
by: Hays, Hasi
Published: (2026)
A Genetic Algorithm-Based Approach for Automated Optimization of Kolmogorov-Arnold Networks in Classification Tasks
by: Long, Quan, et al.
Published: (2025)
by: Long, Quan, et al.
Published: (2025)
Fast Fourier Transform-Based Spectral and Temporal Gradient Filtering for Differential Privacy
by: Shin, Hyeju, et al.
Published: (2025)
by: Shin, Hyeju, et al.
Published: (2025)
Task and Explanation Network
by: Sipper, Moshe
Published: (2024)
by: Sipper, Moshe
Published: (2024)
Bridging Models to Defend: A Population-Based Strategy for Robust Adversarial Defense
by: Wang, Ren, et al.
Published: (2023)
by: Wang, Ren, et al.
Published: (2023)
AlphaGrad: Non-Linear Gradient Normalization Optimizer
by: Sane, Soham
Published: (2025)
by: Sane, Soham
Published: (2025)
Reinforced In-Context Black-Box Optimization
by: Song, Lei, et al.
Published: (2024)
by: Song, Lei, et al.
Published: (2024)
Memory Networks: Towards Fully Biologically Plausible Learning
by: Ruiz, Jacobo, et al.
Published: (2024)
by: Ruiz, Jacobo, et al.
Published: (2024)
Instance-Conditioned Adaptation for Large-scale Generalization of Neural Routing Solver
by: Zhou, Changliang, et al.
Published: (2024)
by: Zhou, Changliang, et al.
Published: (2024)
Long-Sequence Memory with Temporal Kernels and Dense Hopfield Functionals
by: Farooq, Ahmed
Published: (2025)
by: Farooq, Ahmed
Published: (2025)
Modality-Dependent Memory Mechanisms in Cross-Modal Neuromorphic Computing
by: Blessing, Effiong, et al.
Published: (2025)
by: Blessing, Effiong, et al.
Published: (2025)
Attending to Graph Transformers
by: Müller, Luis, et al.
Published: (2023)
by: Müller, Luis, et al.
Published: (2023)
Task-free Adaptive Meta Black-box Optimization
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies
by: Patankar, Dhruv, et al.
Published: (2026)
by: Patankar, Dhruv, et al.
Published: (2026)
Advancing Direct Training for Spiking Neural Networks with Circulate-Firing Neurons and Learnable Gradients
by: Zhou, Feifan, et al.
Published: (2026)
by: Zhou, Feifan, et al.
Published: (2026)
Graph Memory Learning: Imitating Lifelong Remembering and Forgetting of Brain Networks
by: Miao, Jiaxing, et al.
Published: (2024)
by: Miao, Jiaxing, et al.
Published: (2024)
TACOS: Task Agnostic Continual Learning in Spiking Neural Networks
by: Soures, Nicholas, et al.
Published: (2024)
by: Soures, Nicholas, et al.
Published: (2024)
Bridging Synthetic and Real Routing Problems via LLM-Guided Instance Generation and Progressive Adaptation
by: Zhu, Jianghan, et al.
Published: (2025)
by: Zhu, Jianghan, et al.
Published: (2025)
Gradient-Free Continual Learning in Spiking Neural Networks via Inter-Spike Interval Regularization
by: Roy, Samrendra, et al.
Published: (2026)
by: Roy, Samrendra, et al.
Published: (2026)
Similar Items
-
SNAP: Stopping Catastrophic Forgetting in Hebbian Learning with Sigmoidal Neuronal Adaptive Plasticity
by: Xu, Tianyi, et al.
Published: (2024) -
Learning Successor Features with Distributed Hebbian Temporal Memory
by: Dzhivelikian, Evgenii, et al.
Published: (2023) -
Neuron-centric Hebbian Learning
by: Ferigo, Andrea, et al.
Published: (2024) -
A Truly Sparse and General Implementation of Gradient-Based Synaptic Plasticity
by: Lohoff, Jamie, et al.
Published: (2025) -
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)