Guardado en:
| Autores principales: | Lufkin, Leon, Figliolia, Tomás, Millidge, Beren, Krishnamurthy, Kamesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2603.22325 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Online Vector Quantized Attention
por: Alonso, Nick, et al.
Publicado: (2026)
por: Alonso, Nick, et al.
Publicado: (2026)
Toward Conversational Agents with Context and Time Sensitive Long-term Memory
por: Alonso, Nick, et al.
Publicado: (2024)
por: Alonso, Nick, et al.
Publicado: (2024)
Compressed Convolutional Attention: Efficient Attention in a Compressed Latent Space
por: Figliolia, Tomas, et al.
Publicado: (2025)
por: Figliolia, Tomas, et al.
Publicado: (2025)
Zamba: A Compact 7B SSM Hybrid Model
por: Glorioso, Paolo, et al.
Publicado: (2024)
por: Glorioso, Paolo, et al.
Publicado: (2024)
ZUNA: Flexible EEG Superresolution with Position-Aware Diffusion Autoencoders
por: Warner, Christopher, et al.
Publicado: (2026)
por: Warner, Christopher, et al.
Publicado: (2026)
BlackMamba: Mixture of Experts for State-Space Models
por: Anthony, Quentin, et al.
Publicado: (2024)
por: Anthony, Quentin, et al.
Publicado: (2024)
Generalising E-prop to Deep Networks
por: Millidge, Beren
Publicado: (2025)
por: Millidge, Beren
Publicado: (2025)
Associative Memories in the Feature Space
por: Salvatori, Tommaso, et al.
Publicado: (2024)
por: Salvatori, Tommaso, et al.
Publicado: (2024)
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG
por: Alonso, Nicholas, et al.
Publicado: (2024)
por: Alonso, Nicholas, et al.
Publicado: (2024)
The Zamba2 Suite: Technical Report
por: Glorioso, Paolo, et al.
Publicado: (2024)
por: Glorioso, Paolo, et al.
Publicado: (2024)
A Stable, Fast, and Fully Automatic Learning Algorithm for Predictive Coding Networks
por: Salvatori, Tommaso, et al.
Publicado: (2022)
por: Salvatori, Tommaso, et al.
Publicado: (2022)
A Group Theoretic Analysis of the Symmetries Underlying Base Addition and Their Learnability by Neural Networks
por: Dawes, Cutter, et al.
Publicado: (2025)
por: Dawes, Cutter, et al.
Publicado: (2025)
Zyda-2: a 5 Trillion Token High-Quality Dataset
por: Tokpanov, Yury, et al.
Publicado: (2024)
por: Tokpanov, Yury, et al.
Publicado: (2024)
Exploring Action-Centric Representations Through the Lens of Rate-Distortion Theory
por: Varona, Miguel de Llanza, et al.
Publicado: (2024)
por: Varona, Miguel de Llanza, et al.
Publicado: (2024)
Mamba4Cast: Efficient Zero-Shot Time Series Forecasting with State Space Models
por: Bhethanabhotla, Sathya Kamesh, et al.
Publicado: (2024)
por: Bhethanabhotla, Sathya Kamesh, et al.
Publicado: (2024)
ZAYA1-VL-8B Technical Report
por: Shapourian, Hassan, et al.
Publicado: (2026)
por: Shapourian, Hassan, et al.
Publicado: (2026)
Learning Associative Memories with Gradient Descent
por: Cabannes, Vivien, et al.
Publicado: (2024)
por: Cabannes, Vivien, et al.
Publicado: (2024)
Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs
por: Lin, Jun-Liang, et al.
Publicado: (2026)
por: Lin, Jun-Liang, et al.
Publicado: (2026)
Understanding Transformer from the Perspective of Associative Memory
por: Zhong, Shu, et al.
Publicado: (2025)
por: Zhong, Shu, et al.
Publicado: (2025)
Hybrid Self-evolving Structured Memory for GUI Agents
por: Zhu, Sibo, et al.
Publicado: (2026)
por: Zhu, Sibo, et al.
Publicado: (2026)
Tree Attention: Topology-aware Decoding for Long-Context Attention on GPU clusters
por: Shyam, Vasudev, et al.
Publicado: (2024)
por: Shyam, Vasudev, et al.
Publicado: (2024)
The Role of Environment Access in Agnostic Reinforcement Learning
por: Krishnamurthy, Akshay, et al.
Publicado: (2025)
por: Krishnamurthy, Akshay, et al.
Publicado: (2025)
Achieving the Tightest Relaxation of Sigmoids for Formal Verification
por: Chevalier, Samuel, et al.
Publicado: (2024)
por: Chevalier, Samuel, et al.
Publicado: (2024)
Predictive Coding beyond Correlations
por: Salvatori, Tommaso, et al.
Publicado: (2023)
por: Salvatori, Tommaso, et al.
Publicado: (2023)
Tensor Cache: Eviction-conditioned Associative Memory for Transformers
por: Swain, Kabir, et al.
Publicado: (2026)
por: Swain, Kabir, et al.
Publicado: (2026)
Re:Frame -- Retrieving Experience From Associative Memory
por: Zelezetsky, Daniil, et al.
Publicado: (2025)
por: Zelezetsky, Daniil, et al.
Publicado: (2025)
A Review of Neuroscience-Inspired Machine Learning
por: Ororbia, Alexander, et al.
Publicado: (2024)
por: Ororbia, Alexander, et al.
Publicado: (2024)
Large-scale Time-Varying Portfolio Optimisation using Graph Attention Networks
por: Korangi, Kamesh, et al.
Publicado: (2024)
por: Korangi, Kamesh, et al.
Publicado: (2024)
BroadGen: A Framework for Generating Effective and Efficient Advertiser Broad Match Keyphrase Recommendations
por: Mishra, Ashirbad, et al.
Publicado: (2025)
por: Mishra, Ashirbad, et al.
Publicado: (2025)
Equivalence of Personalized PageRank and Successor Representations
por: Millidge, Beren
Publicado: (2025)
por: Millidge, Beren
Publicado: (2025)
Adversarial Robustness of VAEs across Intersectional Subgroups
por: Ramanaik, Chethan Krishnamurthy, et al.
Publicado: (2024)
por: Ramanaik, Chethan Krishnamurthy, et al.
Publicado: (2024)
Nonlinear dynamics of localization in neural receptive fields
por: Lufkin, Leon, et al.
Publicado: (2025)
por: Lufkin, Leon, et al.
Publicado: (2025)
LLMs as High-Dimensional Nonlinear Autoregressive Models with Attention: Training, Alignment and Inference
por: Krishnamurthy, Vikram
Publicado: (2026)
por: Krishnamurthy, Vikram
Publicado: (2026)
In-Memory Learning Automata Architecture using Y-Flash Cell
por: Ghazal, Omar, et al.
Publicado: (2024)
por: Ghazal, Omar, et al.
Publicado: (2024)
A Unifying View of Coverage in Linear Off-Policy Evaluation
por: Amortila, Philip, et al.
Publicado: (2026)
por: Amortila, Philip, et al.
Publicado: (2026)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
por: Liu, Zeyuan, et al.
Publicado: (2026)
por: Liu, Zeyuan, et al.
Publicado: (2026)
Robust Bidirectional Associative Memory via Regularization Inspired by the Subspace Rotation Algorithm
por: Lin, Ci, et al.
Publicado: (2025)
por: Lin, Ci, et al.
Publicado: (2025)
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory
por: Li, Sijia, et al.
Publicado: (2025)
por: Li, Sijia, et al.
Publicado: (2025)
Memory in Plain Sight: Surveying the Uncanny Resemblances of Associative Memories and Diffusion Models
por: Hoover, Benjamin, et al.
Publicado: (2023)
por: Hoover, Benjamin, et al.
Publicado: (2023)
Learning Hidden Markov Models Using Conditional Samples
por: Kakade, Sham M., et al.
Publicado: (2023)
por: Kakade, Sham M., et al.
Publicado: (2023)
Ejemplares similares
-
Online Vector Quantized Attention
por: Alonso, Nick, et al.
Publicado: (2026) -
Toward Conversational Agents with Context and Time Sensitive Long-term Memory
por: Alonso, Nick, et al.
Publicado: (2024) -
Compressed Convolutional Attention: Efficient Attention in a Compressed Latent Space
por: Figliolia, Tomas, et al.
Publicado: (2025) -
Zamba: A Compact 7B SSM Hybrid Model
por: Glorioso, Paolo, et al.
Publicado: (2024) -
ZUNA: Flexible EEG Superresolution with Position-Aware Diffusion Autoencoders
por: Warner, Christopher, et al.
Publicado: (2026)