Guardado en:
| Autor principal: | Chenebaux, Maixent |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.24809 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MKA: Memory-Keyed Attention for Efficient Long-Context Reasoning
por: Liu, Dong, et al.
Publicado: (2026)
por: Liu, Dong, et al.
Publicado: (2026)
When Reasoning Meets Compression: Understanding the Effects of LLMs Compression on Large Reasoning Models
por: Zhang, Nan, et al.
Publicado: (2025)
por: Zhang, Nan, et al.
Publicado: (2025)
Pay Attention to Small Weights
por: Zhou, Chao, et al.
Publicado: (2025)
por: Zhou, Chao, et al.
Publicado: (2025)
Paged Attention Meets FlexAttention: Unlocking Long-Context Efficiency in Deployed Inference
por: Joshi, Thomas, et al.
Publicado: (2025)
por: Joshi, Thomas, et al.
Publicado: (2025)
Probing the Limits of Compressive Memory: A Study of Infini-Attention in Small-Scale Pretraining
por: Huang, Ruizhe, et al.
Publicado: (2025)
por: Huang, Ruizhe, et al.
Publicado: (2025)
Predicting LLM Reasoning Performance with Small Proxy Model
por: Koh, Woosung, et al.
Publicado: (2025)
por: Koh, Woosung, et al.
Publicado: (2025)
Large Language Models Meet Graph Neural Networks for Text-Numeric Graph Reasoning
por: Song, Haoran, et al.
Publicado: (2025)
por: Song, Haoran, et al.
Publicado: (2025)
Enhancing Reasoning with Collaboration and Memory
por: Michelman, Julie, et al.
Publicado: (2025)
por: Michelman, Julie, et al.
Publicado: (2025)
Quantization Meets Reasoning: Exploring and Mitigating Degradation of Low-Bit LLMs in Mathematical Reasoning
por: Li, Zhen, et al.
Publicado: (2025)
por: Li, Zhen, et al.
Publicado: (2025)
Adaptive Memory Decay for Log-Linear Attention
por: Amin, Yaxita, et al.
Publicado: (2026)
por: Amin, Yaxita, et al.
Publicado: (2026)
SeerAttention-R: Sparse Attention Adaptation for Long Reasoning
por: Gao, Yizhao, et al.
Publicado: (2025)
por: Gao, Yizhao, et al.
Publicado: (2025)
Interpretable Concept-Based Memory Reasoning
por: Debot, David, et al.
Publicado: (2024)
por: Debot, David, et al.
Publicado: (2024)
RaaS: Reasoning-Aware Attention Sparsity for Efficient LLM Reasoning
por: Hu, Junhao, et al.
Publicado: (2025)
por: Hu, Junhao, et al.
Publicado: (2025)
Echo State Transformer: Attention Over Finite Memories
por: Bendi-Ouis, Yannis, et al.
Publicado: (2025)
por: Bendi-Ouis, Yannis, et al.
Publicado: (2025)
Towards Reasoning Ability of Small Language Models
por: Srivastava, Gaurav, et al.
Publicado: (2025)
por: Srivastava, Gaurav, et al.
Publicado: (2025)
Enhancing Reasoning Capabilities of Small Language Models with Blueprints and Prompt Template Search
por: Han, Dongge, et al.
Publicado: (2025)
por: Han, Dongge, et al.
Publicado: (2025)
A Comprehensive Benchmark on Spectral GNNs: The Impact on Efficiency, Memory, and Effectiveness
por: Liao, Ningyi, et al.
Publicado: (2024)
por: Liao, Ningyi, et al.
Publicado: (2024)
Introducing Spectral Attention for Long-Range Dependency in Time Series Forecasting
por: Kang, Bong Gyun, et al.
Publicado: (2024)
por: Kang, Bong Gyun, et al.
Publicado: (2024)
Attend or Perish: Benchmarking Attention in Algorithmic Reasoning
por: Spiegel, Michal, et al.
Publicado: (2025)
por: Spiegel, Michal, et al.
Publicado: (2025)
LOOKAT: Lookup-Optimized Key-Attention for Memory-Efficient Transformers
por: Karmore, Aryan
Publicado: (2026)
por: Karmore, Aryan
Publicado: (2026)
QFlash: Bridging Quantization and Memory Efficiency in Vision Transformer Attention
por: Oh, Sehyeon, et al.
Publicado: (2026)
por: Oh, Sehyeon, et al.
Publicado: (2026)
Disentangling Recall and Reasoning in Transformer Models through Layer-wise Attention and Activation Analysis
por: Fartale, Harshwardhan, et al.
Publicado: (2025)
por: Fartale, Harshwardhan, et al.
Publicado: (2025)
Rank-Aware Spectral Bounds on Attention Logits for Stable Low-Precision Training
por: Emadi, Seyed Morteza
Publicado: (2026)
por: Emadi, Seyed Morteza
Publicado: (2026)
Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models
por: Qu, Yun, et al.
Publicado: (2026)
por: Qu, Yun, et al.
Publicado: (2026)
Diffusion Models Meet Contextual Bandits
por: Aouali, Imad
Publicado: (2024)
por: Aouali, Imad
Publicado: (2024)
RAST: Reasoning Activation in LLMs via Small-model Transfer
por: Ouyang, Siru, et al.
Publicado: (2025)
por: Ouyang, Siru, et al.
Publicado: (2025)
AdaTKG: Adaptive Memory for Temporal Knowledge Graph Reasoning
por: Lee, Seunghan, et al.
Publicado: (2026)
por: Lee, Seunghan, et al.
Publicado: (2026)
Scaling Reasoning without Attention
por: Zhao, Xueliang, et al.
Publicado: (2025)
por: Zhao, Xueliang, et al.
Publicado: (2025)
AtMan: Understanding Transformer Predictions Through Memory Efficient Attention Manipulation
por: Deiseroth, Björn, et al.
Publicado: (2023)
por: Deiseroth, Björn, et al.
Publicado: (2023)
DiffCLIP: Differential Attention Meets CLIP
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2025)
por: Hammoud, Hasan Abed Al Kader, et al.
Publicado: (2025)
Attention as Binding: A Vector-Symbolic Perspective on Transformer Reasoning
por: Dhayalkar, Sahil Rajesh
Publicado: (2025)
por: Dhayalkar, Sahil Rajesh
Publicado: (2025)
FROST: Filtering Reasoning Outliers with Attention for Efficient Reasoning
por: Luo, Haozheng, et al.
Publicado: (2026)
por: Luo, Haozheng, et al.
Publicado: (2026)
Can only LLMs do Reasoning?: Potential of Small Language Models in Task Planning
por: Choi, Gawon, et al.
Publicado: (2024)
por: Choi, Gawon, et al.
Publicado: (2024)
PeSANet: Physics-encoded Spectral Attention Network for Simulating PDE-Governed Complex Systems
por: Wan, Han, et al.
Publicado: (2025)
por: Wan, Han, et al.
Publicado: (2025)
E2Former-V2: On-the-Fly Equivariant Attention with Linear Activation Memory
por: Huang, Lin, et al.
Publicado: (2026)
por: Huang, Lin, et al.
Publicado: (2026)
Identity Bridge: Enabling Implicit Reasoning via Shared Latent Memory
por: Lin, Pengxiao, et al.
Publicado: (2025)
por: Lin, Pengxiao, et al.
Publicado: (2025)
Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks
por: Xu, Yifei, et al.
Publicado: (2025)
por: Xu, Yifei, et al.
Publicado: (2025)
When Linear Attention Meets Autoregressive Decoding: Towards More Effective and Efficient Linearized Large Language Models
por: You, Haoran, et al.
Publicado: (2024)
por: You, Haoran, et al.
Publicado: (2024)
Improving Chain-of-Thought for Logical Reasoning via Attention-Aware Intervention
por: Phuong, Nguyen Minh, et al.
Publicado: (2026)
por: Phuong, Nguyen Minh, et al.
Publicado: (2026)
Interpretable Hierarchical Concept Reasoning through Attention-Guided Graph Learning
por: Debot, David, et al.
Publicado: (2025)
por: Debot, David, et al.
Publicado: (2025)
Ejemplares similares
-
MKA: Memory-Keyed Attention for Efficient Long-Context Reasoning
por: Liu, Dong, et al.
Publicado: (2026) -
When Reasoning Meets Compression: Understanding the Effects of LLMs Compression on Large Reasoning Models
por: Zhang, Nan, et al.
Publicado: (2025) -
Pay Attention to Small Weights
por: Zhou, Chao, et al.
Publicado: (2025) -
Paged Attention Meets FlexAttention: Unlocking Long-Context Efficiency in Deployed Inference
por: Joshi, Thomas, et al.
Publicado: (2025) -
Probing the Limits of Compressive Memory: A Study of Infini-Attention in Small-Scale Pretraining
por: Huang, Ruizhe, et al.
Publicado: (2025)