Attend or Perish: Benchmarking Attention in Algorithmic Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Spiegel, Michal, Štefánik, Michal, Kadlčík, Marek, Kuchař, Josef |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VectorEdits: A Dataset and Benchmark for Instruction-Based Editing of Vector Graphics
by: Kuchař, Josef, et al.
Published: (2025)
by: Kuchař, Josef, et al.
Published: (2025)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)
by: Kadlčík, Marek, et al.
Published: (2025)
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
by: Štefánik, Michal, et al.
Published: (2025)
by: Štefánik, Michal, et al.
Published: (2025)
Self-training Language Models for Arithmetic Reasoning
by: Kadlčík, Marek, et al.
Published: (2024)
by: Kadlčík, Marek, et al.
Published: (2024)
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
by: Štefánik, Michal, et al.
Published: (2025)
by: Štefánik, Michal, et al.
Published: (2025)
Concept-aware Data Construction Improves In-context Learning of Language Models
by: Štefánik, Michal, et al.
Published: (2024)
by: Štefánik, Michal, et al.
Published: (2024)
Negation: A Pink Elephant in the Large Language Models' Room?
by: Vrabcová, Tereza, et al.
Published: (2025)
by: Vrabcová, Tereza, et al.
Published: (2025)
How to Protect Models against Adversarial Unlearning?
by: Jasiorski, Patryk, et al.
Published: (2025)
by: Jasiorski, Patryk, et al.
Published: (2025)
Towards consistency of rule-based explainer and black box model -- fusion of rule induction and XAI-based feature importance
by: Kozielski, Michał, et al.
Published: (2024)
by: Kozielski, Michał, et al.
Published: (2024)
Contrastive Representations for Temporal Reasoning
by: Ziarko, Alicja, et al.
Published: (2025)
by: Ziarko, Alicja, et al.
Published: (2025)
Revisiting Prompt Sensitivity in Large Language Models for Text Classification: The Role of Prompt Underspecification
by: Pecher, Branislav, et al.
Published: (2026)
by: Pecher, Branislav, et al.
Published: (2026)
Subgoal Search For Complex Reasoning Tasks
by: Czechowski, Konrad, et al.
Published: (2021)
by: Czechowski, Konrad, et al.
Published: (2021)
BoTTA: Benchmarking on-device Test Time Adaptation
by: Danilowski, Michal, et al.
Published: (2025)
by: Danilowski, Michal, et al.
Published: (2025)
Think Twice: Measuring the Efficiency of Eliminating Prediction Shortcuts of Question Answering Models
by: Mikula, Lukáš, et al.
Published: (2023)
by: Mikula, Lukáš, et al.
Published: (2023)
Scout Before You Attend: Sketch-and-Walk Sparse Attention for Efficient LLM Inference
by: Le, Hoang Anh Duy, et al.
Published: (2026)
by: Le, Hoang Anh Duy, et al.
Published: (2026)
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
by: Estermann, Benjamin, et al.
Published: (2024)
by: Estermann, Benjamin, et al.
Published: (2024)
Deep Dictionary-Free Method for Identifying Linear Model of Nonlinear System with Input Delay
by: Valábek, Patrik, et al.
Published: (2025)
by: Valábek, Patrik, et al.
Published: (2025)
IMGTB: A Framework for Machine-Generated Text Detection Benchmarking
by: Spiegel, Michal, et al.
Published: (2023)
by: Spiegel, Michal, et al.
Published: (2023)
Representation-based Broad Hallucination Detectors Fail to Generalize Out of Distribution
by: Dubanowska, Zuzanna, et al.
Published: (2025)
by: Dubanowska, Zuzanna, et al.
Published: (2025)
Benchmarking ChatGPT on Algorithmic Reasoning
by: McLeish, Sean, et al.
Published: (2024)
by: McLeish, Sean, et al.
Published: (2024)
Automatically Differentiable Nonlinear Tensor Networks (ADNTNs) for Exponential Compression of Deep Neural Networks
by: Cichocki, Andrzej, et al.
Published: (2026)
by: Cichocki, Andrzej, et al.
Published: (2026)
A Comparative Study of Text Retrieval Models on DaReCzech
by: Stetina, Jakub, et al.
Published: (2024)
by: Stetina, Jakub, et al.
Published: (2024)
When Does Non-Uniform Replay Matter in Reinforcement Learning?
by: Korniak, Michal, et al.
Published: (2026)
by: Korniak, Michal, et al.
Published: (2026)
Accelerating Goal-Conditioned RL Algorithms and Research
by: Bortkiewicz, Michał, et al.
Published: (2024)
by: Bortkiewicz, Michał, et al.
Published: (2024)
Should We Attend More or Less? Modulating Attention for Fairness
by: Zayed, Abdelrahman, et al.
Published: (2023)
by: Zayed, Abdelrahman, et al.
Published: (2023)
TimeSeriesGym: A Scalable Benchmark for (Time Series) Machine Learning Engineering Agents
by: Cai, Yifu, et al.
Published: (2025)
by: Cai, Yifu, et al.
Published: (2025)
BiblioPage: A Dataset of Scanned Title Pages for Bibliographic Metadata Extraction
by: Kohút, Jan, et al.
Published: (2025)
by: Kohút, Jan, et al.
Published: (2025)
FractalBench: Diagnosing Visual-Mathematical Reasoning Through Recursive Program Synthesis
by: Ondras, Jan, et al.
Published: (2025)
by: Ondras, Jan, et al.
Published: (2025)
LeanTree: Accelerating White-Box Proof Search with Factorized States in Lean 4
by: Kripner, Matěj, et al.
Published: (2025)
by: Kripner, Matěj, et al.
Published: (2025)
Adaptive Compression of the Latent Space in Variational Autoencoders
by: Sejnova, Gabriela, et al.
Published: (2023)
by: Sejnova, Gabriela, et al.
Published: (2023)
What Does Flow Matching Bring To TD Learning?
by: Agrawalla, Bhavya, et al.
Published: (2026)
by: Agrawalla, Bhavya, et al.
Published: (2026)
Attending to Graph Transformers
by: Müller, Luis, et al.
Published: (2023)
by: Müller, Luis, et al.
Published: (2023)
SeerAttention-R: Sparse Attention Adaptation for Long Reasoning
by: Gao, Yizhao, et al.
Published: (2025)
by: Gao, Yizhao, et al.
Published: (2025)
Classical and Deep Reinforcement Learning Inventory Control Policies for Pharmaceutical Supply Chains with Perishability and Non-Stationarity
by: Stranieri, Francesco, et al.
Published: (2025)
by: Stranieri, Francesco, et al.
Published: (2025)
RaaS: Reasoning-Aware Attention Sparsity for Efficient LLM Reasoning
by: Hu, Junhao, et al.
Published: (2025)
by: Hu, Junhao, et al.
Published: (2025)
Filter then Attend: Improving attention-based Time Series Forecasting with Spectral Filtering
by: Dayag, Elisha, et al.
Published: (2025)
by: Dayag, Elisha, et al.
Published: (2025)
Sparse Model Soups: A Recipe for Improved Pruning via Model Averaging
by: Zimmer, Max, et al.
Published: (2023)
by: Zimmer, Max, et al.
Published: (2023)
Post-Norm can Resharpen Attention
by: Zsámboki, Pál, et al.
Published: (2025)
by: Zsámboki, Pál, et al.
Published: (2025)
How Can We Synthesize High-Quality Pretraining Data? A Systematic Study of Prompt Design, Generator Model, and Source Data
by: Niklaus, Joel, et al.
Published: (2026)
by: Niklaus, Joel, et al.
Published: (2026)
Progressive Sparse Attention: Algorithm and System Co-design for Efficient Attention in LLM Serving
by: Zhou, Qihui, et al.
Published: (2025)
by: Zhou, Qihui, et al.
Published: (2025)
Similar Items
-
VectorEdits: A Dataset and Benchmark for Instruction-Based Editing of Vector Graphics
by: Kuchař, Josef, et al.
Published: (2025) -
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025) -
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
by: Štefánik, Michal, et al.
Published: (2025) -
Self-training Language Models for Arithmetic Reasoning
by: Kadlčík, Marek, et al.
Published: (2024) -
Language Models Learn Universal Representations of Numbers and Here's Why You Should Care
by: Štefánik, Michal, et al.
Published: (2025)