Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning
Fuente:
arXiv
Salvato in:
| Autore principale: | Cazares, Manuel Israel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cognitive Load Limits in Large Language Models: Benchmarking Multi-Hop Reasoning
di: Adapala, Sai Teja Reddy
Pubblicazione: (2025)
di: Adapala, Sai Teja Reddy
Pubblicazione: (2025)
Sarcasm Detection in a Less-Resourced Language
di: Đoković, Lazar, et al.
Pubblicazione: (2024)
di: Đoković, Lazar, et al.
Pubblicazione: (2024)
Integrating Expert Labels into LLM-based Emission Goal Detection: Example Selection vs Automatic Prompt Design
di: Wrzalik, Marco, et al.
Pubblicazione: (2024)
di: Wrzalik, Marco, et al.
Pubblicazione: (2024)
Language Models, Graph Searching, and Supervision Adulteration: When More Supervision is Less and How to Make More More
di: Frydenlund, Arvid
Pubblicazione: (2025)
di: Frydenlund, Arvid
Pubblicazione: (2025)
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
Mathador-LM: A Dynamic Benchmark for Mathematical Reasoning on Large Language Models
di: Kurtic, Eldar, et al.
Pubblicazione: (2024)
di: Kurtic, Eldar, et al.
Pubblicazione: (2024)
Is Less More? Quality, Quantity and Context in Idiom Processing with Natural Language Models
di: Knietaite, Agne, et al.
Pubblicazione: (2024)
di: Knietaite, Agne, et al.
Pubblicazione: (2024)
ZERA: Zero-init Instruction Evolving Refinement Agent -- From Zero Instructions to Structured Prompts via Principle-based Optimization
di: Yi, Seungyoun, et al.
Pubblicazione: (2025)
di: Yi, Seungyoun, et al.
Pubblicazione: (2025)
Thread Detection and Response Generation using Transformers with Prompt Optimisation
di: T, Kevin Joshua, et al.
Pubblicazione: (2024)
di: T, Kevin Joshua, et al.
Pubblicazione: (2024)
QUAD: Quantization and Parameter-Efficient Tuning of LLM with Activation Decomposition
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
di: Hu, Yuxuan, et al.
Pubblicazione: (2025)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
di: Garg, Saloni, et al.
Pubblicazione: (2026)
di: Garg, Saloni, et al.
Pubblicazione: (2026)
Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment
di: Zhang, Yizhuo, et al.
Pubblicazione: (2025)
di: Zhang, Yizhuo, et al.
Pubblicazione: (2025)
External Hippocampus: Topological Cognitive Maps for Guiding Large Language Model Reasoning
di: Yan, Jian
Pubblicazione: (2025)
di: Yan, Jian
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability
di: Du, Yucheng
Pubblicazione: (2026)
di: Du, Yucheng
Pubblicazione: (2026)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
di: Özeren, Enes, et al.
Pubblicazione: (2025)
di: Özeren, Enes, et al.
Pubblicazione: (2025)
WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training
di: Tian, Changxin, et al.
Pubblicazione: (2025)
di: Tian, Changxin, et al.
Pubblicazione: (2025)
SocraSynth: Multi-LLM Reasoning with Conditional Statistics
di: Chang, Edward Y.
Pubblicazione: (2024)
di: Chang, Edward Y.
Pubblicazione: (2024)
Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance
di: Naser-Moghadasi, Mahdi, et al.
Pubblicazione: (2026)
di: Naser-Moghadasi, Mahdi, et al.
Pubblicazione: (2026)
Reasoning Over the Glyphs: Evaluation of LLM's Decipherment of Rare Scripts
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
di: Shih, Yu-Fei, et al.
Pubblicazione: (2025)
Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols
di: Hu, Julia, et al.
Pubblicazione: (2026)
di: Hu, Julia, et al.
Pubblicazione: (2026)
PersonalLLM: Tailoring LLMs to Individual Preferences
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
LLM Vocabulary Compression for Low-Compute Environments
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
TIME: Temporally Intelligent Meta-reasoning Engine for Context-Triggered Explicit Reasoning
di: Das, Susmit
Pubblicazione: (2026)
di: Das, Susmit
Pubblicazione: (2026)
Quantization-Robust LLM Unlearning via Low-Rank Adaptation
di: Abitante, João Vitor Boer, et al.
Pubblicazione: (2026)
di: Abitante, João Vitor Boer, et al.
Pubblicazione: (2026)
HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools
di: Garg, Aashna, et al.
Pubblicazione: (2026)
di: Garg, Aashna, et al.
Pubblicazione: (2026)
Digital Guardians: Can GPT-4, Perspective API, and Moderation API reliably detect hate speech in reader comments of German online newspapers?
di: Weber, Manuel, et al.
Pubblicazione: (2025)
di: Weber, Manuel, et al.
Pubblicazione: (2025)
Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning
di: Xu, Shuyao, et al.
Pubblicazione: (2025)
di: Xu, Shuyao, et al.
Pubblicazione: (2025)
Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders
di: Patel, Het, et al.
Pubblicazione: (2026)
di: Patel, Het, et al.
Pubblicazione: (2026)
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
di: Saparkhan, Raman, et al.
Pubblicazione: (2026)
di: Saparkhan, Raman, et al.
Pubblicazione: (2026)
Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data
di: de Campos, Andresa Rodrigues, et al.
Pubblicazione: (2026)
di: de Campos, Andresa Rodrigues, et al.
Pubblicazione: (2026)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
Engineering A Large Language Model From Scratch
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2024)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2024)
Prompted Contextual Vectors for Spear-Phishing Detection
di: Nahmias, Daniel, et al.
Pubblicazione: (2024)
di: Nahmias, Daniel, et al.
Pubblicazione: (2024)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
di: Wang, Xintao, et al.
Pubblicazione: (2026)
di: Wang, Xintao, et al.
Pubblicazione: (2026)
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
di: Szilvasy, Gergely, et al.
Pubblicazione: (2026)
di: Szilvasy, Gergely, et al.
Pubblicazione: (2026)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
Benchmarking Cognitive Biases in Large Language Models as Evaluators
di: Koo, Ryan, et al.
Pubblicazione: (2023)
di: Koo, Ryan, et al.
Pubblicazione: (2023)
Diagnosing and Addressing Pitfalls in KG-RAG Datasets: Toward More Reliable Benchmarking
di: Zhang, Liangliang, et al.
Pubblicazione: (2025)
di: Zhang, Liangliang, et al.
Pubblicazione: (2025)
Towards Latent Diffusion Suitable For Text
di: Midavaine, Nesta, et al.
Pubblicazione: (2026)
di: Midavaine, Nesta, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Cognitive Load Limits in Large Language Models: Benchmarking Multi-Hop Reasoning
di: Adapala, Sai Teja Reddy
Pubblicazione: (2025) -
Sarcasm Detection in a Less-Resourced Language
di: Đoković, Lazar, et al.
Pubblicazione: (2024) -
Integrating Expert Labels into LLM-based Emission Goal Detection: Example Selection vs Automatic Prompt Design
di: Wrzalik, Marco, et al.
Pubblicazione: (2024) -
Language Models, Graph Searching, and Supervision Adulteration: When More Supervision is Less and How to Make More More
di: Frydenlund, Arvid
Pubblicazione: (2025) -
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)