Why LLMs Cannot Think and How to Fix It
Fuente:
arXiv
Salvato in:
| Autori principali: | Jahrens, Marius, Martinetz, Thomas |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Why Fine-Tuning Encourages Hallucinations and How to Fix It
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
di: Kaplan, Guy, et al.
Pubblicazione: (2026)
Rethinking generalization of classifiers in separable classes scenarios and over-parameterized regimes
di: Martinetz, Julius, et al.
Pubblicazione: (2024)
di: Martinetz, Julius, et al.
Pubblicazione: (2024)
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
di: Rizvi-Martel, Michael, et al.
Pubblicazione: (2026)
di: Rizvi-Martel, Michael, et al.
Pubblicazione: (2026)
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
di: Meeus, Matthieu, et al.
Pubblicazione: (2024)
di: Meeus, Matthieu, et al.
Pubblicazione: (2024)
OptimalThinkingBench: Evaluating Over and Underthinking in LLMs
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2025)
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2025)
Asynchronous Reasoning: Training-Free Interactive Thinking LLMs
di: Yakushev, George, et al.
Pubblicazione: (2025)
di: Yakushev, George, et al.
Pubblicazione: (2025)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
di: Wang, Shouren, et al.
Pubblicazione: (2025)
di: Wang, Shouren, et al.
Pubblicazione: (2025)
Nearest-Neighbor Density Estimation for Dependency Suppression
di: Anderson, Kathleen, et al.
Pubblicazione: (2026)
di: Anderson, Kathleen, et al.
Pubblicazione: (2026)
Do LLMs Understand Romanian Driving Laws? A Study on Multimodal and Fine-Tuned Question Answering
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
di: Barbu, Eduard, et al.
Pubblicazione: (2025)
Forecasting Downstream Performance of LLMs With Proxy Metrics
di: Patel, Arkil, et al.
Pubblicazione: (2026)
di: Patel, Arkil, et al.
Pubblicazione: (2026)
CooperBench: Why Coding Agents Cannot be Your Teammates Yet
di: Khatua, Arpandeep, et al.
Pubblicazione: (2026)
di: Khatua, Arpandeep, et al.
Pubblicazione: (2026)
Unknown Unknowns: Why Hidden Intentions in LLMs Evade Detection
di: Srivastav, Devansh, et al.
Pubblicazione: (2026)
di: Srivastav, Devansh, et al.
Pubblicazione: (2026)
Can We Count on LLMs? The Fixed-Effect Fallacy and Claims of GPT-4 Capabilities
di: Ball, Thomas, et al.
Pubblicazione: (2024)
di: Ball, Thomas, et al.
Pubblicazione: (2024)
The Right Answer, the Wrong Direction: Why Transformers Fail at Counting and How to Fix It
di: Garcia, Gabriel
Pubblicazione: (2026)
di: Garcia, Gabriel
Pubblicazione: (2026)
Do Multilingual LLMs Think In English?
di: Schut, Lisa, et al.
Pubblicazione: (2025)
di: Schut, Lisa, et al.
Pubblicazione: (2025)
Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?
di: Kim, Jeonghye, et al.
Pubblicazione: (2026)
di: Kim, Jeonghye, et al.
Pubblicazione: (2026)
Rethinking Thinking Tokens: Understanding Why They Underperform in Practice
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
Perplexity Cannot Always Tell Right from Wrong
di: Veličković, Petar, et al.
Pubblicazione: (2026)
di: Veličković, Petar, et al.
Pubblicazione: (2026)
Reverse Thinking Makes LLMs Stronger Reasoners
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
Measuring What LLMs Think They Do: SHAP Faithfulness and Deployability on Financial Tabular Classification
di: AlMarri, Saeed, et al.
Pubblicazione: (2025)
di: AlMarri, Saeed, et al.
Pubblicazione: (2025)
Non-Halting Queries: Exploiting Fixed Points in LLMs
di: Hammouri, Ghaith, et al.
Pubblicazione: (2024)
di: Hammouri, Ghaith, et al.
Pubblicazione: (2024)
Multimodal Language Models Cannot Spot Spatial Inconsistencies
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
di: Khangaonkar, Om, et al.
Pubblicazione: (2026)
MillStone: How Open-Minded Are LLMs?
di: Triedman, Harold, et al.
Pubblicazione: (2025)
di: Triedman, Harold, et al.
Pubblicazione: (2025)
How Does Quantization Affect Multilingual LLMs?
di: Marchisio, Kelly, et al.
Pubblicazione: (2024)
di: Marchisio, Kelly, et al.
Pubblicazione: (2024)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition
di: Ye, Guanghao, et al.
Pubblicazione: (2025)
di: Ye, Guanghao, et al.
Pubblicazione: (2025)
Why Are Web AI Agents More Vulnerable Than Standalone LLMs? A Security Analysis
di: Chiang, Jeffrey Yang Fan, et al.
Pubblicazione: (2025)
di: Chiang, Jeffrey Yang Fan, et al.
Pubblicazione: (2025)
Why is "Chicago" Predictive of Deceptive Reviews? Using LLMs to Discover Language Phenomena from Lexical Cues
di: Qu, Jiaming, et al.
Pubblicazione: (2025)
di: Qu, Jiaming, et al.
Pubblicazione: (2025)
Reinforcement Learning for Latent-Space Thinking in LLMs
di: Özeren, Enes, et al.
Pubblicazione: (2025)
di: Özeren, Enes, et al.
Pubblicazione: (2025)
Thinking in Latents: Adaptive Anchor Refinement for Implicit Reasoning in LLMs
di: Sheshanarayana, Disha, et al.
Pubblicazione: (2026)
di: Sheshanarayana, Disha, et al.
Pubblicazione: (2026)
RLVR Training of LLMs Does Not Improve Thinking Ability for General QA: Evaluation Method and a Simple Solution
di: Li, Kaiyuan, et al.
Pubblicazione: (2026)
di: Li, Kaiyuan, et al.
Pubblicazione: (2026)
Uncovering Gaps in How Humans and LLMs Interpret Subjective Language
di: Jones, Erik, et al.
Pubblicazione: (2025)
di: Jones, Erik, et al.
Pubblicazione: (2025)
Slot Machines: How LLMs Keep Track of Multiple Entities
di: Bogdan, Paul C., et al.
Pubblicazione: (2026)
di: Bogdan, Paul C., et al.
Pubblicazione: (2026)
Think Before You Lie: How Reasoning Leads to Honesty
di: Yuan, Ann, et al.
Pubblicazione: (2026)
di: Yuan, Ann, et al.
Pubblicazione: (2026)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
di: Lucchetti, Francesca, et al.
Pubblicazione: (2024)
di: Lucchetti, Francesca, et al.
Pubblicazione: (2024)
What Makes the Preferred Thinking Direction for LLMs in Multiple-choice Questions?
di: Zhang, Yizhe, et al.
Pubblicazione: (2025)
di: Zhang, Yizhe, et al.
Pubblicazione: (2025)
When Two LLMs Debate, Both Think They'll Win
di: Prasad, Pradyumna Shyama, et al.
Pubblicazione: (2025)
di: Prasad, Pradyumna Shyama, et al.
Pubblicazione: (2025)
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
di: Okutomi, Akira
Pubblicazione: (2025)
di: Okutomi, Akira
Pubblicazione: (2025)
SciLitLLM: How to Adapt LLMs for Scientific Literature Understanding
di: Li, Sihang, et al.
Pubblicazione: (2024)
di: Li, Sihang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Semantic Anchors in In-Context Learning: Why Small LLMs Cannot Flip Their Labels
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025) -
Why Fine-Tuning Encourages Hallucinations and How to Fix It
di: Kaplan, Guy, et al.
Pubblicazione: (2026) -
Rethinking generalization of classifiers in separable classes scenarios and over-parameterized regimes
di: Martinetz, Julius, et al.
Pubblicazione: (2024) -
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
di: Rizvi-Martel, Michael, et al.
Pubblicazione: (2026) -
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
di: Meeus, Matthieu, et al.
Pubblicazione: (2024)