The Illusion of Insight in Reasoning Models
Fuente:
arXiv
Salvato in:
| Autori principali: | d'Aliberti, Liv G., Ribeiro, Manoel Horta |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Behavior-Consistent Deep Reinforcement Learning
di: Hussing, Marcel, et al.
Pubblicazione: (2026)
di: Hussing, Marcel, et al.
Pubblicazione: (2026)
Commercial Persuasion in AI-Mediated Conversations
di: Salvi, Francesco, et al.
Pubblicazione: (2026)
di: Salvi, Francesco, et al.
Pubblicazione: (2026)
Accumulating Context Changes the Beliefs of Language Models
di: Geng, Jiayi, et al.
Pubblicazione: (2025)
di: Geng, Jiayi, et al.
Pubblicazione: (2025)
Speculative Decoding: Performance or Illusion?
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2025)
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2025)
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
di: Shojaee, Parshin, et al.
Pubblicazione: (2025)
di: Shojaee, Parshin, et al.
Pubblicazione: (2025)
Control Illusion: The Failure of Instruction Hierarchies in Large Language Models
di: Geng, Yilin, et al.
Pubblicazione: (2025)
di: Geng, Yilin, et al.
Pubblicazione: (2025)
A Comment On "The Illusion of Thinking": Reframing the Reasoning Cliff as an Agentic Gap
di: Khan, Sheraz, et al.
Pubblicazione: (2025)
di: Khan, Sheraz, et al.
Pubblicazione: (2025)
Privacy-Enhancing Technologies for Artificial Intelligence-Enabled Systems
di: d'Aliberti, Liv, et al.
Pubblicazione: (2024)
di: d'Aliberti, Liv, et al.
Pubblicazione: (2024)
The Illusion of Latent Generalization: Bi-directionality and the Reversal Curse
di: Coda-Forno, Julian, et al.
Pubblicazione: (2026)
di: Coda-Forno, Julian, et al.
Pubblicazione: (2026)
Aligning to Illusions: Choice Blindness in Human and AI Feedback
di: Wu, Wenbin
Pubblicazione: (2026)
di: Wu, Wenbin
Pubblicazione: (2026)
An Illusion of Progress? Assessing the Current State of Web Agents
di: Xue, Tianci, et al.
Pubblicazione: (2025)
di: Xue, Tianci, et al.
Pubblicazione: (2025)
The Leaderboard Illusion
di: Singh, Shivalika, et al.
Pubblicazione: (2025)
di: Singh, Shivalika, et al.
Pubblicazione: (2025)
The Illusion of Readiness in Health AI
di: Gu, Yu, et al.
Pubblicazione: (2025)
di: Gu, Yu, et al.
Pubblicazione: (2025)
Are UFOs Driving Innovation? The Illusion of Causality in Large Language Models
di: Carro, María Victoria, et al.
Pubblicazione: (2024)
di: Carro, María Victoria, et al.
Pubblicazione: (2024)
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning
di: Niu, Jingcheng, et al.
Pubblicazione: (2025)
di: Niu, Jingcheng, et al.
Pubblicazione: (2025)
When Incentives Backfire, Data Stops Being Human
di: Santy, Sebastin, et al.
Pubblicazione: (2025)
di: Santy, Sebastin, et al.
Pubblicazione: (2025)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
di: Munir, Sheza, et al.
Pubblicazione: (2026)
di: Munir, Sheza, et al.
Pubblicazione: (2026)
Can Large Language Models Integrate Spatial Data? Empirical Insights into Reasoning Strengths and Computational Weaknesses
di: Han, Bin, et al.
Pubblicazione: (2025)
di: Han, Bin, et al.
Pubblicazione: (2025)
Exploring Reasoning Biases in Large Language Models Through Syllogism: Insights from the NeuBAROCO Dataset
di: Ozeki, Kentaro, et al.
Pubblicazione: (2024)
di: Ozeki, Kentaro, et al.
Pubblicazione: (2024)
Pareidolic Illusions of Meaning: ChatGPT, Pseudolaw and the Triumph of Form over Substance
di: McIntyre, Joe
Pubblicazione: (2025)
di: McIntyre, Joe
Pubblicazione: (2025)
Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Observations
di: Chen, Nuo, et al.
Pubblicazione: (2023)
di: Chen, Nuo, et al.
Pubblicazione: (2023)
Modeling Hierarchical Thinking in Large Reasoning Models
di: Shahariar, G M, et al.
Pubblicazione: (2025)
di: Shahariar, G M, et al.
Pubblicazione: (2025)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
di: Janiak, Denis, et al.
Pubblicazione: (2025)
di: Janiak, Denis, et al.
Pubblicazione: (2025)
GIVE: Structured Reasoning of Large Language Models with Knowledge Graph Inspired Veracity Extrapolation
di: He, Jiashu, et al.
Pubblicazione: (2024)
di: He, Jiashu, et al.
Pubblicazione: (2024)
From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
di: Zhang, Jue, et al.
Pubblicazione: (2025)
di: Zhang, Jue, et al.
Pubblicazione: (2025)
A Rational Analysis of the Speech-to-Song Illusion
di: Marjieh, Raja, et al.
Pubblicazione: (2024)
di: Marjieh, Raja, et al.
Pubblicazione: (2024)
Beyond the Needle's Illusion: Decoupled Evaluation of Evidence Access and Use under Semantic Interference at 326M-Token Scale
di: Lin, Tianwei, et al.
Pubblicazione: (2026)
di: Lin, Tianwei, et al.
Pubblicazione: (2026)
Reasoning Models Will Sometimes Lie About Their Reasoning
di: Walden, William, et al.
Pubblicazione: (2026)
di: Walden, William, et al.
Pubblicazione: (2026)
Reasoning Models Reason Well, Until They Don't
di: Rameshkumar, Revanth, et al.
Pubblicazione: (2025)
di: Rameshkumar, Revanth, et al.
Pubblicazione: (2025)
A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
di: Wu, Zhengxuan, et al.
Pubblicazione: (2024)
Conversational Control with Ontologies for Large Language Models: A Lightweight Framework for Constrained Generation
di: Gendron, Barbara, et al.
Pubblicazione: (2026)
di: Gendron, Barbara, et al.
Pubblicazione: (2026)
ReasoningShield: Safety Detection over Reasoning Traces of Large Reasoning Models
di: Li, Changyi, et al.
Pubblicazione: (2025)
di: Li, Changyi, et al.
Pubblicazione: (2025)
ReasonGRM: Enhancing Generative Reward Models through Large Reasoning Models
di: Chen, Bin, et al.
Pubblicazione: (2025)
di: Chen, Bin, et al.
Pubblicazione: (2025)
Route-and-Reason: Scaling Large Language Model Reasoning with Reinforced Model Router
di: Shao, Chenyang, et al.
Pubblicazione: (2025)
di: Shao, Chenyang, et al.
Pubblicazione: (2025)
RFEval: Benchmarking Reasoning Faithfulness under Counterfactual Reasoning Intervention in Large Reasoning Models
di: Han, Yunseok, et al.
Pubblicazione: (2026)
di: Han, Yunseok, et al.
Pubblicazione: (2026)
The Semantic Illusion: Certified Limits of Embedding-Based Hallucination Detection in RAG Systems
di: Sinha, Debu
Pubblicazione: (2025)
di: Sinha, Debu
Pubblicazione: (2025)
TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models
di: Zhang, Liancheng, et al.
Pubblicazione: (2026)
di: Zhang, Liancheng, et al.
Pubblicazione: (2026)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
di: Luo, Linhao, et al.
Pubblicazione: (2023)
di: Luo, Linhao, et al.
Pubblicazione: (2023)
Safety Through Reasoning: An Empirical Study of Reasoning Guardrail Models
di: Sreedhar, Makesh Narsimhan, et al.
Pubblicazione: (2025)
di: Sreedhar, Makesh Narsimhan, et al.
Pubblicazione: (2025)
Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
di: Liu, Ruikang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Behavior-Consistent Deep Reinforcement Learning
di: Hussing, Marcel, et al.
Pubblicazione: (2026) -
Commercial Persuasion in AI-Mediated Conversations
di: Salvi, Francesco, et al.
Pubblicazione: (2026) -
Accumulating Context Changes the Beliefs of Language Models
di: Geng, Jiayi, et al.
Pubblicazione: (2025) -
Speculative Decoding: Performance or Illusion?
di: Liu, Xiaoxuan, et al.
Pubblicazione: (2025) -
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
di: Shojaee, Parshin, et al.
Pubblicazione: (2025)