Can We Trust AI Explanations? Evidence of Systematic Underreporting in Chain-of-Thought Reasoning
Fuente:
arXiv
Saved in:
| Main Author: | Mehta, Deep Pankajbhai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026)
by: Mehta, Deep Pankajbhai
Published: (2026)
Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs
by: Mehta, Deep
Published: (2026)
by: Mehta, Deep
Published: (2026)
When Reasoning Traces Become Performative: Step-Level Evidence that Chain-of-Thought Is an Imperfect Oversight Channel
by: Li, Wenkai, et al.
Published: (2026)
by: Li, Wenkai, et al.
Published: (2026)
Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction
by: Lucas, Ryan, et al.
Published: (2025)
by: Lucas, Ryan, et al.
Published: (2025)
Can Reasoning Models Obfuscate Reasoning? Stress-Testing Chain-of-Thought Monitorability
by: Zolkowski, Artur, et al.
Published: (2025)
by: Zolkowski, Artur, et al.
Published: (2025)
CanaryBench: Stress Testing Privacy Leakage in Cluster-Level Conversation Summaries
by: Mehta, Deep
Published: (2026)
by: Mehta, Deep
Published: (2026)
Can We Trust LLM Detectors?
by: Sandhan, Jivnesh, et al.
Published: (2026)
by: Sandhan, Jivnesh, et al.
Published: (2026)
LLM Reasoning Is Latent, Not the Chain of Thought
by: Wang, Wenshuo
Published: (2026)
by: Wang, Wenshuo
Published: (2026)
Can We Trust AI to Govern AI? Benchmarking LLM Performance on Privacy and AI Governance Exams
by: Witherspoon, Zane, et al.
Published: (2025)
by: Witherspoon, Zane, et al.
Published: (2025)
Teaching AI Stepwise Diagnostic Reasoning with Report-Guided Chain-of-Thought Learning
by: Luo, Yihong, et al.
Published: (2025)
by: Luo, Yihong, et al.
Published: (2025)
Tracing Thought: Using Chain-of-Thought Reasoning to Identify the LLM Behind AI-Generated Text
by: Agrahari, Shifali, et al.
Published: (2025)
by: Agrahari, Shifali, et al.
Published: (2025)
Bench-2-CoP: Can We Trust Benchmarking for EU AI Compliance?
by: Prandi, Matteo, et al.
Published: (2025)
by: Prandi, Matteo, et al.
Published: (2025)
Are Reasoning LLMs Robust to Interventions on Their Chain-of-Thought?
by: von Recum, Alexander, et al.
Published: (2026)
by: von Recum, Alexander, et al.
Published: (2026)
Diagnosing Pathological Chain-of-Thought in Reasoning Models
by: Liu, Manqing, et al.
Published: (2026)
by: Liu, Manqing, et al.
Published: (2026)
Reasoning Models Struggle to Control their Chains of Thought
by: Yueh-Han, Chen, et al.
Published: (2026)
by: Yueh-Han, Chen, et al.
Published: (2026)
Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation
by: Eriksson, Maria, et al.
Published: (2025)
by: Eriksson, Maria, et al.
Published: (2025)
Can Separators Improve Chain-of-Thought Prompting?
by: Park, Yoonjeong, et al.
Published: (2024)
by: Park, Yoonjeong, et al.
Published: (2024)
Uncertainty Awareness and Trust in Explainable AI- On Trust Calibration using Local and Global Explanations
by: Newen, Carina, et al.
Published: (2025)
by: Newen, Carina, et al.
Published: (2025)
Latent Chain-of-Thought for Visual Reasoning
by: Sun, Guohao, et al.
Published: (2025)
by: Sun, Guohao, et al.
Published: (2025)
Fractured Chain-of-Thought Reasoning
by: Liao, Baohao, et al.
Published: (2025)
by: Liao, Baohao, et al.
Published: (2025)
Reason from Future: Reverse Thought Chain Enhances LLM Reasoning
by: Xu, Yinlong, et al.
Published: (2025)
by: Xu, Yinlong, et al.
Published: (2025)
Streaming Hallucination Detection in Long Chain-of-Thought Reasoning
by: Lu, Haolang, et al.
Published: (2026)
by: Lu, Haolang, et al.
Published: (2026)
Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
by: Zaman, Kerem, et al.
Published: (2025)
by: Zaman, Kerem, et al.
Published: (2025)
LLM-REVal: Can We Trust LLM Reviewers Yet?
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
by: Jiang, Gangwei, et al.
Published: (2025)
by: Jiang, Gangwei, et al.
Published: (2025)
The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoning
by: Chen, Qiguang, et al.
Published: (2026)
by: Chen, Qiguang, et al.
Published: (2026)
Unveiling Confirmation Bias in Chain-of-Thought Reasoning
by: Wan, Yue, et al.
Published: (2025)
by: Wan, Yue, et al.
Published: (2025)
Markov Chain of Thought for Efficient Mathematical Reasoning
by: Yang, Wen, et al.
Published: (2024)
by: Yang, Wen, et al.
Published: (2024)
GRACE: Discriminator-Guided Chain-of-Thought Reasoning
by: Khalifa, Muhammad, et al.
Published: (2023)
by: Khalifa, Muhammad, et al.
Published: (2023)
Chain-of-Trust: A Progressive Trust Evaluation Framework Enabled by Generative AI
by: Zhu, Botao, et al.
Published: (2025)
by: Zhu, Botao, et al.
Published: (2025)
Thought Manipulation: External Thought Can Be Efficient for Large Reasoning Models
by: Liu, Yule, et al.
Published: (2025)
by: Liu, Yule, et al.
Published: (2025)
LaRS: Latent Reasoning Skills for Chain-of-Thought Reasoning
by: Xu, Zifan, et al.
Published: (2023)
by: Xu, Zifan, et al.
Published: (2023)
Deep Hidden Cognition Facilitates Reliable Chain-of-Thought Reasoning
by: Chen, Zijun, et al.
Published: (2025)
by: Chen, Zijun, et al.
Published: (2025)
Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
by: Shao, Yifei, et al.
Published: (2026)
by: Shao, Yifei, et al.
Published: (2026)
Output Supervision Can Obfuscate the Chain of Thought
by: Drori, Jacob, et al.
Published: (2025)
by: Drori, Jacob, et al.
Published: (2025)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
by: Ye, Jiacheng, et al.
Published: (2024)
by: Ye, Jiacheng, et al.
Published: (2024)
Thoughts without Thinking: Reconsidering the Explanatory Value of Chain-of-Thought Reasoning in LLMs through Agentic Pipelines
by: Manuvinakurike, Ramesh, et al.
Published: (2025)
by: Manuvinakurike, Ramesh, et al.
Published: (2025)
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-Thought
by: Lee, Jooyoung, et al.
Published: (2024)
by: Lee, Jooyoung, et al.
Published: (2024)
When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making
by: Fan, Shutong, et al.
Published: (2026)
by: Fan, Shutong, et al.
Published: (2026)
Similar Items
-
Placement Semantics for Distributed Deep Learning: A Systematic Framework for Analyzing Parallelism Strategies
by: Mehta, Deep Pankajbhai
Published: (2026) -
Does Inference Scaling Improve Reasoning Faithfulness? A Multi-Model Analysis of Self-Consistency Tradeoffs
by: Mehta, Deep
Published: (2026) -
When Reasoning Traces Become Performative: Step-Level Evidence that Chain-of-Thought Is an Imperfect Oversight Channel
by: Li, Wenkai, et al.
Published: (2026) -
Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction
by: Lucas, Ryan, et al.
Published: (2025) -
Can Reasoning Models Obfuscate Reasoning? Stress-Testing Chain-of-Thought Monitorability
by: Zolkowski, Artur, et al.
Published: (2025)