Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kambhampati, Subbarao, Valmeekam, Karthik, Bhambri, Siddhant, Palod, Vardhan, Saldyt, Lucas, Stechly, Kaya, Samineni, Soumya Rani, Kalwar, Durgesh, Biswas, Upasana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025)
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025)
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
von: Palod, Vardhan, et al.
Veröffentlicht: (2025)
von: Palod, Vardhan, et al.
Veröffentlicht: (2025)
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2025)
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2025)
Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025)
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025)
Who is Helping Whom? Analyzing Inter-dependencies to Evaluate Cooperation in Human-AI Teaming
von: Biswas, Upasana, et al.
Veröffentlicht: (2025)
von: Biswas, Upasana, et al.
Veröffentlicht: (2025)
Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
Do Cognitively Interpretable Reasoning Traces Improve LLM Performance?
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2025)
LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
von: Kambhampati, Subbarao, et al.
Veröffentlicht: (2024)
von: Kambhampati, Subbarao, et al.
Veröffentlicht: (2024)
Chain of Thoughtlessness? An Analysis of CoT in Planning
von: Stechly, Kaya, et al.
Veröffentlicht: (2024)
von: Stechly, Kaya, et al.
Veröffentlicht: (2024)
(How) Do reasoning models reason?
von: Subbarao Kambhampati, et al.
Veröffentlicht: (2025)
von: Subbarao Kambhampati, et al.
Veröffentlicht: (2025)
LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
On the Self-Verification Limitations of Large Language Models on Reasoning and Planning Tasks
von: Stechly, Kaya, et al.
Veröffentlicht: (2024)
von: Stechly, Kaya, et al.
Veröffentlicht: (2024)
Evaluating the False Trust Engendered by LLM Explanations
von: Palod, Vardhan, et al.
Veröffentlicht: (2026)
von: Palod, Vardhan, et al.
Veröffentlicht: (2026)
Planning in Strawberry Fields: Evaluating and Improving the Planning and Scheduling Capabilities of LRM o1
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2023)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2023)
Robust Planning with LLM-Modulo Framework: Case Study in Travel Planning
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
Algorithmic Language Models with Neurally Compiled Libraries
von: Saldyt, Lucas, et al.
Veröffentlicht: (2024)
von: Saldyt, Lucas, et al.
Veröffentlicht: (2024)
Mind The Gap: Quantifying Mechanistic Gaps in Algorithmic Reasoning via Neural Compilation
von: Saldyt, Lucas, et al.
Veröffentlicht: (2025)
von: Saldyt, Lucas, et al.
Veröffentlicht: (2025)
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
On the Brittle Foundations of ReAct Prompting for Agentic Large Language Models
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
von: Verma, Mudit, et al.
Veröffentlicht: (2024)
Robust Planning with Compound LLM Architectures: An LLM-Modulo Approach
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
von: Gundawar, Atharva, et al.
Veröffentlicht: (2024)
Can Large Language Models Reason and Plan?
von: Kambhampati, Subbarao
Veröffentlicht: (2024)
von: Kambhampati, Subbarao
Veröffentlicht: (2024)
Can large language models reason and plan?
von: Subbarao Kambhampati
Veröffentlicht: (2024)
von: Subbarao Kambhampati
Veröffentlicht: (2024)
Efficiency Boost in Decentralized Optimization: Reimagining Neighborhood Aggregation with Minimal Overhead
von: Kalwar, Durgesh, et al.
Veröffentlicht: (2025)
von: Kalwar, Durgesh, et al.
Veröffentlicht: (2025)
Learning from Ambiguous Demonstrations with Self-Explanation Guided Reinforcement Learning
von: Zha, Yantian, et al.
Veröffentlicht: (2021)
von: Zha, Yantian, et al.
Veröffentlicht: (2021)
Explainable Human-AI Interaction: A Planning Perspective
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
von: Sreedharan, Sarath, et al.
Veröffentlicht: (2024)
Discounting and Drug Seeking in Biological Hierarchical Reinforcement Learning
von: Palod, Vardhan, et al.
Veröffentlicht: (2025)
von: Palod, Vardhan, et al.
Veröffentlicht: (2025)
Evolution of the Ratio of Mg II Intensities During Solar Flares
von: Roy, Soumya, et al.
Veröffentlicht: (2024)
von: Roy, Soumya, et al.
Veröffentlicht: (2024)
Introduction to Variational Quantum Algorithms
von: Stęchły, Michał
Veröffentlicht: (2024)
von: Stęchły, Michał
Veröffentlicht: (2024)
Architecturally Refined Cerium‐Integrated Hydroxyapatite/CNT Nanocomposite Coatings: Enhanced Mechanics and Biofunction for Orthopaedic Implantation
von: Durgesh Phogat, et al.
Veröffentlicht: (2025)
von: Durgesh Phogat, et al.
Veröffentlicht: (2025)
OPTIMIZING BATCH ETL PIPELINES FOR MULTI-BRAND RETAIL ANALYTICS ACROSS LARGE STORE NETWORKS
von: Narasimha Chaitanya Samineni
Veröffentlicht: (2020)
von: Narasimha Chaitanya Samineni
Veröffentlicht: (2020)
Quantifying the Impact of Epigallocatechin Gallate and Piperine on D. tigrina Regeneration
von: Polepalli, Vinil, et al.
Veröffentlicht: (2025)
von: Polepalli, Vinil, et al.
Veröffentlicht: (2025)
Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
von: Guan, Lin, et al.
Veröffentlicht: (2024)
von: Guan, Lin, et al.
Veröffentlicht: (2024)
Deep Learning Approaches for Network Traffic Classification in the Internet of Things (IoT): A Survey
von: Kalwar, Jawad Hussain, et al.
Veröffentlicht: (2024)
von: Kalwar, Jawad Hussain, et al.
Veröffentlicht: (2024)
Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers
von: Yang, Wang, et al.
Veröffentlicht: (2026)
von: Yang, Wang, et al.
Veröffentlicht: (2026)
Building Trust Through Empathetic AI: A Mixed‐Methods Study on Anthropomorphic Artificial Intelligence Driving Sustainable Consumer Behavior in Tourism
von: Satinder Kumar, et al.
Veröffentlicht: (2026)
von: Satinder Kumar, et al.
Veröffentlicht: (2026)
Payoff Continuity in Games of Incomplete Information Across Models of Knowledge
von: Kambhampati, Ashwin
Veröffentlicht: (2025)
von: Kambhampati, Ashwin
Veröffentlicht: (2025)
Lexicographic Robustness and the Efficiency of Optimal Mechanisms
von: Kambhampati, Ashwin
Veröffentlicht: (2026)
von: Kambhampati, Ashwin
Veröffentlicht: (2026)
Robust Performance Evaluation of Independent and Identical Agents
von: Kambhampati, Ashwin
Veröffentlicht: (2024)
von: Kambhampati, Ashwin
Veröffentlicht: (2024)
Ähnliche Einträge
-
RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025) -
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
von: Palod, Vardhan, et al.
Veröffentlicht: (2025) -
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2025) -
Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains
von: Samineni, Soumya Rani, et al.
Veröffentlicht: (2025) -
Who is Helping Whom? Analyzing Inter-dependencies to Evaluate Cooperation in Human-AI Teaming
von: Biswas, Upasana, et al.
Veröffentlicht: (2025)