How Likely Do LLMs with CoT Mimic Human Reasoning?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bao, Guangsheng, Zhang, Hongbo, Wang, Cunxiang, Yang, Linyi, Zhang, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Nash CoT: Multi-Path Inference with Preference Equilibrium
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
CycleResearcher: Improving Automated Research via Automated Review
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
von: Deng, Yuntian, et al.
Veröffentlicht: (2024)
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)
RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?
von: Xu, Haotian, et al.
Veröffentlicht: (2025)
von: Xu, Haotian, et al.
Veröffentlicht: (2025)
To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
von: Sprague, Zayne, et al.
Veröffentlicht: (2024)
Detecting RLVR Training Data via Structural Convergence of Reasoning
von: Zhang, Hongbo, et al.
Veröffentlicht: (2026)
von: Zhang, Hongbo, et al.
Veröffentlicht: (2026)
Steer Like the LLM: Activation Steering that Mimics Prompting
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
von: Heyman, Geert, et al.
Veröffentlicht: (2026)
Knowledge Conflicts for LLMs: A Survey
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
von: Ye, Xinwu, et al.
Veröffentlicht: (2026)
Compositional Generalization from Learned Skills via CoT Training: A Theoretical and Structural Analysis for Reasoning
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
von: Yao, Xinhao, et al.
Veröffentlicht: (2025)
Is continuous CoT better suited for multi-lingual reasoning?
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Examples
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2025)
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2025)
How Do Latent Reasoning Methods Perform Under Weak and Strong Supervision?
von: Cui, Yingqian, et al.
Veröffentlicht: (2026)
von: Cui, Yingqian, et al.
Veröffentlicht: (2026)
Do LLMs Encode Functional Importance of Reasoning Tokens?
von: Singh, Janvijay, et al.
Veröffentlicht: (2026)
von: Singh, Janvijay, et al.
Veröffentlicht: (2026)
AI Scientists Fail Without Strong Implementation Capability
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
Catching rationalization in the act: detecting motivated reasoning before and after CoT via activation probing
von: Mirtaheri, Parsa, et al.
Veröffentlicht: (2026)
von: Mirtaheri, Parsa, et al.
Veröffentlicht: (2026)
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
von: Wang, Qibin, et al.
Veröffentlicht: (2025)
When Do LLMs Reason? A Dynamical Systems View via Entropy Phase Transitions
von: Xia, Wei, et al.
Veröffentlicht: (2026)
von: Xia, Wei, et al.
Veröffentlicht: (2026)
DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
Batched Contextual Reinforcement: A Task-Scaling Law for Efficient Reasoning
von: Yang, Bangji, et al.
Veröffentlicht: (2026)
von: Yang, Bangji, et al.
Veröffentlicht: (2026)
CoT-ICL Lab: A Synthetic Framework for Studying Chain-of-Thought Learning from In-Context Demonstrations
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2025)
von: Kothapalli, Vignesh, et al.
Veröffentlicht: (2025)
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
von: Saparkhan, Raman, et al.
Veröffentlicht: (2026)
von: Saparkhan, Raman, et al.
Veröffentlicht: (2026)
CoRT: Code-integrated Reasoning within Thinking
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
von: Li, Chengpeng, et al.
Veröffentlicht: (2025)
A Human-Like Reasoning Framework for Multi-Phases Planning Task with Large Language Models
von: Xie, Chengxing, et al.
Veröffentlicht: (2024)
von: Xie, Chengxing, et al.
Veröffentlicht: (2024)
DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search
von: Yue, Murong, et al.
Veröffentlicht: (2024)
von: Yue, Murong, et al.
Veröffentlicht: (2024)
MedCLM: Learning to Localize and Reason via a CoT-Curriculum in Medical Vision-Language Models
von: Kim, Soo Yong, et al.
Veröffentlicht: (2025)
von: Kim, Soo Yong, et al.
Veröffentlicht: (2025)
A Decomposition Perspective to Long-context Reasoning for LLMs
von: Xiao, Yanling, et al.
Veröffentlicht: (2026)
von: Xiao, Yanling, et al.
Veröffentlicht: (2026)
How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
von: Yang, Minglai, et al.
Veröffentlicht: (2025)
von: Yang, Minglai, et al.
Veröffentlicht: (2025)
Delving into RL for Image Generation with CoT: A Study on DPO vs. GRPO
von: Tong, Chengzhuo, et al.
Veröffentlicht: (2025)
von: Tong, Chengzhuo, et al.
Veröffentlicht: (2025)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning for LLM Multi-Step Reasoning
von: Wang, Huaijie, et al.
Veröffentlicht: (2024)
von: Wang, Huaijie, et al.
Veröffentlicht: (2024)
CoVerRL: Breaking the Consensus Trap in Label-Free Reasoning via Generator-Verifier Co-Evolution
von: Pan, Teng, et al.
Veröffentlicht: (2026)
von: Pan, Teng, et al.
Veröffentlicht: (2026)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
How Do LLMs Persuade? Linear Probes Can Uncover Persuasion Dynamics in Multi-Turn Conversations
von: Jaipersaud, Brandon, et al.
Veröffentlicht: (2025)
von: Jaipersaud, Brandon, et al.
Veröffentlicht: (2025)
Do Not Let Low-Probability Tokens Over-Dominate in RL for LLMs
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
von: Yang, Zhihe, et al.
Veröffentlicht: (2025)
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
von: Huang, Kaixuan, et al.
Veröffentlicht: (2025)
StepFun-Formalizer: Unlocking the Autoformalization Potential of LLMs through Knowledge-Reasoning Fusion
von: Wu, Yutong, et al.
Veröffentlicht: (2025)
von: Wu, Yutong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Nash CoT: Multi-Path Inference with Preference Equilibrium
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024) -
CycleResearcher: Improving Automated Research via Automated Review
von: Weng, Yixuan, et al.
Veröffentlicht: (2024) -
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025) -
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
von: Deng, Yuntian, et al.
Veröffentlicht: (2024) -
Exploring the Limitations of Mamba in COPY and CoT Reasoning
von: Ren, Ruifeng, et al.
Veröffentlicht: (2024)