Scheherazade: Evaluating Chain-of-Thought Math Reasoning in LLMs with Chain-of-Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miner, Stephen, Takashima, Yoshiki, Han, Simeng, Kouteili, Sam, Erata, Ferhat, Piskac, Ruzica, Shapiro, Scott J |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning How to Cube
von: Erata, Ferhat, et al.
Veröffentlicht: (2026)
von: Erata, Ferhat, et al.
Veröffentlicht: (2026)
Mining Beyond the Bools: Learning Data Transformations and Temporal Specifications
von: Kouteili, Sam Nicholas, et al.
Veröffentlicht: (2026)
von: Kouteili, Sam Nicholas, et al.
Veröffentlicht: (2026)
Quantum Circuit Reconstruction from Power Side-Channel Attacks on Quantum Computer Controllers
von: Erata, Ferhat, et al.
Veröffentlicht: (2024)
von: Erata, Ferhat, et al.
Veröffentlicht: (2024)
Learning Randomized Reductions
von: Erata, Ferhat, et al.
Veröffentlicht: (2024)
von: Erata, Ferhat, et al.
Veröffentlicht: (2024)
Evaluating GRPO and DPO for Faithful Chain-of-Thought Reasoning in LLMs
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Bengali Math Word Problem Solving with Chain of Thought Reasoning
von: Paul, Bidyarthi, et al.
Veröffentlicht: (2025)
von: Paul, Bidyarthi, et al.
Veröffentlicht: (2025)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
von: Shi, Weikang, et al.
Veröffentlicht: (2025)
Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
'Put the Car on the Stand': SMT-based Oracles for Investigating Decisions
von: Judson, Samuel, et al.
Veröffentlicht: (2023)
von: Judson, Samuel, et al.
Veröffentlicht: (2023)
Efficient Reasoning for LLMs through Speculative Chain-of-Thought
von: Wang, Jikai, et al.
Veröffentlicht: (2025)
von: Wang, Jikai, et al.
Veröffentlicht: (2025)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs
von: Xu, Yige, et al.
Veröffentlicht: (2025)
von: Xu, Yige, et al.
Veröffentlicht: (2025)
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs
von: Cao, Shidong, et al.
Veröffentlicht: (2026)
von: Cao, Shidong, et al.
Veröffentlicht: (2026)
Premise-Augmented Reasoning Chains Improve Error Identification in Math reasoning with LLMs
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2025)
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2025)
Safer Reasoning Traces: Measuring and Mitigating Chain-of-Thought Leakage in LLMs
von: Ahrend, Patrick, et al.
Veröffentlicht: (2026)
von: Ahrend, Patrick, et al.
Veröffentlicht: (2026)
Direct Value Optimization: Improving Chain-of-Thought Reasoning in LLMs with Refined Values
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhang, Hongbo, et al.
Veröffentlicht: (2025)
Chain-of-Thought Reasoning Without Prompting
von: Wang, Xuezhi, et al.
Veröffentlicht: (2024)
von: Wang, Xuezhi, et al.
Veröffentlicht: (2024)
Long Grounded Thoughts: Synthesizing Visual Problems and Reasoning Chains at Scale
von: Acuna, David, et al.
Veröffentlicht: (2025)
von: Acuna, David, et al.
Veröffentlicht: (2025)
ReasonFlux-PRM: Trajectory-Aware PRMs for Long Chain-of-Thought Reasoning in LLMs
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
von: Zou, Jiaru, et al.
Veröffentlicht: (2025)
Structured Reasoning with Tree-of-Thoughts for Bengali Math Word Problems
von: Mahmood, Aurprita, et al.
Veröffentlicht: (2025)
von: Mahmood, Aurprita, et al.
Veröffentlicht: (2025)
Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs
von: Yuan, Hongyuan, et al.
Veröffentlicht: (2026)
von: Yuan, Hongyuan, et al.
Veröffentlicht: (2026)
Direct Evaluation of Chain-of-Thought in Multi-hop Reasoning with Knowledge Graphs
von: Nguyen, Minh-Vuong, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh-Vuong, et al.
Veröffentlicht: (2024)
Draft-Thinking: Learning Efficient Reasoning in Long Chain-of-Thought LLMs
von: Cao, Jie, et al.
Veröffentlicht: (2026)
von: Cao, Jie, et al.
Veröffentlicht: (2026)
LLMs can Find Mathematical Reasoning Mistakes by Pedagogical Chain-of-Thought
von: Jiang, Zhuoxuan, et al.
Veröffentlicht: (2024)
von: Jiang, Zhuoxuan, et al.
Veröffentlicht: (2024)
Efficient Reasoning via Chain of Unconscious Thought
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
On the Impact of Fine-Tuning on Chain-of-Thought Reasoning
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
von: Lobo, Elita, et al.
Veröffentlicht: (2024)
Fractured Chain-of-Thought Reasoning
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
von: Liao, Baohao, et al.
Veröffentlicht: (2025)
Evaluating Chain-of-Thought Reasoning through Reusability and Verifiability
von: Aggarwal, Shashank, et al.
Veröffentlicht: (2026)
von: Aggarwal, Shashank, et al.
Veröffentlicht: (2026)
Learning to Reason via Mixture-of-Thought for Logical Reasoning
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
Latent Chain-of-Thought for Visual Reasoning
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
Beyond Chain-of-Thought: A Survey of Chain-of-X Paradigms for LLMs
von: Xia, Yu, et al.
Veröffentlicht: (2024)
von: Xia, Yu, et al.
Veröffentlicht: (2024)
S3-CoT: Self-Sampled Succinct Reasoning Enables Efficient Chain-of-Thought LLMs
von: Du, Yanrui, et al.
Veröffentlicht: (2026)
von: Du, Yanrui, et al.
Veröffentlicht: (2026)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
Prompting Science Report 2: The Decreasing Value of Chain of Thought in Prompting
von: Meincke, Lennart, et al.
Veröffentlicht: (2025)
von: Meincke, Lennart, et al.
Veröffentlicht: (2025)
Beyond Chain-of-Thought, Effective Graph-of-Thought Reasoning in Language Models
von: Yao, Yao, et al.
Veröffentlicht: (2023)
von: Yao, Yao, et al.
Veröffentlicht: (2023)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ye, Jiacheng, et al.
Veröffentlicht: (2024)
Keypoint-based Progressive Chain-of-Thought Distillation for LLMs
von: Feng, Kaituo, et al.
Veröffentlicht: (2024)
von: Feng, Kaituo, et al.
Veröffentlicht: (2024)
DICE: Structured Reasoning in LLMs through SLM-Guided Chain-of-Thought Correction
von: Li, Yiqi, et al.
Veröffentlicht: (2025)
von: Li, Yiqi, et al.
Veröffentlicht: (2025)
Facilitating Long Context Understanding via Supervised Chain-of-Thought Reasoning
von: Lin, Jingyang, et al.
Veröffentlicht: (2025)
von: Lin, Jingyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning How to Cube
von: Erata, Ferhat, et al.
Veröffentlicht: (2026) -
Mining Beyond the Bools: Learning Data Transformations and Temporal Specifications
von: Kouteili, Sam Nicholas, et al.
Veröffentlicht: (2026) -
Quantum Circuit Reconstruction from Power Side-Channel Attacks on Quantum Computer Controllers
von: Erata, Ferhat, et al.
Veröffentlicht: (2024) -
Learning Randomized Reductions
von: Erata, Ferhat, et al.
Veröffentlicht: (2024) -
Evaluating GRPO and DPO for Faithful Chain-of-Thought Reasoning in LLMs
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)