Can LLMs Learn from Previous Mistakes? Investigating LLMs' Errors to Boost for Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tong, Yongqi, Li, Dawei, Wang, Sizhe, Wang, Yujia, Teng, Fei, Shang, Jingbo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimizing Language Model's Reasoning Abilities with Weak Supervision
von: Tong, Yongqi, et al.
Veröffentlicht: (2024)
von: Tong, Yongqi, et al.
Veröffentlicht: (2024)
Retrieved In-Context Principles from Previous Mistakes
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
BPO: Towards Balanced Preference Optimization between Knowledge Breadth and Depth in Alignment
von: Wang, Sizhe, et al.
Veröffentlicht: (2024)
von: Wang, Sizhe, et al.
Veröffentlicht: (2024)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
LLMs can Find Mathematical Reasoning Mistakes by Pedagogical Chain-of-Thought
von: Jiang, Zhuoxuan, et al.
Veröffentlicht: (2024)
von: Jiang, Zhuoxuan, et al.
Veröffentlicht: (2024)
When Can LLMs Actually Correct Their Own Mistakes? A Critical Survey of Self-Correction of LLMs
von: Kamoi, Ryo, et al.
Veröffentlicht: (2024)
von: Kamoi, Ryo, et al.
Veröffentlicht: (2024)
Evaluating the Smooth Control of Attribute Intensity in Text Generation with LLMs
von: Zhou, Shang, et al.
Veröffentlicht: (2024)
von: Zhou, Shang, et al.
Veröffentlicht: (2024)
Can LLMs Reason with Rules? Logic Scaffolding for Stress-Testing and Improving LLMs
von: Wang, Siyuan, et al.
Veröffentlicht: (2024)
von: Wang, Siyuan, et al.
Veröffentlicht: (2024)
READ: Improving Relation Extraction from an ADversarial Perspective
von: Li, Dawei, et al.
Veröffentlicht: (2024)
von: Li, Dawei, et al.
Veröffentlicht: (2024)
The Price of Format: Diversity Collapse in LLMs
von: Yun, Longfei, et al.
Veröffentlicht: (2025)
von: Yun, Longfei, et al.
Veröffentlicht: (2025)
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
von: Singh, Joykirat, et al.
Veröffentlicht: (2024)
Can LLMs Follow Simple Rules?
von: Mu, Norman, et al.
Veröffentlicht: (2023)
von: Mu, Norman, et al.
Veröffentlicht: (2023)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
Learning from Mistakes: Negative Reasoning Samples Enhance Out-of-Domain Generalization
von: Tian, Xueyun, et al.
Veröffentlicht: (2026)
von: Tian, Xueyun, et al.
Veröffentlicht: (2026)
Can LLMs Reason in the Wild with Programs?
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
Do LLMs Make Mistakes Like Students? Exploring Natural Alignment between Language Models and Human Error Patterns
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
von: Liu, Naiming, et al.
Veröffentlicht: (2025)
Reasoning Boosts Opinion Alignment in LLMs
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2026)
When Do LLMs Admit Their Mistakes? Understanding The Role Of Model Belief In Retraction
von: Yang, Yuqing, et al.
Veröffentlicht: (2025)
von: Yang, Yuqing, et al.
Veröffentlicht: (2025)
Relevant or Random: Can LLMs Truly Perform Analogical Reasoning?
von: Qin, Chengwei, et al.
Veröffentlicht: (2024)
von: Qin, Chengwei, et al.
Veröffentlicht: (2024)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs
von: Yang, Wanli, et al.
Veröffentlicht: (2026)
von: Yang, Wanli, et al.
Veröffentlicht: (2026)
MinosEval: Distinguishing Factoid and Non-Factoid for Tailored Open-Ended QA Evaluation with LLMs
von: Fan, Yongqi, et al.
Veröffentlicht: (2025)
von: Fan, Yongqi, et al.
Veröffentlicht: (2025)
Can Hallucinations Help? Boosting LLMs for Drug Discovery
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2025)
von: Yuan, Shuzhou, et al.
Veröffentlicht: (2025)
Investigating Neurons and Heads in Transformer-based LLMs for Typographical Errors
von: Tsuji, Kohei, et al.
Veröffentlicht: (2025)
von: Tsuji, Kohei, et al.
Veröffentlicht: (2025)
LLMs + Persona-Plug = Personalized LLMs
von: Liu, Jiongnan, et al.
Veröffentlicht: (2024)
von: Liu, Jiongnan, et al.
Veröffentlicht: (2024)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
von: Gan, Esther, et al.
Veröffentlicht: (2024)
von: Gan, Esther, et al.
Veröffentlicht: (2024)
AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
von: Xiong, Jian, et al.
Veröffentlicht: (2025)
Can LLMs Solve longer Math Word Problems Better?
von: Xu, Xin, et al.
Veröffentlicht: (2024)
von: Xu, Xin, et al.
Veröffentlicht: (2024)
KDRL: Post-Training Reasoning LLMs via Unified Knowledge Distillation and Reinforcement Learning
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
von: Xu, Hongling, et al.
Veröffentlicht: (2025)
Can LLMs Perceive Time? An Empirical Investigation
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
von: Garikaparthi, Aniketh
Veröffentlicht: (2026)
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
Order Matters: Rethinking Prompt Construction in In-Context Learning
von: Li, Warren, et al.
Veröffentlicht: (2025)
von: Li, Warren, et al.
Veröffentlicht: (2025)
Can LLMs Learn to Map the World from Local Descriptions?
von: Xia, Sirui, et al.
Veröffentlicht: (2025)
von: Xia, Sirui, et al.
Veröffentlicht: (2025)
CodeBoost: Boosting Code LLMs by Squeezing Knowledge from Code Snippets with RL
von: Wang, Sijie, et al.
Veröffentlicht: (2025)
von: Wang, Sijie, et al.
Veröffentlicht: (2025)
EPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via Reinforcement Learning
von: Liu, Xiaoqian, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoqian, et al.
Veröffentlicht: (2025)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
ReSpark: Leveraging Previous Data Reports as References to Generate New Reports with LLMs
von: Tian, Yuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuan, et al.
Veröffentlicht: (2025)
Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs
von: Hong, Yining, et al.
Veröffentlicht: (2026)
von: Hong, Yining, et al.
Veröffentlicht: (2026)
Can LLMs Compute with Reasons?
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Optimizing Language Model's Reasoning Abilities with Weak Supervision
von: Tong, Yongqi, et al.
Veröffentlicht: (2024) -
Retrieved In-Context Principles from Previous Mistakes
von: Sun, Hao, et al.
Veröffentlicht: (2024) -
BPO: Towards Balanced Preference Optimization between Knowledge Breadth and Depth in Alignment
von: Wang, Sizhe, et al.
Veröffentlicht: (2024) -
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
von: Yu, Erxin, et al.
Veröffentlicht: (2025) -
LLMs can Find Mathematical Reasoning Mistakes by Pedagogical Chain-of-Thought
von: Jiang, Zhuoxuan, et al.
Veröffentlicht: (2024)