Error Detection and Correction for Interpretable Mathematics in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yijin, Cornelio, Cristina, Leiva, Mario, Shakarian, Paulo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rule-Based Error Detection and Correction to Operationalize Movement Trajectory Classification
by: Xi, Bowen, et al.
Published: (2023)
by: Xi, Bowen, et al.
Published: (2023)
Machine Learning Model Integration with Open World Temporal Logic for Process Automation
by: Aditya, Dyuman, et al.
Published: (2025)
by: Aditya, Dyuman, et al.
Published: (2025)
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
by: Ngu, Noel, et al.
Published: (2023)
by: Ngu, Noel, et al.
Published: (2023)
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
by: Li, Xiaoyuan, et al.
Published: (2024)
by: Li, Xiaoyuan, et al.
Published: (2024)
Mathematical Computation and Reasoning Errors by Large Language Models
by: Zhang, Liang, et al.
Published: (2025)
by: Zhang, Liang, et al.
Published: (2025)
Probabilistic Foundations for Metacognition via Hybrid-AI
by: Shakarian, Paulo, et al.
Published: (2025)
by: Shakarian, Paulo, et al.
Published: (2025)
Better Later Than Sooner: Neuro-Symbolic Knowledge Graph Construction via Ontology-grounded Post-extraction Correction
by: Loconte, Lorenzo, et al.
Published: (2026)
by: Loconte, Lorenzo, et al.
Published: (2026)
Recover: A Neuro-Symbolic Framework for Failure Detection and Recovery
by: Cornelio, Cristina, et al.
Published: (2024)
by: Cornelio, Cristina, et al.
Published: (2024)
A Systematic Analysis of Large Language Models with RAG-enabled Dynamic Prompting for Medical Error Detection and Correction
by: Ahmed, Farzad, et al.
Published: (2025)
by: Ahmed, Farzad, et al.
Published: (2025)
Improving MPI Error Detection and Repair with Large Language Models and Bug References
by: Piersall, Scott, et al.
Published: (2026)
by: Piersall, Scott, et al.
Published: (2026)
Error Detection and Constraint Recovery in Hierarchical Multi-Label Classification without Prior Knowledge
by: Kricheli, Joshua Shay, et al.
Published: (2024)
by: Kricheli, Joshua Shay, et al.
Published: (2024)
Embedding Self-Correction as an Inherent Ability in Large Language Models for Enhanced Mathematical Reasoning
by: Gao, Kuofeng, et al.
Published: (2024)
by: Gao, Kuofeng, et al.
Published: (2024)
CorBenchX: Large-Scale Chest X-Ray Error Dataset and Vision-Language Model Benchmark for Report Error Correction
by: Zou, Jing, et al.
Published: (2025)
by: Zou, Jing, et al.
Published: (2025)
Enhancing Mathematical Reasoning in Large Language Models with Self-Consistency-Based Hallucination Detection
by: Liu, MingShan, et al.
Published: (2025)
by: Liu, MingShan, et al.
Published: (2025)
Large Language Models and Mathematical Reasoning Failures
by: Boye, Johan, et al.
Published: (2025)
by: Boye, Johan, et al.
Published: (2025)
Metacognitive AI: Framework and the Case for a Neurosymbolic Approach
by: Wei, Hua, et al.
Published: (2024)
by: Wei, Hua, et al.
Published: (2024)
Generative Large Language Models Trained for Detecting Errors in Radiology Reports
by: Sun, Cong, et al.
Published: (2025)
by: Sun, Cong, et al.
Published: (2025)
Beyond the First Error: Process Reward Models for Reflective Mathematical Reasoning
by: Yang, Zhaohui, et al.
Published: (2025)
by: Yang, Zhaohui, et al.
Published: (2025)
Large Language Models for Mathematical Analysis
by: Chen, Ziye, et al.
Published: (2024)
by: Chen, Ziye, et al.
Published: (2024)
Consistency-based Abductive Reasoning over Perceptual Errors of Multiple Pre-trained Models in Novel Environments
by: Leiva, Mario, et al.
Published: (2025)
by: Leiva, Mario, et al.
Published: (2025)
Synthesizing and Adapting Error Correction Data for Mobile Large Language Model Applications
by: Zhang, Yanxiang, et al.
Published: (2025)
by: Zhang, Yanxiang, et al.
Published: (2025)
Large Language Model enabled Mathematical Modeling
by: Zhang, Guoyun
Published: (2025)
by: Zhang, Guoyun
Published: (2025)
DSGram: Dynamic Weighting Sub-Metrics for Grammatical Error Correction in the Era of Large Language Models
by: Xie, Jinxiang, et al.
Published: (2024)
by: Xie, Jinxiang, et al.
Published: (2024)
Concurrent Linguistic Error Detection (CLED): a New Methodology for Error Detection in Large Language Models
by: Zhu, Jinhua, et al.
Published: (2024)
by: Zhu, Jinhua, et al.
Published: (2024)
Position: Artificial Intelligence Needs Meta Intelligence -- the Case for Metacognitive AI
by: Chuprov, Sergei, et al.
Published: (2026)
by: Chuprov, Sergei, et al.
Published: (2026)
Abduction of Domain Relationships from Data for VQA
by: Chowdhury, Al Mehdi Saadat, et al.
Published: (2025)
by: Chowdhury, Al Mehdi Saadat, et al.
Published: (2025)
Sea-cret Agents: Maritime Abduction for Region Generation to Expose Dark Vessel Trajectories
by: Bavikadi, Divyagna, et al.
Published: (2025)
by: Bavikadi, Divyagna, et al.
Published: (2025)
A Survey on Large Language Models for Mathematical Reasoning
by: Wang, Peng-Yuan, et al.
Published: (2025)
by: Wang, Peng-Yuan, et al.
Published: (2025)
Are UFOs Driving Innovation? The Illusion of Causality in Large Language Models
by: Carro, María Victoria, et al.
Published: (2024)
by: Carro, María Victoria, et al.
Published: (2024)
A Survey on Mathematical Reasoning and Optimization with Large Language Models
by: Forootani, Ali
Published: (2025)
by: Forootani, Ali
Published: (2025)
GRPO and Reflection Reward for Mathematical Reasoning in Large Language Models
by: Wang, Zhijie
Published: (2026)
by: Wang, Zhijie
Published: (2026)
Mathematical Reasoning in Large Language Models: Assessing Logical and Arithmetic Errors across Wide Numerical Ranges
by: Shrestha, Safal, et al.
Published: (2025)
by: Shrestha, Safal, et al.
Published: (2025)
What Is Missing: Interpretable Ratings for Large Language Model Outputs
by: Stranges, Nicholas, et al.
Published: (2026)
by: Stranges, Nicholas, et al.
Published: (2026)
GAUSS: Benchmarking Structured Mathematical Skills for Large Language Models
by: Zhang, Yue, et al.
Published: (2025)
by: Zhang, Yue, et al.
Published: (2025)
Internalized Self-Correction for Large Language Models
by: Upadhyaya, Nishanth, et al.
Published: (2024)
by: Upadhyaya, Nishanth, et al.
Published: (2024)
Computational Blueprints: Generating Isomorphic Mathematics Problems with Large Language Models
by: Kim, Jeong-Hoon, et al.
Published: (2025)
by: Kim, Jeong-Hoon, et al.
Published: (2025)
Synthetic Error Injection Fails to Elicit Self-Correction In Language Models
by: Wu, David X., et al.
Published: (2025)
by: Wu, David X., et al.
Published: (2025)
Mathematical Foundation of Interpretable Equivariant Surrogate Models
by: Colombini, Jacopo Joy, et al.
Published: (2025)
by: Colombini, Jacopo Joy, et al.
Published: (2025)
Using Large Language Models to Study Mathematical Practice
by: D'Alessandro, William
Published: (2025)
by: D'Alessandro, William
Published: (2025)
ToolCritic: Detecting and Correcting Tool-Use Errors in Dialogue Systems
by: Hamad, Hassan, et al.
Published: (2025)
by: Hamad, Hassan, et al.
Published: (2025)
Similar Items
-
Rule-Based Error Detection and Correction to Operationalize Movement Trajectory Classification
by: Xi, Bowen, et al.
Published: (2023) -
Machine Learning Model Integration with Open World Temporal Logic for Process Automation
by: Aditya, Dyuman, et al.
Published: (2025) -
Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries
by: Ngu, Noel, et al.
Published: (2023) -
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
by: Li, Xiaoyuan, et al.
Published: (2024) -
Mathematical Computation and Reasoning Errors by Large Language Models
by: Zhang, Liang, et al.
Published: (2025)