MathDivide: Improved mathematical reasoning by large language models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Srivastava, Saksham Sahai, Gandhi, Ashutosh
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929352087699456
author Srivastava, Saksham Sahai
Gandhi, Ashutosh
author_facet Srivastava, Saksham Sahai
Gandhi, Ashutosh
contents Large language models have been proven to be capable of handling complex linguistic and cognitive tasks. Therefore their usage has been extended to tasks requiring logical reasoning ability such as Mathematics. In this paper, we propose a prompting technique called MathDivide that breaks down the mathematical problem into simpler subproblems. Each of the subproblems is formulated as an algebraic expression whose value is evaluated by the Python code generated by the LLM for the corresponding algebraic expression. The values fed to the Python code are the numerical values provided in the problem statement. The solutions for the subproblems are composed together to obtain the final answer for the problem statement. Finally, the final answer is compared to the correct answer. If the final answer matches the correct answer, it is produced as output else a refinement prompt is fed to the LLM. We experiment with this prompting technique on both closed-source LLM models and open-source LLM models using GSM8K dataset. The results obtained demonstrate that MathDivide was able to significantly outperform the leading prompting technique called Math-prompter.
format Preprint
id arxiv_https___arxiv_org_abs_2405_13004
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle MathDivide: Improved mathematical reasoning by large language models
Srivastava, Saksham Sahai
Gandhi, Ashutosh
Computation and Language
Artificial Intelligence
Large language models have been proven to be capable of handling complex linguistic and cognitive tasks. Therefore their usage has been extended to tasks requiring logical reasoning ability such as Mathematics. In this paper, we propose a prompting technique called MathDivide that breaks down the mathematical problem into simpler subproblems. Each of the subproblems is formulated as an algebraic expression whose value is evaluated by the Python code generated by the LLM for the corresponding algebraic expression. The values fed to the Python code are the numerical values provided in the problem statement. The solutions for the subproblems are composed together to obtain the final answer for the problem statement. Finally, the final answer is compared to the correct answer. If the final answer matches the correct answer, it is produced as output else a refinement prompt is fed to the LLM. We experiment with this prompting technique on both closed-source LLM models and open-source LLM models using GSM8K dataset. The results obtained demonstrate that MathDivide was able to significantly outperform the leading prompting technique called Math-prompter.
title MathDivide: Improved mathematical reasoning by large language models
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2405.13004