LOGICPO: Efficient Translation of NL-based Logical Problems to FOL using LLMs and Preference Optimization

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Viswanadha, Koushik, Ghosal, Deepanway, Aditya, Somak
Natura: Preprint
Pubblicazione: 2025
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866913907146227712
author Viswanadha, Koushik
Ghosal, Deepanway
Aditya, Somak
author_facet Viswanadha, Koushik
Ghosal, Deepanway
Aditya, Somak
contents Logical reasoning is a key task for artificial intelligence due to it's role in major downstream tasks such as Question Answering, Summarization. Recent methods in improving the reasoning ability of LLMs fall short in correctly converting a natural language reasoning problem to an equivalent logical formulation, which hinders the framework's overall ability to reason. Towards this, we propose to use finetuning on a preference optimization dataset to learn to parse and represent a natural language problem as a whole to a consistent logical program by 1) introducing a new supervised and preference optimization dataset LogicPO, and 2) adopting popular techniques such as Direct Preference Optimization (DPO), Kahneman-Tversky optimization (KTO) to finetune open-source LLMs. Our best model with Phi-3.5 consistently outperforms GPT-3.5-turbo's (8-shot) by producing 10% more logically correct and with 14% less syntax errors. Through the framework and our improved evaluation metrics, we offer a promising direction in improving the logical reasoning of LLMs by better representing them in their logical formulations.
format Preprint
id arxiv_https___arxiv_org_abs_2506_18383
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle LOGICPO: Efficient Translation of NL-based Logical Problems to FOL using LLMs and Preference Optimization
Viswanadha, Koushik
Ghosal, Deepanway
Aditya, Somak
Machine Learning
Artificial Intelligence
Logical reasoning is a key task for artificial intelligence due to it's role in major downstream tasks such as Question Answering, Summarization. Recent methods in improving the reasoning ability of LLMs fall short in correctly converting a natural language reasoning problem to an equivalent logical formulation, which hinders the framework's overall ability to reason. Towards this, we propose to use finetuning on a preference optimization dataset to learn to parse and represent a natural language problem as a whole to a consistent logical program by 1) introducing a new supervised and preference optimization dataset LogicPO, and 2) adopting popular techniques such as Direct Preference Optimization (DPO), Kahneman-Tversky optimization (KTO) to finetune open-source LLMs. Our best model with Phi-3.5 consistently outperforms GPT-3.5-turbo's (8-shot) by producing 10% more logically correct and with 14% less syntax errors. Through the framework and our improved evaluation metrics, we offer a promising direction in improving the logical reasoning of LLMs by better representing them in their logical formulations.
title LOGICPO: Efficient Translation of NL-based Logical Problems to FOL using LLMs and Preference Optimization
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2506.18383