Fill in the Blank: Exploring and Enhancing LLM Capabilities for Backward Reasoning in Math Word Problems
Fuente:
arXiv
Saved in:
| Main Authors: | Deb, Aniruddha, Oza, Neeva, Singla, Sarthak, Khandelwal, Dinesh, Garg, Dinesh, Singla, Parag |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
by: Oza, Neeva, et al.
Published: (2025)
by: Oza, Neeva, et al.
Published: (2025)
Enhancing LLM Code Generation Capabilities through Test-Driven Development and Code Interpreter
by: Jalil, Sajed, et al.
Published: (2025)
by: Jalil, Sajed, et al.
Published: (2025)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
by: Wei, Kaiwen, et al.
Published: (2025)
by: Wei, Kaiwen, et al.
Published: (2025)
Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game
by: Li, Lixing
Published: (2026)
by: Li, Lixing
Published: (2026)
SSP: Self-Supervised Prompting for Cross-Lingual Transfer to Low-Resource Languages using Large Language Models
by: Rathore, Vipul, et al.
Published: (2024)
by: Rathore, Vipul, et al.
Published: (2024)
Large Language Models Are Not Strong Abstract Reasoners
by: Gendron, Gaël, et al.
Published: (2023)
by: Gendron, Gaël, et al.
Published: (2023)
SQL Query Engine: A Self-Healing LLM Pipeline for Natural Language to PostgreSQL Translation
by: Ijaz, Muhammad Adeel
Published: (2026)
by: Ijaz, Muhammad Adeel
Published: (2026)
Understanding Syllogistic Reasoning in LLMs from Formal and Natural Language Perspectives
by: Poddar, Aheli, et al.
Published: (2025)
by: Poddar, Aheli, et al.
Published: (2025)
DynaSemble: Dynamic Ensembling of Textual and Structure-Based Models for Knowledge Graph Completion
by: Nandi, Ananjan, et al.
Published: (2023)
by: Nandi, Ananjan, et al.
Published: (2023)
NL2LOGIC: AST-Guided Translation of Natural Language into First-Order Logic with Large Language Models
by: Putra, Rizky Ramadhana, et al.
Published: (2026)
by: Putra, Rizky Ramadhana, et al.
Published: (2026)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
by: Abramov, Roman, et al.
Published: (2025)
by: Abramov, Roman, et al.
Published: (2025)
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
by: Vernie, Julius, et al.
Published: (2026)
by: Vernie, Julius, et al.
Published: (2026)
Memory Architectures for Multi-Turn Text-to-SQL: A Benchmark and Empirical Study
by: Tummalapenta, Ravi Kumar, et al.
Published: (2026)
by: Tummalapenta, Ravi Kumar, et al.
Published: (2026)
Inductive First-Order Formula Synthesis by ASP: A Case Study in Invariant Inference
by: Yang, Ziyi, et al.
Published: (2026)
by: Yang, Ziyi, et al.
Published: (2026)
DISCO: DISCovering Overfittings as Causal Rules for Text Classification Models
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
Advancing Explainability in Neural Machine Translation: Analytical Metrics for Attention and Alignment Consistency
by: Mishra, Anurag
Published: (2024)
by: Mishra, Anurag
Published: (2024)
Leveraging Log Probabilities in Language Models to Forecast Future Events
by: Soru, Tommaso, et al.
Published: (2025)
by: Soru, Tommaso, et al.
Published: (2025)
Training Language Models to Use Prolog as a Tool
by: Mellgren, Niklas, et al.
Published: (2025)
by: Mellgren, Niklas, et al.
Published: (2025)
LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics
by: Peyronnet, Antoine, et al.
Published: (2026)
by: Peyronnet, Antoine, et al.
Published: (2026)
Causal Cartographer: From Mapping to Reasoning Over Counterfactual Worlds
by: Gendron, Gaël, et al.
Published: (2025)
by: Gendron, Gaël, et al.
Published: (2025)
Anthem 2.0: Automated Reasoning for Answer Set Programming
by: Fandinno, Jorge, et al.
Published: (2025)
by: Fandinno, Jorge, et al.
Published: (2025)
Mitigating LLM Hallucinations through Domain-Grounded Tiered Retrieval
by: Haque, Md. Asraful, et al.
Published: (2026)
by: Haque, Md. Asraful, et al.
Published: (2026)
Enhancing Large Language Models through Neuro-Symbolic Integration and Ontological Reasoning
by: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Published: (2025)
by: Vsevolodovna, Ruslan Idelfonso Magana, et al.
Published: (2025)
Integration of Contextual Descriptors in Ontology Alignment for Enrichment of Semantic Correspondence
by: Manziuk, Eduard, et al.
Published: (2024)
by: Manziuk, Eduard, et al.
Published: (2024)
Advancing Natural Language Formalization to First Order Logic with Fine-tuned LLMs
by: Vossel, Felix, et al.
Published: (2025)
by: Vossel, Felix, et al.
Published: (2025)
Robust Uncertainty Quantification for Factual Generation of Large Language Models
by: Zhang, Yuhao, et al.
Published: (2026)
by: Zhang, Yuhao, et al.
Published: (2026)
Beyond Prediction -- Structuring Epistemic Integrity in Artificial Reasoning Systems
by: Wright, Craig Steven
Published: (2025)
by: Wright, Craig Steven
Published: (2025)
LeanExplore: A search engine for Lean 4 declarations
by: Asher, Justin
Published: (2025)
by: Asher, Justin
Published: (2025)
LLM-Assisted Formalization Enables Deterministic Detection of Statutory Inconsistency in the Internal Revenue Code
by: Yadamsuren, Borchuluun, et al.
Published: (2025)
by: Yadamsuren, Borchuluun, et al.
Published: (2025)
Reasoning and Planning with Dynamically Changing Norms
by: Olson, Taylor, et al.
Published: (2026)
by: Olson, Taylor, et al.
Published: (2026)
AI Agents-as-Judge: Automated Assessment of Accuracy, Consistency, Completeness and Clarity for Enterprise Documents
by: Dasgupta, Sudip, et al.
Published: (2025)
by: Dasgupta, Sudip, et al.
Published: (2025)
The Stable Model Semantics for Higher-Order Logic Programming
by: Bogaerts, Bart, et al.
Published: (2024)
by: Bogaerts, Bart, et al.
Published: (2024)
Making Logic a First-Class Citizen in Generative ML for Networking
by: Hè, Hongyu, et al.
Published: (2025)
by: Hè, Hongyu, et al.
Published: (2025)
Quantum Abduction: A New Paradigm for Reasoning under Uncertainty
by: Pareschi, Remo
Published: (2025)
by: Pareschi, Remo
Published: (2025)
Can Large Language Models Learn Independent Causal Mechanisms?
by: Gendron, Gaël, et al.
Published: (2024)
by: Gendron, Gaël, et al.
Published: (2024)
CLEV: LLM-Based Evaluation Through Lightweight Efficient Voting for Free-Form Question-Answering
by: Badshah, Sher, et al.
Published: (2025)
by: Badshah, Sher, et al.
Published: (2025)
Retrieval and Augmentation of Domain Knowledge for Text-to-SQL Semantic Parsing
by: Patwardhan, Manasi, et al.
Published: (2025)
by: Patwardhan, Manasi, et al.
Published: (2025)
Datrics Text2SQL: A Framework for Natural Language to SQL Query Generation
by: Gladkykh, Tetiana, et al.
Published: (2025)
by: Gladkykh, Tetiana, et al.
Published: (2025)
A Declarative System for Optimizing AI Workloads
by: Liu, Chunwei, et al.
Published: (2024)
by: Liu, Chunwei, et al.
Published: (2024)
On the Equivalence between Logic Programming and SETAF
by: Alcântara, João, et al.
Published: (2024)
by: Alcântara, João, et al.
Published: (2024)
Similar Items
-
CETBench: A Novel Dataset constructed via Transformations over Programs for Benchmarking LLMs for Code-Equivalence Checking
by: Oza, Neeva, et al.
Published: (2025) -
Enhancing LLM Code Generation Capabilities through Test-Driven Development and Code Interpreter
by: Jalil, Sajed, et al.
Published: (2025) -
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
by: Wei, Kaiwen, et al.
Published: (2025) -
Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game
by: Li, Lixing
Published: (2026) -
SSP: Self-Supervised Prompting for Cross-Lingual Transfer to Low-Resource Languages using Large Language Models
by: Rathore, Vipul, et al.
Published: (2024)