KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
Fuente:
arXiv
Saved in:
| Main Authors: | Saha, Soumadeep, Chaturvedi, Akshay, Saha, Saptarshi, Garain, Utpal, Asher, Nicholas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Models are Crossword Solvers
by: Saha, Soumadeep, et al.
Published: (2024)
by: Saha, Soumadeep, et al.
Published: (2024)
sudoLLM: On Multi-role Alignment of Language Models
by: Saha, Soumadeep, et al.
Published: (2025)
by: Saha, Soumadeep, et al.
Published: (2025)
Strong hallucinations from negation and how to fix them
by: Asher, Nicholas, et al.
Published: (2024)
by: Asher, Nicholas, et al.
Published: (2024)
On Explaining with Attention Matrices
by: Naim, Omar, et al.
Published: (2024)
by: Naim, Omar, et al.
Published: (2024)
InfinityMATH: A Scalable Instruction Tuning Dataset in Programmatic Mathematical Reasoning
by: Zhang, Bo-Wen, et al.
Published: (2024)
by: Zhang, Bo-Wen, et al.
Published: (2024)
Do Biased Models Have Biased Thoughts?
by: Rajwal, Swati, et al.
Published: (2025)
by: Rajwal, Swati, et al.
Published: (2025)
VALUED -- Vision and Logical Understanding Evaluation Dataset
by: Saha, Soumadeep, et al.
Published: (2023)
by: Saha, Soumadeep, et al.
Published: (2023)
ExpressivityBench: Can LLMs Communicate Implicitly?
by: Tint, Joshua, et al.
Published: (2024)
by: Tint, Joshua, et al.
Published: (2024)
Understanding Syllogistic Reasoning in LLMs from Formal and Natural Language Perspectives
by: Poddar, Aheli, et al.
Published: (2025)
by: Poddar, Aheli, et al.
Published: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
Region Mixup
by: Saha, Saptarshi, et al.
Published: (2024)
by: Saha, Saptarshi, et al.
Published: (2024)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026)
by: Teixeira, Tiago, et al.
Published: (2026)
KACE: Knowledge-Adaptive Context Engineering for Mathematical Reasoning
by: Parashar, Jayant, et al.
Published: (2026)
by: Parashar, Jayant, et al.
Published: (2026)
Measuring Reasoning Utility in LLMs via Conditional Entropy Reduction
by: Guo, Xu
Published: (2025)
by: Guo, Xu
Published: (2025)
Temporal Knowledge Question Answering via Abstract Reasoning Induction
by: Chen, Ziyang, et al.
Published: (2023)
by: Chen, Ziyang, et al.
Published: (2023)
Do LLMs Use Cultural Knowledge Without Being Told? A Multilingual Evaluation of Implicit Pragmatic Adaptation
by: Nasim, Mehwish, et al.
Published: (2026)
by: Nasim, Mehwish, et al.
Published: (2026)
A Survey of Task-Oriented Knowledge Graph Reasoning: Status, Applications, and Prospects
by: Niu, Guanglin, et al.
Published: (2025)
by: Niu, Guanglin, et al.
Published: (2025)
Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improves Decision Alignment
by: Su, Ruoxi, et al.
Published: (2026)
by: Su, Ruoxi, et al.
Published: (2026)
From Guessing to Asking: An Approach to Resolving the Persona Knowledge Gap in LLMs during Multi-Turn Conversations
by: Baskar, Sarvesh, et al.
Published: (2025)
by: Baskar, Sarvesh, et al.
Published: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Conversation Tree Architecture: A Structured Framework for Context-Aware Multi-Branch LLM Conversations
by: Hemanth, Pranav, et al.
Published: (2026)
by: Hemanth, Pranav, et al.
Published: (2026)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
System-Mediated Attention Imbalances Make Vision-Language Models Say Yes
by: Chan, Tsan Tsai, et al.
Published: (2026)
by: Chan, Tsan Tsai, et al.
Published: (2026)
Bielik-Minitron-7B: Compressing Large Language Models via Structured Pruning and Knowledge Distillation for the Polish Language
by: Kinas, Remigiusz, et al.
Published: (2026)
by: Kinas, Remigiusz, et al.
Published: (2026)
Red Teaming for Large Language Models At Scale: Tackling Hallucinations on Mathematics Tasks
by: Buszydlik, Aleksander, et al.
Published: (2023)
by: Buszydlik, Aleksander, et al.
Published: (2023)
Mathador-LM: A Dynamic Benchmark for Mathematical Reasoning on Large Language Models
by: Kurtic, Eldar, et al.
Published: (2024)
by: Kurtic, Eldar, et al.
Published: (2024)
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
by: Le, Van-Truong
Published: (2026)
by: Le, Van-Truong
Published: (2026)
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
by: Ansell, Rebecca, et al.
Published: (2026)
by: Ansell, Rebecca, et al.
Published: (2026)
MAWARITH: A Dataset and Benchmark for Legal Inheritance Reasoning with LLMs
by: Bouchekif, Abdessalam, et al.
Published: (2026)
by: Bouchekif, Abdessalam, et al.
Published: (2026)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
by: Chang, Hoyeon, et al.
Published: (2024)
by: Chang, Hoyeon, et al.
Published: (2024)
Detecting and Steering LLMs' Empathy in Action
by: Cadile, Juan P.
Published: (2025)
by: Cadile, Juan P.
Published: (2025)
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs
by: Mikami, Yusuke, et al.
Published: (2024)
by: Mikami, Yusuke, et al.
Published: (2024)
Reflective Translation: Improving Low-Resource Machine Translation via Structured Self-Reflection
by: Cheng, Nicholas
Published: (2026)
by: Cheng, Nicholas
Published: (2026)
Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?
by: Bajpai, Ashutosh, et al.
Published: (2025)
by: Bajpai, Ashutosh, et al.
Published: (2025)
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
by: Williamson, Dane, et al.
Published: (2025)
by: Williamson, Dane, et al.
Published: (2025)
On Preserving the Knowledge of Long Clinical Texts
by: Hasan, Mohammad Junayed, et al.
Published: (2023)
by: Hasan, Mohammad Junayed, et al.
Published: (2023)
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
by: Mueller, Felix B, et al.
Published: (2024)
by: Mueller, Felix B, et al.
Published: (2024)
Training LLMs to Recognize Hedges in Spontaneous Narratives
by: Paige, Amie J., et al.
Published: (2024)
by: Paige, Amie J., et al.
Published: (2024)
Teaching Probabilistic Logical Reasoning to Transformers
by: Nafar, Aliakbar, et al.
Published: (2023)
by: Nafar, Aliakbar, et al.
Published: (2023)
Similar Items
-
Language Models are Crossword Solvers
by: Saha, Soumadeep, et al.
Published: (2024) -
sudoLLM: On Multi-role Alignment of Language Models
by: Saha, Soumadeep, et al.
Published: (2025) -
Strong hallucinations from negation and how to fix them
by: Asher, Nicholas, et al.
Published: (2024) -
On Explaining with Attention Matrices
by: Naim, Omar, et al.
Published: (2024) -
InfinityMATH: A Scalable Instruction Tuning Dataset in Programmatic Mathematical Reasoning
by: Zhang, Bo-Wen, et al.
Published: (2024)