A Small Math Model: Recasting Strategy Choice Theory in an LLM-Inspired Architecture
Fuente:
arXiv
Saved in:
| Main Authors: | Rahman, Roussel, Shrager, Jeff |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Recasting Continual Learning as Sequence Modeling
by: Lee, Soochan, et al.
Published: (2023)
by: Lee, Soochan, et al.
Published: (2023)
A Fragile Number Sense: Probing the Elemental Limits of Numerical Reasoning in LLMs
by: Rahman, Roussel, et al.
Published: (2025)
by: Rahman, Roussel, et al.
Published: (2025)
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models
by: Park, Taekhyun, et al.
Published: (2026)
by: Park, Taekhyun, et al.
Published: (2026)
LLM-Inspired Pretrain-Then-Finetune for Small-Data, Large-Scale Optimization
by: Zhang, Zishi, et al.
Published: (2026)
by: Zhang, Zishi, et al.
Published: (2026)
Simple and Critical Iterative Denoising: A Recasting of Discrete Diffusion in Graph Generation
by: Boget, Yoann
Published: (2025)
by: Boget, Yoann
Published: (2025)
ELIZA Reinterpreted: The world's first chatbot was not intended as a chatbot at all
by: Shrager, Jeff
Published: (2024)
by: Shrager, Jeff
Published: (2024)
Executable Archaeology: Reanimating the Logic Theorist from its IPL-V Source
by: Shrager, Jeff
Published: (2026)
by: Shrager, Jeff
Published: (2026)
The Importance of Architecture Choice in Deep Learning for Climate Applications
by: Dräger, Simon, et al.
Published: (2024)
by: Dräger, Simon, et al.
Published: (2024)
Human-Inspired Memory Architecture for LLM Agents
by: Kerestecioglu, Doga, et al.
Published: (2026)
by: Kerestecioglu, Doga, et al.
Published: (2026)
MedMamba: Recasting Mamba for Medical Time Series Classification
by: He, ZhengXiao, et al.
Published: (2026)
by: He, ZhengXiao, et al.
Published: (2026)
Foundation Model Self-Play: Open-Ended Strategy Innovation via Foundation Models
by: Dharna, Aaron, et al.
Published: (2025)
by: Dharna, Aaron, et al.
Published: (2025)
Grid2Guide: A* Enabled Small Language Model for Indoor Navigation
by: Haque, Md. Wasiul, et al.
Published: (2025)
by: Haque, Md. Wasiul, et al.
Published: (2025)
Rewriting Pre-Training Data Boosts LLM Performance in Math and Code
by: Fujii, Kazuki, et al.
Published: (2025)
by: Fujii, Kazuki, et al.
Published: (2025)
Developing a Foundation of Vector Symbolic Architectures Using Category Theory
by: Shaw, Nolan P, et al.
Published: (2025)
by: Shaw, Nolan P, et al.
Published: (2025)
Reversing the Lens: Using Explainable AI to Understand Human Expertise
by: Rahman, Roussel, et al.
Published: (2025)
by: Rahman, Roussel, et al.
Published: (2025)
Beyond Pairwise: Empowering LLM Alignment With Ranked Choice Modeling
by: Tang, Yuxuan, et al.
Published: (2025)
by: Tang, Yuxuan, et al.
Published: (2025)
The Sign Estimator: LLM Alignment in the Face of Choice Heterogeneity
by: Aouad, Ali, et al.
Published: (2025)
by: Aouad, Ali, et al.
Published: (2025)
PARAMANU-GANITA: Can Small Math Language Models Rival with Large Language Models on Mathematical Reasoning?
by: Niyogi, Mitodru, et al.
Published: (2024)
by: Niyogi, Mitodru, et al.
Published: (2024)
ControlMath: Controllable Data Generation Promotes Math Generalist Models
by: Chen, Nuo, et al.
Published: (2024)
by: Chen, Nuo, et al.
Published: (2024)
CoFrNets: Interpretable Neural Architecture Inspired by Continued Fractions
by: Puri, Isha, et al.
Published: (2025)
by: Puri, Isha, et al.
Published: (2025)
Predicting LLM Reasoning Performance with Small Proxy Model
by: Koh, Woosung, et al.
Published: (2025)
by: Koh, Woosung, et al.
Published: (2025)
MathMixup: Boosting LLM Mathematical Reasoning with Difficulty-Controllable Data Synthesis and Curriculum Learning
by: Li, Xuchen, et al.
Published: (2026)
by: Li, Xuchen, et al.
Published: (2026)
Synthesizing Attitudes, Predicting Actions (SAPA): Behavioral Theory-Guided LLMs for Ridesourcing Mode Choice Modeling
by: Sameen, Mustafa, et al.
Published: (2025)
by: Sameen, Mustafa, et al.
Published: (2025)
Small Language Models for Agentic Systems: A Survey of Architectures, Capabilities, and Deployment Trade offs
by: Sharma, Raghav, et al.
Published: (2025)
by: Sharma, Raghav, et al.
Published: (2025)
SmallKV: Small Model Assisted Compensation of KV Cache Compression for Efficient LLM Inference
by: Zhao, Yi, et al.
Published: (2025)
by: Zhao, Yi, et al.
Published: (2025)
A Human-Inspired Decoupled Architecture for Efficient Audio Representation Learning
by: Kawano, Harunori, et al.
Published: (2026)
by: Kawano, Harunori, et al.
Published: (2026)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
LeMo-NADe: Multi-Parameter Neural Architecture Discovery with LLMs
by: Rahman, Md Hafizur, et al.
Published: (2024)
by: Rahman, Md Hafizur, et al.
Published: (2024)
MathAtlas: A Benchmark for Autoformalization in the Wild
by: Patel, Nilay, et al.
Published: (2026)
by: Patel, Nilay, et al.
Published: (2026)
ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems
by: Bering, Alexander
Published: (2026)
by: Bering, Alexander
Published: (2026)
Theoretical Tensions in RLHF: Reconciling Empirical Success with Inconsistencies in Social Choice Theory
by: Xiao, Jiancong, et al.
Published: (2025)
by: Xiao, Jiancong, et al.
Published: (2025)
Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models
by: Albalak, Alon, et al.
Published: (2025)
by: Albalak, Alon, et al.
Published: (2025)
Large Language Models in Numberland: A Quick Test of Their Numerical Reasoning Abilities
by: Rahman, Roussel
Published: (2025)
by: Rahman, Roussel
Published: (2025)
Boosting LLM Reasoning via Human-Inspired Reward Shaping
by: Lin, Wenze, et al.
Published: (2026)
by: Lin, Wenze, et al.
Published: (2026)
Hymba: A Hybrid-head Architecture for Small Language Models
by: Dong, Xin, et al.
Published: (2024)
by: Dong, Xin, et al.
Published: (2024)
A Dynamical Systems-Inspired Pruning Strategy for Addressing Oversmoothing in Graph Neural Networks
by: Chakraborty, Biswadeep, et al.
Published: (2024)
by: Chakraborty, Biswadeep, et al.
Published: (2024)
Modeling Choice via Self-Attention
by: Ko, Joohwan, et al.
Published: (2023)
by: Ko, Joohwan, et al.
Published: (2023)
A Survey of Graph Transformers: Architectures, Theories and Applications
by: Yuan, Chaohao, et al.
Published: (2025)
by: Yuan, Chaohao, et al.
Published: (2025)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
by: Wang, Zengzhi, et al.
Published: (2023)
by: Wang, Zengzhi, et al.
Published: (2023)
MathBode: Measuring the Stability of LLM Reasoning using Frequency Response
by: Wang, Charles L.
Published: (2025)
by: Wang, Charles L.
Published: (2025)
Similar Items
-
Recasting Continual Learning as Sequence Modeling
by: Lee, Soochan, et al.
Published: (2023) -
A Fragile Number Sense: Probing the Elemental Limits of Numerical Reasoning in LLMs
by: Rahman, Roussel, et al.
Published: (2025) -
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models
by: Park, Taekhyun, et al.
Published: (2026) -
LLM-Inspired Pretrain-Then-Finetune for Small-Data, Large-Scale Optimization
by: Zhang, Zishi, et al.
Published: (2026) -
Simple and Critical Iterative Denoising: A Recasting of Discrete Diffusion in Graph Generation
by: Boget, Yoann
Published: (2025)