Modular Arithmetic: Language Models Solve Math Digit by Digit
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baeumel, Tanja, Gurgurov, Daniil, Ghussin, Yusser al, van Genabith, Josef, Ostermann, Simon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026)
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026)
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
CLaS-Bench: A Cross-Lingual Alignment and Steering Benchmark
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026)
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026)
The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs
von: Baeumel, Tanja, et al.
Veröffentlicht: (2025)
von: Baeumel, Tanja, et al.
Veröffentlicht: (2025)
Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms
von: Baeumel, Tanja, et al.
Veröffentlicht: (2026)
von: Baeumel, Tanja, et al.
Veröffentlicht: (2026)
On Multilingual Encoder Language Model Compression for Low-Resource Languages
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
Small Models, Big Impact: Efficient Corpus and Graph-Based Adaptation of Small Multilingual Language Models for Low-Resource Languages
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
Adapting Multilingual LLMs to Low-Resource Languages with Knowledge Graphs via Adapters
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
The Latin Substrate: How Language Models Represent and Mediate Script Choice
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
Multilingual Political Views of Large Language Models: Identification and Steering
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
AutoPsyC: Automatic Recognition of Psychodynamic Conflicts from Semi-structured Interviews with Large Language Models
von: Hossain, Sayed Muddashir, et al.
Veröffentlicht: (2025)
von: Hossain, Sayed Muddashir, et al.
Veröffentlicht: (2025)
From Weights to Activations: Is Steering the Next Frontier of Adaptation?
von: Ostermann, Simon, et al.
Veröffentlicht: (2026)
von: Ostermann, Simon, et al.
Veröffentlicht: (2026)
GrEmLIn: A Repository of Green Baseline Embeddings for 87 Low-Resource Languages Injected with Multilingual Graph Knowledge
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
Probing Context Localization of Polysemous Words in Pre-trained Language Model Sub-Layers
von: Vijayakumar, Soniya, et al.
Veröffentlicht: (2024)
von: Vijayakumar, Soniya, et al.
Veröffentlicht: (2024)
Multilingual Large Language Models and Curse of Multilinguality
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
LLMCheckup: Conversational Examination of Large Language Models via Interpretability Tools and Self-Explanations
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
von: Wang, Qianli, et al.
Veröffentlicht: (2024)
DualFact+: A Multimodal Fact Verification Framework for Procedural Video Understanding
von: Oguz, Cennet, et al.
Veröffentlicht: (2026)
von: Oguz, Cennet, et al.
Veröffentlicht: (2026)
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
von: Fang, Meng, et al.
Veröffentlicht: (2024)
von: Fang, Meng, et al.
Veröffentlicht: (2024)
TeleMath: A Benchmark for Large Language Models in Telecom Mathematical Problem Solving
von: Colle, Vincenzo, et al.
Veröffentlicht: (2025)
von: Colle, Vincenzo, et al.
Veröffentlicht: (2025)
MathGLM-Vision: Solving Mathematical Problems with Multi-Modal Large Language Model
von: Yang, Zhen, et al.
Veröffentlicht: (2024)
von: Yang, Zhen, et al.
Veröffentlicht: (2024)
Can Vision-Language Models Solve Visual Math Equations?
von: Choudhury, Monjoy Narayan, et al.
Veröffentlicht: (2025)
von: Choudhury, Monjoy Narayan, et al.
Veröffentlicht: (2025)
REAMS: Reasoning Enhanced Algorithm for Maths Solving
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
von: Singh, Eishkaran, et al.
Veröffentlicht: (2025)
Why Does Reinforcement Learning Generalize? A Feature-Level Mechanistic Study of Post-Training in Large Language Models
von: Shi, Dan, et al.
Veröffentlicht: (2026)
von: Shi, Dan, et al.
Veröffentlicht: (2026)
OpenFActScore: Open-Source Atomic Evaluation of Factuality in Text Generation
von: Lage, Lucas Fonseca, et al.
Veröffentlicht: (2025)
von: Lage, Lucas Fonseca, et al.
Veröffentlicht: (2025)
TabularMath: Understanding Math Reasoning over Tables with Large Language Models
von: Tian, Shi-Yu, et al.
Veröffentlicht: (2025)
von: Tian, Shi-Yu, et al.
Veröffentlicht: (2025)
Solving for X and Beyond: Can Large Language Models Solve Complex Math Problems with More-Than-Two Unknowns?
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2024)
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2024)
A Diversity-Enhanced Knowledge Distillation Model for Practical Math Word Problem Solving
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
Soft Begging: Modular and Efficient Shielding of LLMs against Prompt Injection and Jailbreaking based on Prompt Tuning
von: Ostermann, Simon, et al.
Veröffentlicht: (2024)
von: Ostermann, Simon, et al.
Veröffentlicht: (2024)
Probing for Arithmetic Errors in Language Models
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
von: Sun, Yucheng, et al.
Veröffentlicht: (2025)
DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving
von: Tong, Yuxuan, et al.
Veröffentlicht: (2024)
von: Tong, Yuxuan, et al.
Veröffentlicht: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
Arithmetic with Language Models: from Memorization to Computation
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
von: Maltoni, Davide, et al.
Veröffentlicht: (2023)
Self-training Language Models for Arithmetic Reasoning
von: Kadlčík, Marek, et al.
Veröffentlicht: (2024)
von: Kadlčík, Marek, et al.
Veröffentlicht: (2024)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
von: Alavi, Khashayar, et al.
Veröffentlicht: (2025)
von: Alavi, Khashayar, et al.
Veröffentlicht: (2025)
CresOWLve: Benchmarking Creative Problem-Solving Over Real-World Knowledge
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2026)
von: Ismayilzada, Mete, et al.
Veröffentlicht: (2026)
MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models
von: Peng, Shuai, et al.
Veröffentlicht: (2024)
von: Peng, Shuai, et al.
Veröffentlicht: (2024)
Beyond Solving Math Quiz: Evaluating the Ability of Large Reasoning Models to Ask for Information
von: Huang, Youcheng, et al.
Veröffentlicht: (2025)
von: Huang, Youcheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025) -
Multilingual Steering by Design: Multilingual Sparse Autoencoders and Principled Layer Selection
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026) -
Sparse Subnetwork Enhancement for Underrepresented Languages in Large Language Models
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025) -
CLaS-Bench: A Cross-Lingual Alignment and Steering Benchmark
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026) -
DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge
von: Ghussin, Yusser Al, et al.
Veröffentlicht: (2026)