Token-by-Token Regeneration and Domain Biases: A Benchmark of LLMs on Advanced Mathematical Problem-Solving
Fuente:
arXiv
Saved in:
| Main Author: | Evstafev, Evgenii |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
The Paradox of Stochasticity: Limited Creativity and Computational Decoupling in Temperature-Varied LLM Outputs of Structured Fictional Data
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
Optimizing Humor Generation in Large Language Models: Temperature Configurations and Architectural Trade-offs
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
Benchmarking Multimodal Models for Fine-Grained Image Analysis: A Comparative Study Across Diverse Visual Features
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
An Investigation of Robustness of LLMs in Mathematical Reasoning: Benchmarking with Mathematically-Equivalent Transformation of Advanced Mathematical Problems
by: Hao, Yuren, et al.
Published: (2025)
by: Hao, Yuren, et al.
Published: (2025)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Adaptive Tokenization: On the Hop-Overpriority Problem in Tokenized Graph Learning Models
by: Wang, Zhibiao, et al.
Published: (2025)
by: Wang, Zhibiao, et al.
Published: (2025)
GRAFT: Graph-Tokenized LLMs for Tool Planning
by: Gao, Xinyi, et al.
Published: (2026)
by: Gao, Xinyi, et al.
Published: (2026)
Toward a Theory of Tokenization in LLMs
by: Rajaraman, Nived, et al.
Published: (2024)
by: Rajaraman, Nived, et al.
Published: (2024)
Adaptive Token-Weighted Differential Privacy for LLMs: Not All Tokens Require Equal Protection
by: Yu, Manjiang, et al.
Published: (2025)
by: Yu, Manjiang, et al.
Published: (2025)
XFinBench: Benchmarking LLMs in Complex Financial Problem Solving and Reasoning
by: Zhang, Zhihan, et al.
Published: (2025)
by: Zhang, Zhihan, et al.
Published: (2025)
Physics Informed Token Transformer for Solving Partial Differential Equations
by: Lorsung, Cooper, et al.
Published: (2023)
by: Lorsung, Cooper, et al.
Published: (2023)
Rethinking Thinking Tokens: LLMs as Improvement Operators
by: Madaan, Lovish, et al.
Published: (2025)
by: Madaan, Lovish, et al.
Published: (2025)
Silent Tokens, Loud Effects: Padding in LLMs
by: Himelstein, Rom, et al.
Published: (2025)
by: Himelstein, Rom, et al.
Published: (2025)
Persuasion Tokens for Editing Factual Knowledge in LLMs
by: Youssef, Paul, et al.
Published: (2026)
by: Youssef, Paul, et al.
Published: (2026)
Learning to Route LLMs with Confidence Tokens
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
Robust Hallucination Detection in LLMs via Adaptive Token Selection
by: Niu, Mengjia, et al.
Published: (2025)
by: Niu, Mengjia, et al.
Published: (2025)
QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs
by: Noh, Kanghyun, et al.
Published: (2026)
by: Noh, Kanghyun, et al.
Published: (2026)
Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications
by: Zeinali, Farshad, et al.
Published: (2026)
by: Zeinali, Farshad, et al.
Published: (2026)
Character-level Tokenizations as Powerful Inductive Biases for RNA Foundational Models
by: Morales-Pastor, Adrián, et al.
Published: (2024)
by: Morales-Pastor, Adrián, et al.
Published: (2024)
TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
by: Zhang, Yuxiang, et al.
Published: (2025)
by: Zhang, Yuxiang, et al.
Published: (2025)
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs
by: Singh, Aaditya K., et al.
Published: (2024)
by: Singh, Aaditya K., et al.
Published: (2024)
Improving Self Consistency in LLMs through Probabilistic Tokenization
by: Sathe, Ashutosh, et al.
Published: (2024)
by: Sathe, Ashutosh, et al.
Published: (2024)
TokenFormer: Rethinking Transformer Scaling with Tokenized Model Parameters
by: Wang, Haiyang, et al.
Published: (2024)
by: Wang, Haiyang, et al.
Published: (2024)
Lost in Tokenization: Fundamental Trade-offs in Graph Tokenization for Transformers
by: Bechler-Speicher, Maya, et al.
Published: (2026)
by: Bechler-Speicher, Maya, et al.
Published: (2026)
Bridging the Training-Inference Gap in LLMs by Leveraging Self-Generated Tokens
by: Cen, Zhepeng, et al.
Published: (2024)
by: Cen, Zhepeng, et al.
Published: (2024)
STILL: Selecting Tokens for Intra-Layer Hybrid Attention to Linearize LLMs
by: Meng, Weikang, et al.
Published: (2026)
by: Meng, Weikang, et al.
Published: (2026)
Hypertokens: Holographic Associative Memory in Tokenized LLMs
by: Augeri, Christopher James
Published: (2025)
by: Augeri, Christopher James
Published: (2025)
Protein Structure Tokenization: Benchmarking and New Recipe
by: Yuan, Xinyu, et al.
Published: (2025)
by: Yuan, Xinyu, et al.
Published: (2025)
Tracing Mathematical Proficiency Through Problem-Solving Processes
by: Park, Jungyang, et al.
Published: (2025)
by: Park, Jungyang, et al.
Published: (2025)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
by: Nakkiran, Preetum, et al.
Published: (2025)
by: Nakkiran, Preetum, et al.
Published: (2025)
Token-Controlled Re-ranking for Sequential Recommendation via LLMs
by: Dai, Wenxi, et al.
Published: (2025)
by: Dai, Wenxi, et al.
Published: (2025)
Beyond Multi-Token Prediction: Pretraining LLMs with Future Summaries
by: Mahajan, Divyat, et al.
Published: (2025)
by: Mahajan, Divyat, et al.
Published: (2025)
(Token-Level) InfoRMIA: Stronger Membership Inference and Memorization Assessment for LLMs
by: Tao, Jiashu, et al.
Published: (2025)
by: Tao, Jiashu, et al.
Published: (2025)
Curriculum Design for Trajectory-Constrained Agent: Compressing Chain-of-Thought Tokens in LLMs
by: Tzannetos, Georgios, et al.
Published: (2025)
by: Tzannetos, Georgios, et al.
Published: (2025)
Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs
by: Guo, Tianyu, et al.
Published: (2024)
by: Guo, Tianyu, et al.
Published: (2024)
TokenButler: Token Importance is Predictable
by: Akhauri, Yash, et al.
Published: (2025)
by: Akhauri, Yash, et al.
Published: (2025)
Less Data, More Security: Advancing Cybersecurity LLMs Specialization via Resource-Efficient Domain-Adaptive Continuous Pre-training with Minimal Tokens
by: Salahuddin, Salahuddin, et al.
Published: (2025)
by: Salahuddin, Salahuddin, et al.
Published: (2025)
Do LLMs Encode Functional Importance of Reasoning Tokens?
by: Singh, Janvijay, et al.
Published: (2026)
by: Singh, Janvijay, et al.
Published: (2026)
CAD-Tokenizer: Towards Text-based CAD Prototyping via Modality-Specific Tokenization
by: Wang, Ruiyu, et al.
Published: (2025)
by: Wang, Ruiyu, et al.
Published: (2025)
Similar Items
-
Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH
by: Evstafev, Evgenii
Published: (2025) -
The Paradox of Stochasticity: Limited Creativity and Computational Decoupling in Temperature-Varied LLM Outputs of Structured Fictional Data
by: Evstafev, Evgenii
Published: (2025) -
Optimizing Humor Generation in Large Language Models: Temperature Configurations and Architectural Trade-offs
by: Evstafev, Evgenii
Published: (2025) -
Benchmarking Multimodal Models for Fine-Grained Image Analysis: A Comparative Study Across Diverse Visual Features
by: Evstafev, Evgenii
Published: (2025) -
An Investigation of Robustness of LLMs in Mathematical Reasoning: Benchmarking with Mathematically-Equivalent Transformation of Advanced Mathematical Problems
by: Hao, Yuren, et al.
Published: (2025)