InfinityMATH: A Scalable Instruction Tuning Dataset in Programmatic Mathematical Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Bo-Wen, Yan, Yan, Li, Lin, Liu, Guang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
von: Saha, Soumadeep, et al.
Veröffentlicht: (2025)
von: Saha, Soumadeep, et al.
Veröffentlicht: (2025)
Quantized Side Tuning: Fast and Memory-Efficient Tuning of Quantized Large Language Models
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2024)
On Semantic Loss Fine-Tuning Approach for Preventing Model Collapse in Causal Reasoning
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026)
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026)
External Hippocampus: Topological Cognitive Maps for Guiding Large Language Model Reasoning
von: Yan, Jian
Veröffentlicht: (2025)
von: Yan, Jian
Veröffentlicht: (2025)
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
von: Dai, Qirun, et al.
Veröffentlicht: (2025)
SCULPT: Constraint-Guided Pruned MCTS that Carves Efficient Paths for Mathematical Reasoning
von: Fang, Qitong, et al.
Veröffentlicht: (2026)
von: Fang, Qitong, et al.
Veröffentlicht: (2026)
Mathador-LM: A Dynamic Benchmark for Mathematical Reasoning on Large Language Models
von: Kurtic, Eldar, et al.
Veröffentlicht: (2024)
von: Kurtic, Eldar, et al.
Veröffentlicht: (2024)
A Scalable Communication Protocol for Networks of Large Language Models
von: Marro, Samuele, et al.
Veröffentlicht: (2024)
von: Marro, Samuele, et al.
Veröffentlicht: (2024)
Beyond the Black Box: A Statistical Model for LLM Reasoning and Inference
von: Dalal, Siddhartha, et al.
Veröffentlicht: (2024)
von: Dalal, Siddhartha, et al.
Veröffentlicht: (2024)
FediLoRA: Practical Federated Fine-Tuning of Foundation Models Under Missing-Modality Constraints
von: Yang, Lishan, et al.
Veröffentlicht: (2025)
von: Yang, Lishan, et al.
Veröffentlicht: (2025)
RAVR: Reference-Answer-guided Variational Reasoning for Large Language Models
von: Lin, Tianqianjin, et al.
Veröffentlicht: (2025)
von: Lin, Tianqianjin, et al.
Veröffentlicht: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
Balancing Efficiency and Effectiveness: An LLM-Infused Approach for Optimized CTR Prediction
von: Zhang, Guoxiao, et al.
Veröffentlicht: (2024)
von: Zhang, Guoxiao, et al.
Veröffentlicht: (2024)
CoDA: Coding LM via Diffusion Adaptation
von: Chen, Haolin, et al.
Veröffentlicht: (2025)
von: Chen, Haolin, et al.
Veröffentlicht: (2025)
EnergyMamba: An Uncertainty-Aware Graph-Enhanced Selective State Space Model for Energy Consumption Prediction
von: Yu, Dahai, et al.
Veröffentlicht: (2026)
von: Yu, Dahai, et al.
Veröffentlicht: (2026)
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
von: Eisenstadt, Roy, et al.
Veröffentlicht: (2025)
von: Eisenstadt, Roy, et al.
Veröffentlicht: (2025)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
CircuitProbe: Predicting Reasoning Circuits in Transformers via Stability Zone Detection
von: Panuganti, Rajkiran
Veröffentlicht: (2026)
von: Panuganti, Rajkiran
Veröffentlicht: (2026)
Reasoning Large Language Model Errors Arise from Hallucinating Critical Problem Features
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
von: Heyman, Alex, et al.
Veröffentlicht: (2025)
GNN for Structural Displacement Prediction
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2026)
von: Chang, Hung-Fu, et al.
Veröffentlicht: (2026)
How Does Unfaithful Reasoning Emerge from Autoregressive Training? A Study of Synthetic Experiments
von: Wang, Fuxin, et al.
Veröffentlicht: (2026)
von: Wang, Fuxin, et al.
Veröffentlicht: (2026)
A Llama walks into the 'Bar': Efficient Supervised Fine-Tuning for Legal Reasoning in the Multi-state Bar Exam
von: Fernandes, Rean, et al.
Veröffentlicht: (2025)
von: Fernandes, Rean, et al.
Veröffentlicht: (2025)
KACE: Knowledge-Adaptive Context Engineering for Mathematical Reasoning
von: Parashar, Jayant, et al.
Veröffentlicht: (2026)
von: Parashar, Jayant, et al.
Veröffentlicht: (2026)
Exploring the Effectiveness of Instruction Tuning in Biomedical Language Processing
von: Rohanian, Omid, et al.
Veröffentlicht: (2023)
von: Rohanian, Omid, et al.
Veröffentlicht: (2023)
Meta-Learning at Scale for Large Language Models via Low-Rank Amortized Bayesian Meta-Learning
von: Zhang, Liyi, et al.
Veröffentlicht: (2025)
von: Zhang, Liyi, et al.
Veröffentlicht: (2025)
End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
von: Chen, Guanzhong, et al.
Veröffentlicht: (2025)
SSSD: Simply-Scalable Speculative Decoding
von: Marzollo, Michele, et al.
Veröffentlicht: (2024)
von: Marzollo, Michele, et al.
Veröffentlicht: (2024)
Attention Drift: What Autoregressive Speculative Decoding Models Learn
von: Eldenk, Doğaç, et al.
Veröffentlicht: (2026)
von: Eldenk, Doğaç, et al.
Veröffentlicht: (2026)
Robust Tool Use via Fission-GRPO: Learning to Recover from Execution Errors
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2026)
von: Zhang, Zhiwei, et al.
Veröffentlicht: (2026)
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
von: Chen, Haolin, et al.
Veröffentlicht: (2024)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
R2R: Efficiently Navigating Divergent Reasoning Paths with Small-Large Model Token Routing
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks
von: Zhang, Chuyifei, et al.
Veröffentlicht: (2026)
von: Zhang, Chuyifei, et al.
Veröffentlicht: (2026)
Training Dynamics Underlying Language Model Scaling Laws: Loss Deceleration and Zero-Sum Learning
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
von: Mircea, Andrei, et al.
Veröffentlicht: (2025)
Behavioural Analysis of Alignment Faking
von: Hadida, Nathaniel Mitrani, et al.
Veröffentlicht: (2026)
von: Hadida, Nathaniel Mitrani, et al.
Veröffentlicht: (2026)
REAP the Experts: Why Pruning Prevails for One-Shot MoE compression
von: Lasby, Mike, et al.
Veröffentlicht: (2025)
von: Lasby, Mike, et al.
Veröffentlicht: (2025)
Random Rule Forest (RRF): Interpretable Ensembles of LLM-Generated Questions for Predicting Startup Success
von: Griffin, Ben, et al.
Veröffentlicht: (2025)
von: Griffin, Ben, et al.
Veröffentlicht: (2025)
Beyond Memorization: Violating Privacy Via Inference with Large Language Models
von: Staab, Robin, et al.
Veröffentlicht: (2023)
von: Staab, Robin, et al.
Veröffentlicht: (2023)
Disposition Distillation at Small Scale: A Three-Arc Negative Result
von: Sadasivan, Hari
Veröffentlicht: (2026)
von: Sadasivan, Hari
Veröffentlicht: (2026)
Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space
von: Vasilenko, Vladimir
Veröffentlicht: (2026)
von: Vasilenko, Vladimir
Veröffentlicht: (2026)
Ähnliche Einträge
-
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
von: Saha, Soumadeep, et al.
Veröffentlicht: (2025) -
Quantized Side Tuning: Fast and Memory-Efficient Tuning of Quantized Large Language Models
von: Zhang, Zhengxin, et al.
Veröffentlicht: (2024) -
On Semantic Loss Fine-Tuning Approach for Preventing Model Collapse in Causal Reasoning
von: Deshmukh, Pratik, et al.
Veröffentlicht: (2026) -
External Hippocampus: Topological Cognitive Maps for Guiding Large Language Model Reasoning
von: Yan, Jian
Veröffentlicht: (2025) -
Improving Influence-based Instruction Tuning Data Selection for Balanced Learning of Diverse Capabilities
von: Dai, Qirun, et al.
Veröffentlicht: (2025)