MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Zimu, Zhou, Aojun, Ren, Houxing, Wang, Ke, Shi, Weikang, Pan, Junting, Zhan, Mingjie, Li, Hongsheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
por: Lu, Zimu, et al.
Publicado: (2024)
por: Lu, Zimu, et al.
Publicado: (2024)
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
por: Lu, Zimu, et al.
Publicado: (2024)
por: Lu, Zimu, et al.
Publicado: (2024)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
por: Wang, Ke, et al.
Publicado: (2025)
por: Wang, Ke, et al.
Publicado: (2025)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
por: Yang, Yunqiao, et al.
Publicado: (2025)
por: Yang, Yunqiao, et al.
Publicado: (2025)
Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
por: Wang, Ke, et al.
Publicado: (2024)
por: Wang, Ke, et al.
Publicado: (2024)
Alignment with Fill-In-the-Middle for Enhancing Code Generation
por: Ren, Houxing, et al.
Publicado: (2025)
por: Ren, Houxing, et al.
Publicado: (2025)
From Solver to Tutor: Evaluating the Pedagogical Intelligence of LLMs with KMP-Bench
por: Shi, Weikang, et al.
Publicado: (2026)
por: Shi, Weikang, et al.
Publicado: (2026)
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
por: Ren, Houxing, et al.
Publicado: (2024)
por: Ren, Houxing, et al.
Publicado: (2024)
WebGen-Bench: Evaluating LLMs on Generating Interactive and Functional Websites from Scratch
por: Lu, Zimu, et al.
Publicado: (2025)
por: Lu, Zimu, et al.
Publicado: (2025)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
por: Shi, Weikang, et al.
Publicado: (2025)
por: Shi, Weikang, et al.
Publicado: (2025)
WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning
por: Lu, Zimu, et al.
Publicado: (2025)
por: Lu, Zimu, et al.
Publicado: (2025)
VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
por: Wang, Ke, et al.
Publicado: (2025)
por: Wang, Ke, et al.
Publicado: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
por: Lu, Zimu, et al.
Publicado: (2026)
por: Lu, Zimu, et al.
Publicado: (2026)
Edit-Based Refinement for Parallel Masked Diffusion Language Models
por: Ren, Houxing, et al.
Publicado: (2026)
por: Ren, Houxing, et al.
Publicado: (2026)
Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning
por: Ren, Houxing, et al.
Publicado: (2026)
por: Ren, Houxing, et al.
Publicado: (2026)
Empowering Character-level Text Infilling by Eliminating Sub-Tokens
por: Ren, Houxing, et al.
Publicado: (2024)
por: Ren, Houxing, et al.
Publicado: (2024)
SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
por: Yang, Yunqiao, et al.
Publicado: (2026)
por: Yang, Yunqiao, et al.
Publicado: (2026)
UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents
por: Xiao, Han, et al.
Publicado: (2025)
por: Xiao, Han, et al.
Publicado: (2025)
Navi-plus: Managing Ambiguous GUI Navigation Tasks with Follow-up Questions
por: Cheng, Ziming, et al.
Publicado: (2025)
por: Cheng, Ziming, et al.
Publicado: (2025)
MathSmith: Towards Extremely Hard Mathematical Reasoning by Forging Synthetic Problems with a Reinforced Policy
por: Zhan, Shaoxiong, et al.
Publicado: (2025)
por: Zhan, Shaoxiong, et al.
Publicado: (2025)
GenieBlue: Integrating both Linguistic and Multimodal Capabilities for Large Language Models on Mobile Devices
por: Lu, Xudong, et al.
Publicado: (2025)
por: Lu, Xudong, et al.
Publicado: (2025)
LM-Searcher: Cross-domain Neural Architecture Search with LLMs via Unified Numerical Encoding
por: Hu, Yuxuan, et al.
Publicado: (2025)
por: Hu, Yuxuan, et al.
Publicado: (2025)
MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning
por: Chen, Xinyan, et al.
Publicado: (2025)
por: Chen, Xinyan, et al.
Publicado: (2025)
MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
por: Zhang, Renrui, et al.
Publicado: (2024)
por: Zhang, Renrui, et al.
Publicado: (2024)
MathClean: A Benchmark for Synthetic Mathematical Data Cleaning
por: Liang, Hao, et al.
Publicado: (2025)
por: Liang, Hao, et al.
Publicado: (2025)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models
por: Liu, Wentao, et al.
Publicado: (2024)
por: Liu, Wentao, et al.
Publicado: (2024)
NODI: Out-Of-Distribution Detection with Noise from Diffusion
por: Zhou, Jingqiu, et al.
Publicado: (2024)
por: Zhou, Jingqiu, et al.
Publicado: (2024)
Can LLMs $\textit{understand}$ Math? -- Exploring the Pitfalls in Mathematical Reasoning
por: Roy, Tiasa Singha, et al.
Publicado: (2025)
por: Roy, Tiasa Singha, et al.
Publicado: (2025)
MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs
por: Wang, Lei, et al.
Publicado: (2024)
por: Wang, Lei, et al.
Publicado: (2024)
SAND-Math: Using LLMs to Generate Novel, Difficult and Useful Mathematics Questions and Answers
por: Manem, Chaitanya, et al.
Publicado: (2025)
por: Manem, Chaitanya, et al.
Publicado: (2025)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
por: Ma, Jingkun, et al.
Publicado: (2024)
por: Ma, Jingkun, et al.
Publicado: (2024)
SpiritSight Agent: Advanced GUI Agent with One Look
por: Huang, Zhiyuan, et al.
Publicado: (2025)
por: Huang, Zhiyuan, et al.
Publicado: (2025)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
por: Lu, Pan, et al.
Publicado: (2023)
por: Lu, Pan, et al.
Publicado: (2023)
QuesGenie: Intelligent Multimodal Question Generation
por: Mubarak, Ahmed, et al.
Publicado: (2025)
por: Mubarak, Ahmed, et al.
Publicado: (2025)
Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning
por: Zhuang, Wenwen, et al.
Publicado: (2024)
por: Zhuang, Wenwen, et al.
Publicado: (2024)
DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
por: Shao, Zhihong, et al.
Publicado: (2025)
por: Shao, Zhihong, et al.
Publicado: (2025)
WirelessMathLM: Teaching Mathematical Reasoning for LLMs in Wireless Communications with Reinforcement Learning
por: Li, Xin, et al.
Publicado: (2025)
por: Li, Xin, et al.
Publicado: (2025)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
por: Yu, Longhui, et al.
Publicado: (2023)
por: Yu, Longhui, et al.
Publicado: (2023)
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task
por: Yan, Yuchen, et al.
Publicado: (2025)
por: Yan, Yuchen, et al.
Publicado: (2025)
Ejemplares similares
-
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
por: Lu, Zimu, et al.
Publicado: (2024) -
Step-Controlled DPO: Leveraging Stepwise Error for Enhanced Mathematical Reasoning
por: Lu, Zimu, et al.
Publicado: (2024) -
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
por: Wang, Ke, et al.
Publicado: (2025) -
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
por: Yang, Yunqiao, et al.
Publicado: (2025) -
Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
por: Wang, Ke, et al.
Publicado: (2024)