Augmenting Math Word Problems via Iterative Question Composing
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Haoxiong, Zhang, Yifan, Luo, Yifan, Yao, Andrew Chi-Chih |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Diagram of Thought
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
Autonomous Data Selection with Zero-shot Generative Classifiers for Mathematical Texts
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
di: Zhang, Yifan, et al.
Pubblicazione: (2024)
On the Design of KL-Regularized Policy Gradient Algorithms for LLM Reasoning
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
Tensor Product Attention Is All You Need
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
di: Ouyang, Jialin
Pubblicazione: (2025)
di: Ouyang, Jialin
Pubblicazione: (2025)
Group Representational Position Encoding
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
di: Li, Chengpeng, et al.
Pubblicazione: (2023)
di: Li, Chengpeng, et al.
Pubblicazione: (2023)
Empowering Bengali Education with AI: Solving Bengali Math Word Problems through Transformer Models
di: Era, Jalisha Jashim, et al.
Pubblicazione: (2025)
di: Era, Jalisha Jashim, et al.
Pubblicazione: (2025)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
di: Luo, Haipeng, et al.
Pubblicazione: (2025)
di: Luo, Haipeng, et al.
Pubblicazione: (2025)
Training and Evaluating Language Models with Template-based Data Generation
di: Zhang, Yifan
Pubblicazione: (2024)
di: Zhang, Yifan
Pubblicazione: (2024)
Residual Stream Duality in Modern Transformer Architectures
di: Zhang, Yifan
Pubblicazione: (2026)
di: Zhang, Yifan
Pubblicazione: (2026)
A Markov Categorical Framework for Language Modeling
di: Zhang, Yifan
Pubblicazione: (2025)
di: Zhang, Yifan
Pubblicazione: (2025)
Fill in the Blank: Exploring and Enhancing LLM Capabilities for Backward Reasoning in Math Word Problems
di: Deb, Aniruddha, et al.
Pubblicazione: (2023)
di: Deb, Aniruddha, et al.
Pubblicazione: (2023)
Meta Prompting for AI Systems
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
di: Xiong, Weimin, et al.
Pubblicazione: (2024)
di: Xiong, Weimin, et al.
Pubblicazione: (2024)
Deceiving Question-Answering Models: A Hybrid Word-Level Adversarial Approach
di: Li, Jiyao, et al.
Pubblicazione: (2024)
di: Li, Jiyao, et al.
Pubblicazione: (2024)
Reasoning, Code, or Both? How Large Language Models Handle Variations in Math Questions
di: Kutakh, Matthew
Pubblicazione: (2026)
di: Kutakh, Matthew
Pubblicazione: (2026)
MegaMath: Pushing the Limits of Open Math Corpora
di: Zhou, Fan, et al.
Pubblicazione: (2025)
di: Zhou, Fan, et al.
Pubblicazione: (2025)
Leveraging Online Olympiad-Level Math Problems for LLMs Training and Contamination-Resistant Evaluation
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2025)
di: Mahdavi, Sadegh, et al.
Pubblicazione: (2025)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
di: Wang, Zengzhi, et al.
Pubblicazione: (2023)
di: Wang, Zengzhi, et al.
Pubblicazione: (2023)
MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
di: Opedal, Andreas, et al.
Pubblicazione: (2024)
di: Opedal, Andreas, et al.
Pubblicazione: (2024)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
di: Qin, Tian, et al.
Pubblicazione: (2025)
di: Qin, Tian, et al.
Pubblicazione: (2025)
Executable Functional Abstractions: Inferring Generative Programs for Advanced Math Problems
di: Khan, Zaid, et al.
Pubblicazione: (2025)
di: Khan, Zaid, et al.
Pubblicazione: (2025)
Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling
di: Liu, Zihan, et al.
Pubblicazione: (2024)
di: Liu, Zihan, et al.
Pubblicazione: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
di: Yao, Siyang, et al.
Pubblicazione: (2026)
di: Yao, Siyang, et al.
Pubblicazione: (2026)
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
di: Luo, Haipeng, et al.
Pubblicazione: (2023)
di: Luo, Haipeng, et al.
Pubblicazione: (2023)
Investigating Bias: A Multilingual Pipeline for Generating, Solving, and Evaluating Math Problems with LLMs
di: Mahran, Mariam, et al.
Pubblicazione: (2025)
di: Mahran, Mariam, et al.
Pubblicazione: (2025)
Same Question, Different Words: A Latent Adversarial Framework for Prompt Robustness
di: Fu, Tingchen, et al.
Pubblicazione: (2025)
di: Fu, Tingchen, et al.
Pubblicazione: (2025)
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
di: Liu, Dongqi, et al.
Pubblicazione: (2023)
A Toolbox, Not a Hammer -- Multi-TAG: Scaling Math Reasoning with Multi-Tool Aggregation
di: Yao, Bohan, et al.
Pubblicazione: (2025)
di: Yao, Bohan, et al.
Pubblicazione: (2025)
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
di: Wang, Zhiyuan, et al.
Pubblicazione: (2024)
ControlMath: Controllable Data Generation Promotes Math Generalist Models
di: Chen, Nuo, et al.
Pubblicazione: (2024)
di: Chen, Nuo, et al.
Pubblicazione: (2024)
M-QUEST -- Meme Question-Understanding Evaluation on Semantics and Toxicity
di: De Giorgis, Stefano, et al.
Pubblicazione: (2026)
di: De Giorgis, Stefano, et al.
Pubblicazione: (2026)
Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
di: Yang, An, et al.
Pubblicazione: (2024)
di: Yang, An, et al.
Pubblicazione: (2024)
ChemAmp: Amplified Chemistry Tools via Composable Agents
di: Li, Zhucong, et al.
Pubblicazione: (2025)
di: Li, Zhucong, et al.
Pubblicazione: (2025)
MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
di: Zhang, Renrui, et al.
Pubblicazione: (2024)
di: Zhang, Renrui, et al.
Pubblicazione: (2024)
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Solving Formal Math Problems by Decomposition and Iterative Reflection
di: Zhou, Yichi, et al.
Pubblicazione: (2025)
di: Zhou, Yichi, et al.
Pubblicazione: (2025)
Word Embeddings Are Steers for Language Models
di: Han, Chi, et al.
Pubblicazione: (2023)
di: Han, Chi, et al.
Pubblicazione: (2023)
Documenti analoghi
-
On the Diagram of Thought
di: Zhang, Yifan, et al.
Pubblicazione: (2024) -
Autonomous Data Selection with Zero-shot Generative Classifiers for Mathematical Texts
di: Zhang, Yifan, et al.
Pubblicazione: (2024) -
On the Design of KL-Regularized Policy Gradient Algorithms for LLM Reasoning
di: Zhang, Yifan, et al.
Pubblicazione: (2025) -
Tensor Product Attention Is All You Need
di: Zhang, Yifan, et al.
Pubblicazione: (2025) -
TreeCut: A Synthetic Unanswerable Math Word Problem Dataset for LLM Hallucination Evaluation
di: Ouyang, Jialin
Pubblicazione: (2025)