Verifier-Backed Hard Problem Generation for Mathematical Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lai, Yuhang, Feng, Jiazhan, Teh, Yee Whye, Miao, Ning |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
L3Ms -- Lagrange Large Language Models
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024)
Prompting Strategies for Enabling Large Language Models to Infer Causation from Correlation
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024)
von: Sims, Anya, et al.
Veröffentlicht: (2024)
When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
von: Singhi, Nishad, et al.
Veröffentlicht: (2025)
AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning
von: Bruno, Alessio
Veröffentlicht: (2026)
von: Bruno, Alessio
Veröffentlicht: (2026)
From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
von: Huang, Yuzhen, et al.
Veröffentlicht: (2025)
von: Huang, Yuzhen, et al.
Veröffentlicht: (2025)
An Investigation of Robustness of LLMs in Mathematical Reasoning: Benchmarking with Mathematically-Equivalent Transformation of Advanced Mathematical Problems
von: Hao, Yuren, et al.
Veröffentlicht: (2025)
von: Hao, Yuren, et al.
Veröffentlicht: (2025)
POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
von: Qu, Yuxiao, et al.
Veröffentlicht: (2026)
von: Qu, Yuxiao, et al.
Veröffentlicht: (2026)
ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning
von: Pei, Qizhi, et al.
Veröffentlicht: (2025)
von: Pei, Qizhi, et al.
Veröffentlicht: (2025)
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
Do We Need Frontier Models to Verify Mathematical Proofs?
von: Naik, Aaditya, et al.
Veröffentlicht: (2026)
von: Naik, Aaditya, et al.
Veröffentlicht: (2026)
Cog-DRIFT: Exploration on Adaptively Reformulated Instances Enables Learning from Hard Reasoning Problems
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2026)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2026)
Incorporating Unlabelled Data into Bayesian Neural Networks
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
von: Zhao, Xueliang, et al.
Veröffentlicht: (2025)
Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Models
von: Li, Miao, et al.
Veröffentlicht: (2026)
von: Li, Miao, et al.
Veröffentlicht: (2026)
SCI-Verifier: Scientific Verifier with Thinking
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
von: Zheng, Shenghe, et al.
Veröffentlicht: (2025)
RLPR: Extrapolating RLVR to General Domains without Verifiers
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
von: Yu, Tianyu, et al.
Veröffentlicht: (2025)
Hard2Verify: A Step-Level Verification Benchmark for Open-Ended Frontier Math
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
Rewarding Graph Reasoning Process makes LLMs more Generalized Reasoners
von: Peng, Miao, et al.
Veröffentlicht: (2025)
von: Peng, Miao, et al.
Veröffentlicht: (2025)
Manifold Aware Denoising Score Matching (MAD)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
von: Levy-Jurgenson, Alona, et al.
Veröffentlicht: (2026)
V-STaR: Training Verifiers for Self-Taught Reasoners
von: Hosseini, Arian, et al.
Veröffentlicht: (2024)
von: Hosseini, Arian, et al.
Veröffentlicht: (2024)
ProofSketch: Efficient Verified Reasoning for Large Language Models
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2025)
von: Sheshanarayana, Disha, et al.
Veröffentlicht: (2025)
Verifying Chain-of-Thought Reasoning via Its Computational Graph
von: Zhao, Zheng, et al.
Veröffentlicht: (2025)
von: Zhao, Zheng, et al.
Veröffentlicht: (2025)
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
von: Zhou, Jin Peng, et al.
Veröffentlicht: (2024)
REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards
von: Stojanovski, Zafir, et al.
Veröffentlicht: (2025)
von: Stojanovski, Zafir, et al.
Veröffentlicht: (2025)
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
von: Li, Xiaoyuan, et al.
Veröffentlicht: (2024)
CoVerRL: Breaking the Consensus Trap in Label-Free Reasoning via Generator-Verifier Co-Evolution
von: Pan, Teng, et al.
Veröffentlicht: (2026)
von: Pan, Teng, et al.
Veröffentlicht: (2026)
Schoenfeld's Anatomy of Mathematical Reasoning by Language Models
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
Evaluating Robustness of Reward Models for Mathematical Reasoning
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sunghwan, et al.
Veröffentlicht: (2024)
Verifying the Verifiers: Unveiling Pitfalls and Potentials in Fact Verifiers
von: Seo, Wooseok, et al.
Veröffentlicht: (2025)
von: Seo, Wooseok, et al.
Veröffentlicht: (2025)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
von: Jiang, Xitai, et al.
Veröffentlicht: (2026)
Towards Verifiable Text Generation with Symbolic References
von: Hennigen, Lucas Torroba, et al.
Veröffentlicht: (2023)
von: Hennigen, Lucas Torroba, et al.
Veröffentlicht: (2023)
ProcessBench: Identifying Process Errors in Mathematical Reasoning
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
von: Zheng, Chujie, et al.
Veröffentlicht: (2024)
Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
MathScale: Scaling Instruction Tuning for Mathematical Reasoning
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2024)
The Lessons of Developing Process Reward Models in Mathematical Reasoning
von: Zhang, Zhenru, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenru, et al.
Veröffentlicht: (2025)
Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards
von: Ma, Zhengzhao, et al.
Veröffentlicht: (2026)
von: Ma, Zhengzhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
L3Ms -- Lagrange Large Language Models
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2024) -
Prompting Strategies for Enabling Large Language Models to Infer Causation from Correlation
von: Sgouritsa, Eleni, et al.
Veröffentlicht: (2024) -
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024) -
When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning
von: Singhi, Nishad, et al.
Veröffentlicht: (2025) -
AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning
von: Bruno, Alessio
Veröffentlicht: (2026)