Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Xingguang, Liu, Yahui, Wang, Qi, Zhang, Jingyuan, Yue, Yang, Shi, Rui, Sun, Chenxi, Zhang, Fuzheng, Zhou, Guorui, Gai, Kun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leanabell-Prover: Posttraining Scaling in Formal Reasoning
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
Klear-AgentForge: Forging Agentic Intelligence through Posttraining Scaling
von: Wang, Qi, et al.
Veröffentlicht: (2025)
von: Wang, Qi, et al.
Veröffentlicht: (2025)
OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving
von: Li, Chenyi, et al.
Veröffentlicht: (2026)
von: Li, Chenyi, et al.
Veröffentlicht: (2026)
Capybara-OMNI: An Efficient Paradigm for Building Omni-Modal Language Models
von: Ji, Xingguang, et al.
Veröffentlicht: (2025)
von: Ji, Xingguang, et al.
Veröffentlicht: (2025)
Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving
von: Chen, Luoxin, et al.
Veröffentlicht: (2025)
von: Chen, Luoxin, et al.
Veröffentlicht: (2025)
PhysProver: Advancing Automatic Theorem Proving for Physics
von: Zhang, Hanning, et al.
Veröffentlicht: (2026)
von: Zhang, Hanning, et al.
Veröffentlicht: (2026)
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
von: Zhang, Hongzhi, et al.
Veröffentlicht: (2025)
von: Zhang, Hongzhi, et al.
Veröffentlicht: (2025)
AR-GRPO: Training Autoregressive Image Generation Models via Reinforcement Learning
von: Yuan, Shihao, et al.
Veröffentlicht: (2025)
von: Yuan, Shihao, et al.
Veröffentlicht: (2025)
Spark-Prover-X1: Formal Theorem Proving Through Diverse Data Training
von: Zhou, Xinyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Xinyuan, et al.
Veröffentlicht: (2025)
RLEP: Reinforcement Learning with Experience Replay for LLM Reasoning
von: Zhang, Hongzhi, et al.
Veröffentlicht: (2025)
von: Zhang, Hongzhi, et al.
Veröffentlicht: (2025)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
von: Fu, Jia, et al.
Veröffentlicht: (2025)
von: Fu, Jia, et al.
Veröffentlicht: (2025)
Goedel-Prover-V2: Scaling Formal Theorem Proving with Scaffolded Data Synthesis and Self-Correction
von: Lin, Yong, et al.
Veröffentlicht: (2025)
von: Lin, Yong, et al.
Veröffentlicht: (2025)
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
GAR: Generative Adversarial Reinforcement Learning for Formal Theorem Proving
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
EvolProver: Advancing Automated Theorem Proving by Evolving Formalized Problems via Symmetry and Difficulty
von: Tian, Yuchen, et al.
Veröffentlicht: (2025)
von: Tian, Yuchen, et al.
Veröffentlicht: (2025)
Proving Properties of $φ$-Representations with the Walnut Theorem-Prover
von: Shallit, Jeffrey
Veröffentlicht: (2023)
von: Shallit, Jeffrey
Veröffentlicht: (2023)
HybridProver: Augmenting Theorem Proving with LLM-Driven Proof Synthesis and Refinement
von: Hu, Jilin, et al.
Veröffentlicht: (2025)
von: Hu, Jilin, et al.
Veröffentlicht: (2025)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
von: Li, Guchan, et al.
Veröffentlicht: (2026)
von: Li, Guchan, et al.
Veröffentlicht: (2026)
Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics
von: Breen, Benjamin, et al.
Veröffentlicht: (2025)
von: Breen, Benjamin, et al.
Veröffentlicht: (2025)
STP: Self-play LLM Theorem Provers with Iterative Conjecturing and Proving
von: Dong, Kefan, et al.
Veröffentlicht: (2025)
von: Dong, Kefan, et al.
Veröffentlicht: (2025)
EconProver: Towards More Economical Test-Time Scaling for Automated Theorem Proving
von: Li, Mukai, et al.
Veröffentlicht: (2025)
von: Li, Mukai, et al.
Veröffentlicht: (2025)
MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2025)
DreamProver: Evolving Transferable Lemma Libraries via a Wake-Sleep Theorem-Proving Agent
von: Zhang, Youyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Youyuan, et al.
Veröffentlicht: (2026)
CE-GPPO: Coordinating Entropy via Gradient-Preserving Clipping Policy Optimization in Reinforcement Learning
von: Su, Zhenpeng, et al.
Veröffentlicht: (2025)
von: Su, Zhenpeng, et al.
Veröffentlicht: (2025)
Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving
von: Lin, Yong, et al.
Veröffentlicht: (2025)
von: Lin, Yong, et al.
Veröffentlicht: (2025)
3D-Prover: Diversity Driven Theorem Proving With Determinantal Point Processes
von: Lamont, Sean, et al.
Veröffentlicht: (2024)
von: Lamont, Sean, et al.
Veröffentlicht: (2024)
OProver: A Unified Framework for Agentic Formal Theorem Proving
von: Ma, David, et al.
Veröffentlicht: (2026)
von: Ma, David, et al.
Veröffentlicht: (2026)
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning
von: Zhang, Ziyin, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyin, et al.
Veröffentlicht: (2025)
BFS-Prover: Scalable Best-First Tree Search for LLM-based Automatic Theorem Proving
von: Xin, Ran, et al.
Veröffentlicht: (2025)
von: Xin, Ran, et al.
Veröffentlicht: (2025)
Routing to the Right Expertise: A Trustworthy Judge for Instruction-based Image Editing
von: Sun, Chenxi, et al.
Veröffentlicht: (2025)
von: Sun, Chenxi, et al.
Veröffentlicht: (2025)
On Reasoning-Centric LLM-based Automated Theorem Proving
von: Sun, Yican, et al.
Veröffentlicht: (2026)
von: Sun, Yican, et al.
Veröffentlicht: (2026)
Steering LLMs for Formal Theorem Proving
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
von: Kirtania, Shashank, et al.
Veröffentlicht: (2025)
Evaluating Multimodal Large Language Models on Video Captioning via Monte Carlo Tree Search
von: Yu, Linhao, et al.
Veröffentlicht: (2025)
von: Yu, Linhao, et al.
Veröffentlicht: (2025)
MA-LoT: Model-Collaboration Lean-based Long Chain-of-Thought Reasoning enhances Formal Theorem Proving
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
von: Wang, Ruida, et al.
Veröffentlicht: (2025)
MerLean-Prover: A Recursive Looping Harness for Lean 4 Theorem Proving
von: Li, Jinzheng, et al.
Veröffentlicht: (2026)
von: Li, Jinzheng, et al.
Veröffentlicht: (2026)
Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning
von: Wang, Haiming, et al.
Veröffentlicht: (2025)
von: Wang, Haiming, et al.
Veröffentlicht: (2025)
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
von: Wang, Jianing, et al.
Veröffentlicht: (2026)
von: Wang, Jianing, et al.
Veröffentlicht: (2026)
DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
von: Ren, Z. Z., et al.
Veröffentlicht: (2025)
von: Ren, Z. Z., et al.
Veröffentlicht: (2025)
Beyond Theorem Proving: Formulation, Framework and Benchmark for Formal Problem-Solving
von: Liu, Qi, et al.
Veröffentlicht: (2025)
von: Liu, Qi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Leanabell-Prover: Posttraining Scaling in Formal Reasoning
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025) -
Klear-AgentForge: Forging Agentic Intelligence through Posttraining Scaling
von: Wang, Qi, et al.
Veröffentlicht: (2025) -
OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving
von: Li, Chenyi, et al.
Veröffentlicht: (2026) -
Capybara-OMNI: An Efficient Paradigm for Building Omni-Modal Language Models
von: Ji, Xingguang, et al.
Veröffentlicht: (2025) -
Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving
von: Chen, Luoxin, et al.
Veröffentlicht: (2025)