Numina-Lean-Agent: An Open and General Agentic Reasoning System for Formal Mathematics
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Junqi, Zhou, Zihao, Zhu, Zekai, Santos, Marco Dos, He, Weikun, Liu, Jiawei, Wang, Ran, Xie, Yunzhou, Zhao, Junqiao, Wang, Qiufeng, Zhi, Lihong, Li, Jia, Li, Wenda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Formal Proof of the Irrationality of $ζ(3)$ in Lean 4
di: Liu, Junqi, et al.
Pubblicazione: (2025)
di: Liu, Junqi, et al.
Pubblicazione: (2025)
Formalizing Gröbner Basis Theory in Lean
di: Guo, Junyu, et al.
Pubblicazione: (2026)
di: Guo, Junyu, et al.
Pubblicazione: (2026)
Automated Tactics for Polynomial Reasoning in Lean 4
di: Shen, Hao, et al.
Pubblicazione: (2026)
di: Shen, Hao, et al.
Pubblicazione: (2026)
LeanGeo: Formalizing Competitional Geometry problems in Lean
di: Song, Chendong, et al.
Pubblicazione: (2025)
di: Song, Chendong, et al.
Pubblicazione: (2025)
Kimina Lean Server: A High-Performance Lean Server for Large-Scale Verification
di: Santos, Marco Dos, et al.
Pubblicazione: (2025)
di: Santos, Marco Dos, et al.
Pubblicazione: (2025)
Formalizing Wu-Ritt Method in Lean 4
di: Xiao, Yuxuan, et al.
Pubblicazione: (2026)
di: Xiao, Yuxuan, et al.
Pubblicazione: (2026)
CombiBench: Benchmarking LLM Capability for Combinatorial Mathematics
di: Liu, Junqi, et al.
Pubblicazione: (2025)
di: Liu, Junqi, et al.
Pubblicazione: (2025)
CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization
di: Peng, Zhongyuan, et al.
Pubblicazione: (2025)
di: Peng, Zhongyuan, et al.
Pubblicazione: (2025)
Formal Mathematical Reasoning: A New Frontier in AI
di: Yang, Kaiyu, et al.
Pubblicazione: (2024)
di: Yang, Kaiyu, et al.
Pubblicazione: (2024)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
di: Zhou, Zihao, et al.
Pubblicazione: (2024)
Automated Formal Proofs of Combinatorial Identities via Wilf-Zeilberger Guidance and LLMs
di: Xiong, Beibei, et al.
Pubblicazione: (2026)
di: Xiong, Beibei, et al.
Pubblicazione: (2026)
An Algorithm for Diagonalizing Matrices of Formal Power Series
di: Dai, Zihao, et al.
Pubblicazione: (2026)
di: Dai, Zihao, et al.
Pubblicazione: (2026)
Mechanic: Sorrifier-Driven Formal Decomposition Workflow for Automated Theorem Proving
di: Qiu, Ruichen, et al.
Pubblicazione: (2026)
di: Qiu, Ruichen, et al.
Pubblicazione: (2026)
OpenHA: A Series of Open-Source Hierarchical Agentic Models in Minecraft
di: Wang, Zihao, et al.
Pubblicazione: (2025)
di: Wang, Zihao, et al.
Pubblicazione: (2025)
FormalMATH: Benchmarking Formal Mathematical Reasoning of Large Language Models
di: Yu, Zhouliang, et al.
Pubblicazione: (2025)
di: Yu, Zhouliang, et al.
Pubblicazione: (2025)
A Formally Verified Library of Mathematical Finance in Lean 4
di: Coelho, Raphael
Pubblicazione: (2026)
di: Coelho, Raphael
Pubblicazione: (2026)
Scaling the Scaling Logic: Agentic Meta-Synthesis of Logic Reasoning
di: Liu, Bowen, et al.
Pubblicazione: (2026)
di: Liu, Bowen, et al.
Pubblicazione: (2026)
Rethinking Wireless Communications through Formal Mathematical AI Reasoning
di: Zhao, Changyuan, et al.
Pubblicazione: (2026)
di: Zhao, Changyuan, et al.
Pubblicazione: (2026)
MA-LoT: Model-Collaboration Lean-based Long Chain-of-Thought Reasoning enhances Formal Theorem Proving
di: Wang, Ruida, et al.
Pubblicazione: (2025)
di: Wang, Ruida, et al.
Pubblicazione: (2025)
Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning
di: Wang, Haiming, et al.
Pubblicazione: (2025)
di: Wang, Haiming, et al.
Pubblicazione: (2025)
Can MLLMs Absorb Math Reasoning Abilities from LLMs as Free Lunch?
di: Hu, Yijie, et al.
Pubblicazione: (2025)
di: Hu, Yijie, et al.
Pubblicazione: (2025)
Allocating Mixed Goods with Customized Fairness and Indivisibility Ratio
di: Li, Bo, et al.
Pubblicazione: (2024)
di: Li, Bo, et al.
Pubblicazione: (2024)
Prover Agent: An Agent-Based Framework for Formal Mathematical Proofs
di: Baba, Kaito, et al.
Pubblicazione: (2025)
di: Baba, Kaito, et al.
Pubblicazione: (2025)
Whitney Stratification of Algebraic Boundaries of Convex Semi-algebraic Sets
di: Dai, Zihao, et al.
Pubblicazione: (2024)
di: Dai, Zihao, et al.
Pubblicazione: (2024)
Discover and Prove: An Open-source Agentic Framework for Hard Mode Automated Theorem Proving in Lean 4
di: Liu, Chengwu, et al.
Pubblicazione: (2026)
di: Liu, Chengwu, et al.
Pubblicazione: (2026)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
di: Shen, Ziju, et al.
Pubblicazione: (2025)
di: Shen, Ziju, et al.
Pubblicazione: (2025)
Open-Ended Video Game Glitch Detection with Agentic Reasoning and Temporal Grounding
di: Zheng, Muyang, et al.
Pubblicazione: (2026)
di: Zheng, Muyang, et al.
Pubblicazione: (2026)
FANS -- Formal Answer Selection for Natural Language Math Reasoning Using Lean4
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
LeanAgent: Lifelong Learning for Formal Theorem Proving
di: Kumarappan, Adarsh, et al.
Pubblicazione: (2024)
di: Kumarappan, Adarsh, et al.
Pubblicazione: (2024)
LeanCat: A Benchmark Suite for Formal Category Theory in Lean (Part I: 1-Categories)
di: Xu, Rongge, et al.
Pubblicazione: (2025)
di: Xu, Rongge, et al.
Pubblicazione: (2025)
Mathematical Formalized Problem Solving and Theorem Proving in Different Fields in Lean 4
di: Tang, Xichen
Pubblicazione: (2024)
di: Tang, Xichen
Pubblicazione: (2024)
Construction-Verification: A Benchmark for Applied Mathematics in Lean 4
di: Yang, Bowen, et al.
Pubblicazione: (2026)
di: Yang, Bowen, et al.
Pubblicazione: (2026)
Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency
di: Yang, Shu, et al.
Pubblicazione: (2026)
di: Yang, Shu, et al.
Pubblicazione: (2026)
APOLLO: Automated LLM and Lean Collaboration for Advanced Formal Reasoning
di: Ospanov, Azim, et al.
Pubblicazione: (2025)
di: Ospanov, Azim, et al.
Pubblicazione: (2025)
FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean
di: Meadows, Jordan, et al.
Pubblicazione: (2026)
di: Meadows, Jordan, et al.
Pubblicazione: (2026)
Lean4Physics: Comprehensive Reasoning Framework for College-level Physics in Lean4
di: Li, Yuxin, et al.
Pubblicazione: (2025)
di: Li, Yuxin, et al.
Pubblicazione: (2025)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
di: Bourigault, Pauline, et al.
Pubblicazione: (2026)
di: Bourigault, Pauline, et al.
Pubblicazione: (2026)
Formally Solving Answer-Construction Problems in Lean
di: Sun, Jialiang, et al.
Pubblicazione: (2025)
di: Sun, Jialiang, et al.
Pubblicazione: (2025)
ViRC: Enhancing Visual Interleaved Mathematical CoT with Reason Chunking
di: Wang, Lihong, et al.
Pubblicazione: (2025)
di: Wang, Lihong, et al.
Pubblicazione: (2025)
Learning to Seek Evidence: A Verifiable Reasoning Agent with Causal Faithfulness Analysis
di: Huang, Yuhang, et al.
Pubblicazione: (2025)
di: Huang, Yuhang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Formal Proof of the Irrationality of $ζ(3)$ in Lean 4
di: Liu, Junqi, et al.
Pubblicazione: (2025) -
Formalizing Gröbner Basis Theory in Lean
di: Guo, Junyu, et al.
Pubblicazione: (2026) -
Automated Tactics for Polynomial Reasoning in Lean 4
di: Shen, Hao, et al.
Pubblicazione: (2026) -
LeanGeo: Formalizing Competitional Geometry problems in Lean
di: Song, Chendong, et al.
Pubblicazione: (2025) -
Kimina Lean Server: A High-Performance Lean Server for Large-Scale Verification
di: Santos, Marco Dos, et al.
Pubblicazione: (2025)