DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ren, Z. Z., Shao, Zhihong, Song, Junxiao, Xin, Huajian, Wang, Haocheng, Zhao, Wanjia, Zhang, Liyue, Fu, Zhe, Zhu, Qihao, Yang, Dejian, Wu, Z. F., Gou, Zhibin, Ma, Shirong, Tang, Hongxuan, Liu, Yuxuan, Gao, Wenjun, Guo, Daya, Ruan, Chong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
von: Xin, Huajian, et al.
Veröffentlicht: (2024)
DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
von: Shao, Zhihong, et al.
Veröffentlicht: (2025)
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
von: Zhao, Chenggang, et al.
Veröffentlicht: (2025)
von: Zhao, Chenggang, et al.
Veröffentlicht: (2025)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
von: Shao, Zhihong, et al.
Veröffentlicht: (2024)
von: Shao, Zhihong, et al.
Veröffentlicht: (2024)
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
von: Guo, Daya, et al.
Veröffentlicht: (2024)
von: Guo, Daya, et al.
Veröffentlicht: (2024)
DeepSeek-V3 Technical Report
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
Fine-tuning DeepSeek-OCR-2 for Molecular Structure Recognition
von: Tang, Haocheng, et al.
Veröffentlicht: (2026)
von: Tang, Haocheng, et al.
Veröffentlicht: (2026)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
von: Shen, Ziju, et al.
Veröffentlicht: (2025)
von: Shen, Ziju, et al.
Veröffentlicht: (2025)
Prover Agent: An Agent-Based Framework for Formal Mathematical Proofs
von: Baba, Kaito, et al.
Veröffentlicht: (2025)
von: Baba, Kaito, et al.
Veröffentlicht: (2025)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
Collisionless zonal-flow dynamics in quasisymmetric stellarators
von: Zhu, Hongxuan, et al.
Veröffentlicht: (2024)
von: Zhu, Hongxuan, et al.
Veröffentlicht: (2024)
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
von: DeepSeek-AI, et al.
Veröffentlicht: (2024)
DeepSeek-VL: Towards Real-World Vision-Language Understanding
von: Lu, Haoyu, et al.
Veröffentlicht: (2024)
von: Lu, Haoyu, et al.
Veröffentlicht: (2024)
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
von: DeepSeek-AI, et al.
Veröffentlicht: (2025)
A Comparison of DeepSeek and Other LLMs
von: Gao, Tianchen, et al.
Veröffentlicht: (2025)
von: Gao, Tianchen, et al.
Veröffentlicht: (2025)
DeepSeek-OCR: Contexts Optical Compression
von: Wei, Haoran, et al.
Veröffentlicht: (2025)
von: Wei, Haoran, et al.
Veröffentlicht: (2025)
Global linear drift-wave eigenmode structures on flux surfaces in stellarators: ion temperature gradient mode
von: Zhu, Hongxuan, et al.
Veröffentlicht: (2025)
von: Zhu, Hongxuan, et al.
Veröffentlicht: (2025)
Leanabell-Prover: Posttraining Scaling in Formal Reasoning
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyuan, et al.
Veröffentlicht: (2025)
A Case Study of Information-Seeking Behavior in 7-Year-Old Children in a Semistructured Situation.
von: Cooper, Linda Z.
Veröffentlicht: (2002)
von: Cooper, Linda Z.
Veröffentlicht: (2002)
FormalML: A Benchmark for Evaluating Formal Subgoal Completion in Machine Learning Theory
von: Yang, Xiao-Wen, et al.
Veröffentlicht: (2025)
von: Yang, Xiao-Wen, et al.
Veröffentlicht: (2025)
ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
von: Gou, Zhibin, et al.
Veröffentlicht: (2023)
Teaching "Foundations of Mathematics" with the Lean Theorem Prover
von: Bottoni, Mattia Luciano, et al.
Veröffentlicht: (2025)
von: Bottoni, Mattia Luciano, et al.
Veröffentlicht: (2025)
DeepSeek reshaping healthcare in China's tertiary hospitals
von: Chen, Jishizhan, et al.
Veröffentlicht: (2025)
von: Chen, Jishizhan, et al.
Veröffentlicht: (2025)
Memory Analysis on the Training Course of DeepSeek Models
von: Zhang, Ping, et al.
Veröffentlicht: (2025)
von: Zhang, Ping, et al.
Veröffentlicht: (2025)
Safety Evaluation of DeepSeek Models in Chinese Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
DeepSeek-OCR 2: Visual Causal Flow
von: Wei, Haoran, et al.
Veröffentlicht: (2026)
von: Wei, Haoran, et al.
Veröffentlicht: (2026)
Mathematical modelling and experimental investigation of dehumidifier drying of radiata pine timber¿
von: Z. F. Sun
Veröffentlicht: (2005)
von: Z. F. Sun
Veröffentlicht: (2005)
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
von: Wu, Zhiyu, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyu, et al.
Veröffentlicht: (2024)
A Review of DeepSeek Models' Key Innovative Techniques
von: Wang, Chengen, et al.
Veröffentlicht: (2025)
von: Wang, Chengen, et al.
Veröffentlicht: (2025)
Safety Evaluation and Enhancement of DeepSeek Models in Chinese Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2025)
Evaluating the Performance of the DeepSeek Model in Confidential Computing Environment
von: Dong, Ben, et al.
Veröffentlicht: (2025)
von: Dong, Ben, et al.
Veröffentlicht: (2025)
An evaluation of DeepSeek Models in Biomedical Natural Language Processing
von: Zhan, Zaifu, et al.
Veröffentlicht: (2025)
von: Zhan, Zaifu, et al.
Veröffentlicht: (2025)
Research on Collaborative Governance of AIGC Applications in the DeepSeek Era
von: Shengli Deng, et al.
Veröffentlicht: (2025)
von: Shengli Deng, et al.
Veröffentlicht: (2025)
DeepSeek Powered Solid Dosage Formulation Design and Development
von: Lin, Leqi, et al.
Veröffentlicht: (2025)
von: Lin, Leqi, et al.
Veröffentlicht: (2025)
Quantitative Analysis of Performance Drop in DeepSeek Model Quantization
von: Zhao, Enbo, et al.
Veröffentlicht: (2025)
von: Zhao, Enbo, et al.
Veröffentlicht: (2025)
Formalizing Automated Market Makers in the Lean 4 Theorem Prover
von: Pusceddu, Daniele, et al.
Veröffentlicht: (2024)
von: Pusceddu, Daniele, et al.
Veröffentlicht: (2024)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
von: Li, Guchan, et al.
Veröffentlicht: (2026)
von: Li, Guchan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
von: Xin, Huajian, et al.
Veröffentlicht: (2024) -
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
von: Xin, Huajian, et al.
Veröffentlicht: (2024) -
DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
von: Shao, Zhihong, et al.
Veröffentlicht: (2025) -
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
von: DeepSeek-AI, et al.
Veröffentlicht: (2024) -
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
von: Zhao, Chenggang, et al.
Veröffentlicht: (2025)