InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ying, Huaiyuan, Zhang, Shuo, Li, Linyang, Zhou, Zhejian, Shao, Yunfan, Fei, Zhaoye, Ma, Yichuan, Hong, Jiawei, Liu, Kuikun, Wang, Ziyi, Wang, Yudong, Wu, Zijian, Li, Shuaibin, Zhou, Fengzhe, Liu, Hongwei, Zhang, Songyang, Zhang, Wenwei, Yan, Hang, Qiu, Xipeng, Wang, Jiayu, Chen, Kai, Lin, Dahua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
InternLM2.5-StepProver: Advancing Automated Theorem Proving via Critic-Guided Search
von: Wu, Zijian, et al.
Veröffentlicht: (2024)
von: Wu, Zijian, et al.
Veröffentlicht: (2024)
InternLM-Law: An Open Source Chinese Legal Large Language Model
von: Fei, Zhiwei, et al.
Veröffentlicht: (2024)
von: Fei, Zhiwei, et al.
Veröffentlicht: (2024)
InternLM2 Technical Report
von: Cai, Zheng, et al.
Veröffentlicht: (2024)
von: Cai, Zheng, et al.
Veröffentlicht: (2024)
MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
von: Liu, Hongwei, et al.
Veröffentlicht: (2024)
InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model
von: Zang, Yuhang, et al.
Veröffentlicht: (2025)
von: Zang, Yuhang, et al.
Veröffentlicht: (2025)
InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
von: Dong, Xiaoyi, et al.
Veröffentlicht: (2024)
von: Dong, Xiaoyi, et al.
Veröffentlicht: (2024)
InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
von: Zhang, Pan, et al.
Veröffentlicht: (2024)
von: Zhang, Pan, et al.
Veröffentlicht: (2024)
InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD
von: Dong, Xiaoyi, et al.
Veröffentlicht: (2024)
von: Dong, Xiaoyi, et al.
Veröffentlicht: (2024)
Balanced Data Sampling for Language Model Training with Clustering
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
von: Shao, Yunfan, et al.
Veröffentlicht: (2024)
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
von: Zhang, Pan, et al.
Veröffentlicht: (2024)
von: Zhang, Pan, et al.
Veröffentlicht: (2024)
CIBench: Evaluating Your LLMs with a Code Interpreter Plugin
von: Zhang, Chuyu, et al.
Veröffentlicht: (2024)
von: Zhang, Chuyu, et al.
Veröffentlicht: (2024)
InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling
von: Li, Peiji, et al.
Veröffentlicht: (2025)
von: Li, Peiji, et al.
Veröffentlicht: (2025)
Scaling Behavior for Large Language Models regarding Numeral Systems: An Example using Pythia
von: Zhou, Zhejian, et al.
Veröffentlicht: (2024)
von: Zhou, Zhejian, et al.
Veröffentlicht: (2024)
Are Your LLMs Capable of Stable Reasoning?
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
von: Liu, Junnan, et al.
Veröffentlicht: (2024)
The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner
von: Hua, Zhouqi, et al.
Veröffentlicht: (2025)
von: Hua, Zhouqi, et al.
Veröffentlicht: (2025)
Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning
von: Lyu, Chengqi, et al.
Veröffentlicht: (2025)
von: Lyu, Chengqi, et al.
Veröffentlicht: (2025)
Unearthing Large Scale Domain-Specific Knowledge from Public Corpora
von: Fei, Zhaoye, et al.
Veröffentlicht: (2024)
von: Fei, Zhaoye, et al.
Veröffentlicht: (2024)
MegaMath: Pushing the Limits of Open Math Corpora
von: Zhou, Fan, et al.
Veröffentlicht: (2025)
von: Zhou, Fan, et al.
Veröffentlicht: (2025)
Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning
von: Zhao, Haiteng, et al.
Veröffentlicht: (2025)
von: Zhao, Haiteng, et al.
Veröffentlicht: (2025)
Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
von: Chen, Zehui, et al.
Veröffentlicht: (2024)
T-Eval: Evaluating the Tool Utilization Capability of Large Language Models Step by Step
von: Chen, Zehui, et al.
Veröffentlicht: (2023)
von: Chen, Zehui, et al.
Veröffentlicht: (2023)
Let's Verify Math Questions Step by Step
von: Shen, Chengyu, et al.
Veröffentlicht: (2025)
von: Shen, Chengyu, et al.
Veröffentlicht: (2025)
CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
von: Liu, Shudong, et al.
Veröffentlicht: (2025)
von: Liu, Shudong, et al.
Veröffentlicht: (2025)
Semi-off-Policy Reinforcement Learning for Vision-Language Slow-Thinking Reasoning
von: Shen, Junhao, et al.
Veröffentlicht: (2025)
von: Shen, Junhao, et al.
Veröffentlicht: (2025)
WirelessMathLM: Teaching Mathematical Reasoning for LLMs in Wireless Communications with Reinforcement Learning
von: Li, Xin, et al.
Veröffentlicht: (2025)
von: Li, Xin, et al.
Veröffentlicht: (2025)
CMM-Math: A Chinese Multimodal Math Dataset To Evaluate and Enhance the Mathematics Reasoning of Large Multimodal Models
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
von: Liu, Wentao, et al.
Veröffentlicht: (2024)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains
von: Zhao, Yilun, et al.
Veröffentlicht: (2023)
von: Zhao, Yilun, et al.
Veröffentlicht: (2023)
CoinMath: Harnessing the Power of Coding Instruction for Math LLMs
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?
von: Guo, Dadi, et al.
Veröffentlicht: (2026)
von: Guo, Dadi, et al.
Veröffentlicht: (2026)
AlchemistCoder: Harmonizing and Eliciting Code Capability by Hindsight Tuning on Multi-source Data
von: Song, Zifan, et al.
Veröffentlicht: (2024)
von: Song, Zifan, et al.
Veröffentlicht: (2024)
MathChat: Converse to Tackle Challenging Math Problems with LLM Agents
von: Wu, Yiran, et al.
Veröffentlicht: (2023)
von: Wu, Yiran, et al.
Veröffentlicht: (2023)
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
von: Li, Chengpeng, et al.
Veröffentlicht: (2023)
Parental Anxiety and Math Engagement: A Moderated Mediation Model of Math Anxiety and Perceived Teacher Support
von: Yu Zhou, et al.
Veröffentlicht: (2025)
von: Yu Zhou, et al.
Veröffentlicht: (2025)
TabularMath: Evaluating Computational Extrapolation in Tabular Learning via Program-Verified Synthesis
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
Turn Waste into Worth: Rectifying Top-$k$ Router of MoE
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zeng, Zhiyuan, et al.
Veröffentlicht: (2024)
Solving Formal Math Problems by Decomposition and Iterative Reflection
von: Zhou, Yichi, et al.
Veröffentlicht: (2025)
von: Zhou, Yichi, et al.
Veröffentlicht: (2025)
MathReal: We Keep It Real! A Real Scene Benchmark for Evaluating Math Reasoning in Multimodal Large Language Models
von: Feng, Jun, et al.
Veröffentlicht: (2025)
von: Feng, Jun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
InternLM2.5-StepProver: Advancing Automated Theorem Proving via Critic-Guided Search
von: Wu, Zijian, et al.
Veröffentlicht: (2024) -
InternLM-Law: An Open Source Chinese Legal Large Language Model
von: Fei, Zhiwei, et al.
Veröffentlicht: (2024) -
InternLM2 Technical Report
von: Cai, Zheng, et al.
Veröffentlicht: (2024) -
MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
von: Liu, Hongwei, et al.
Veröffentlicht: (2024) -
InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model
von: Zang, Yuhang, et al.
Veröffentlicht: (2025)