Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Di, Huang, Xiaoshui, Zhou, Dongzhan, Li, Yuqiang, Ouyang, Wanli |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
A Deep User Interface for Exploring LLaMa
by: Perumal, Divya, et al.
Published: (2025)
by: Perumal, Divya, et al.
Published: (2025)
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ
by: Allard, Marc-Antoine, et al.
Published: (2024)
by: Allard, Marc-Antoine, et al.
Published: (2024)
Training Dynamics of a 1.7B LLaMa Model: A Data-Efficient Approach
by: Li, Miles Q., et al.
Published: (2024)
by: Li, Miles Q., et al.
Published: (2024)
Benchmarking quantized LLaMa-based models on the Brazilian Secondary School Exam
by: Santos, Matheus L. O., et al.
Published: (2023)
by: Santos, Matheus L. O., et al.
Published: (2023)
Evaluating LLMs for Quotation Attribution in Literary Texts: A Case Study of LLaMa3
by: Michel, Gaspard, et al.
Published: (2024)
by: Michel, Gaspard, et al.
Published: (2024)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
by: Wu, Mengsong, et al.
Published: (2025)
by: Wu, Mengsong, et al.
Published: (2025)
PhysicsMinions: Winning Gold Medals in the Latest Physics Olympiads with a Coevolutionary Multimodal Multi-Agent System
by: Yu, Fangchen, et al.
Published: (2025)
by: Yu, Fangchen, et al.
Published: (2025)
MOOSE-Chem3: Toward Experiment-Guided Hypothesis Ranking via Simulated Experimental Feedback
by: Liu, Wanhao, et al.
Published: (2025)
by: Liu, Wanhao, et al.
Published: (2025)
Revisiting the Broken Symmetry Phase of Solid Hydrogen: A Neural Network Variational Monte Carlo Study
by: Chai, Shengdu, et al.
Published: (2025)
by: Chai, Shengdu, et al.
Published: (2025)
Online Test-time Adaptation for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2024)
by: Cui, Taoyong, et al.
Published: (2024)
ChemMLLM: Chemical Multimodal Large Language Model
by: Tan, Qian, et al.
Published: (2025)
by: Tan, Qian, et al.
Published: (2025)
Iterative Pretraining Framework for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2025)
by: Cui, Taoyong, et al.
Published: (2025)
A CLIP-Powered Framework for Robust and Generalizable Data Selection
by: Yang, Suorong, et al.
Published: (2024)
by: Yang, Suorong, et al.
Published: (2024)
ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition
by: Liu, Yujie, et al.
Published: (2025)
by: Liu, Yujie, et al.
Published: (2025)
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
by: Yang, Zonglin, et al.
Published: (2024)
by: Yang, Zonglin, et al.
Published: (2024)
NeRF-Det++: Incorporating Semantic Cues and Perspective-aware Depth Supervision for Indoor Multi-View 3D Detection
by: Huang, Chenxi, et al.
Published: (2024)
by: Huang, Chenxi, et al.
Published: (2024)
MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs
by: He, Xianglong, et al.
Published: (2025)
by: He, Xianglong, et al.
Published: (2025)
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search
by: Li, Yanjie, et al.
Published: (2024)
by: Li, Yanjie, et al.
Published: (2024)
EEFSUVA: A New Mathematical Olympiad Benchmark
by: Khatibi, Nicole N, et al.
Published: (2025)
by: Khatibi, Nicole N, et al.
Published: (2025)
GPT4Vis: What Can GPT-4 Do for Zero-shot Visual Recognition?
by: Wu, Wenhao, et al.
Published: (2023)
by: Wu, Wenhao, et al.
Published: (2023)
CrystalX: High-accuracy Crystal Structure Analysis Using Deep Learning
by: Zheng, Kaipeng, et al.
Published: (2024)
by: Zheng, Kaipeng, et al.
Published: (2024)
Report on the 61st Annual International Mathematical Olympiad
by: Bajnok, Bela, et al.
Published: (2024)
by: Bajnok, Bela, et al.
Published: (2024)
Report on the 63rd Annual International Mathematical Olympiad
by: Bajnok, Béla
Published: (2025)
by: Bajnok, Béla
Published: (2025)
Math Matters of the Past (The Very First Mathematical Olympiads)
by: Fomin, Dmitri
Published: (2025)
by: Fomin, Dmitri
Published: (2025)
Report on the 50th Annual USA Mathematical Olympiad
by: Bajnok, Bela
Published: (2024)
by: Bajnok, Bela
Published: (2024)
LOCR: Location-Guided Transformer for Optical Character Recognition
by: Sun, Yu, et al.
Published: (2024)
by: Sun, Yu, et al.
Published: (2024)
GVGEN: Text-to-3D Generation with Volumetric Representation
by: He, Xianglong, et al.
Published: (2024)
by: He, Xianglong, et al.
Published: (2024)
PromptCoT: Synthesizing Olympiad-level Problems for Mathematical Reasoning in Large Language Models
by: Zhao, Xueliang, et al.
Published: (2025)
by: Zhao, Xueliang, et al.
Published: (2025)
Evidential Deep Learning for Interatomic Potentials
by: Xu, Han, et al.
Published: (2024)
by: Xu, Han, et al.
Published: (2024)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
by: Tu, Chongjun, et al.
Published: (2025)
by: Tu, Chongjun, et al.
Published: (2025)
Brains vs. Bytes: Evaluating LLM Proficiency in Olympiad Mathematics
by: Mahdavi, Hamed, et al.
Published: (2025)
by: Mahdavi, Hamed, et al.
Published: (2025)
Report on the 12th Annual USA Junior Mathematical Olympiad
by: Bajnok, Bela, et al.
Published: (2024)
by: Bajnok, Bela, et al.
Published: (2024)
Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
by: Cherian, Anoop, et al.
Published: (2024)
by: Cherian, Anoop, et al.
Published: (2024)
SBSC: Step-By-Step Coding for Improving Mathematical Olympiad Performance
by: Singh, Kunal, et al.
Published: (2025)
by: Singh, Kunal, et al.
Published: (2025)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
CogAtom: From Cognitive Atoms to Olympiad-level Mathematical Reasoning in Large Language Models
by: Chen, Zhuofan, et al.
Published: (2025)
by: Chen, Zhuofan, et al.
Published: (2025)
Control-R: Towards controllable test-time scaling
by: Zhang, Di, et al.
Published: (2025)
by: Zhang, Di, et al.
Published: (2025)
Incremental Structure Discovery of Classification via Sequential Monte Carlo
by: Huang, Changze, et al.
Published: (2024)
by: Huang, Changze, et al.
Published: (2024)
Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving
by: Tahmid, Saad, et al.
Published: (2024)
by: Tahmid, Saad, et al.
Published: (2024)
Similar Items
-
LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
by: Zhang, Di, et al.
Published: (2024) -
A Deep User Interface for Exploring LLaMa
by: Perumal, Divya, et al.
Published: (2025) -
LLaMa-SciQ: An Educational Chatbot for Answering Science MCQ
by: Allard, Marc-Antoine, et al.
Published: (2024) -
Training Dynamics of a 1.7B LLaMa Model: A Data-Efficient Approach
by: Li, Miles Q., et al.
Published: (2024) -
Benchmarking quantized LLaMa-based models on the Brazilian Secondary School Exam
by: Santos, Matheus L. O., et al.
Published: (2023)