CMPhysBench: A Benchmark for Evaluating Large Language Models in Condensed Matter Physics
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Weida, Huang, Dongchen, Li, Jiatong, Yang, Tengchao, Zheng, Ziyang, Zhang, Di, Han, Dong, Chen, Benteng, Luo, Binzhao, Liu, Zhiyu, Liu, Kunling, Gao, Zhiyuan, Geng, Shiqi, Ma, Wei, Su, Jiaming, Li, Xin, Pu, Shuchen, Shui, Yuhan, Cheng, Qianjia, Dou, Zhihao, Cui, Dongfei, He, Changyong, Zeng, Jin, Xie, Zeke, Su, Mao, Zhou, Dongzhan, Li, Yuqiang, Ouyang, Wanli, Cai, Yunqi, Dai, Xi, Zhang, Shufei, Bai, Lei, Cheng, Jinguang, Fang, Zhong, Weng, Hongming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025)
by: Dou, Zhihao, et al.
Published: (2025)
Online Test-time Adaptation for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2024)
by: Cui, Taoyong, et al.
Published: (2024)
Iterative Pretraining Framework for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2025)
by: Cui, Taoyong, et al.
Published: (2025)
Control-R: Towards controllable test-time scaling
by: Zhang, Di, et al.
Published: (2025)
by: Zhang, Di, et al.
Published: (2025)
Chem-R: Learning to Reason as a Chemist
by: Wang, Weida, et al.
Published: (2025)
by: Wang, Weida, et al.
Published: (2025)
Step-GRPO: Internalizing Dynamic Early Exit for Efficient Reasoning
by: Chen, Benteng, et al.
Published: (2026)
by: Chen, Benteng, et al.
Published: (2026)
PolyReal: A Benchmark for Real-World Polymer Science Workflows
by: Liu, Wanhao, et al.
Published: (2026)
by: Liu, Wanhao, et al.
Published: (2026)
LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
Physical formula enhanced multi-task learning for pharmacokinetics prediction
by: Li, Ruifeng, et al.
Published: (2024)
by: Li, Ruifeng, et al.
Published: (2024)
Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
Evidential Deep Learning for Interatomic Potentials
by: Xu, Han, et al.
Published: (2024)
by: Xu, Han, et al.
Published: (2024)
An Agentic Framework for Autonomous Materials Computation
by: Xia, Zeyu, et al.
Published: (2025)
by: Xia, Zeyu, et al.
Published: (2025)
ChemLLM: A Chemical Large Language Model
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
Geometry-enhanced Pre-training on Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2023)
by: Cui, Taoyong, et al.
Published: (2023)
Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery
by: Li, Jiatong, et al.
Published: (2025)
by: Li, Jiatong, et al.
Published: (2025)
QCBench: Evaluating Large Language Models on Domain-Specific Quantitative Chemistry
by: Xie, Jiaqing, et al.
Published: (2025)
by: Xie, Jiaqing, et al.
Published: (2025)
MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts
by: Li, Jiatong, et al.
Published: (2024)
by: Li, Jiatong, et al.
Published: (2024)
ChemBOMAS: Accelerated BO in Chemistry with LLM-Enhanced Multi-Agent System
by: Han, Dong, et al.
Published: (2025)
by: Han, Dong, et al.
Published: (2025)
ChemMLLM: Chemical Multimodal Large Language Model
by: Tan, Qian, et al.
Published: (2025)
by: Tan, Qian, et al.
Published: (2025)
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
$C^1$-robust homoclinic tangencies
by: Li, Dongchen
Published: (2024)
by: Li, Dongchen
Published: (2024)
Blender-producing mechanisms and a dichotomy for local dynamics for heterodimensional cycles
by: Li, Dongchen
Published: (2024)
by: Li, Dongchen
Published: (2024)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
by: Wu, Mengsong, et al.
Published: (2025)
by: Wu, Mengsong, et al.
Published: (2025)
Terraced Compression Method with Automated Threshold Selection for Multidimensional Image Clustering of Heterogeneous Bodies
by: Li, Jiatong, et al.
Published: (2024)
by: Li, Jiatong, et al.
Published: (2024)
Biology-Instructions: A Dataset and Benchmark for Multi-Omics Sequence Understanding Capability of Large Language Models
by: He, Haonan, et al.
Published: (2024)
by: He, Haonan, et al.
Published: (2024)
ChemVLM: Exploring the Power of Multimodal Large Language Models in Chemistry Area
by: Li, Junxian, et al.
Published: (2024)
by: Li, Junxian, et al.
Published: (2024)
MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
by: Yang, Zonglin, et al.
Published: (2024)
by: Yang, Zonglin, et al.
Published: (2024)
MOOSE-Chem3: Toward Experiment-Guided Hypothesis Ranking via Simulated Experimental Feedback
by: Liu, Wanhao, et al.
Published: (2025)
by: Liu, Wanhao, et al.
Published: (2025)
Large Language Models are In-Context Molecule Learners
by: Li, Jiatong, et al.
Published: (2024)
by: Li, Jiatong, et al.
Published: (2024)
Multiphysics Bench: Benchmarking and Investigating Scientific Machine Learning for Multiphysics PDEs
by: Yang, Changfan, et al.
Published: (2025)
by: Yang, Changfan, et al.
Published: (2025)
Consistent Time-of-Flight Depth Denoising via Graph-Informed Geometric Attention
by: Wang, Weida, et al.
Published: (2025)
by: Wang, Weida, et al.
Published: (2025)
Dynamic reversible evolution of vicinal/bonding heteronuclear diatoms drives relay reductive C–N coupling for enhancive urea electrosynthesis
by: Su Wang, et al.
Published: (2025)
by: Su Wang, et al.
Published: (2025)
The neutral scalars of type-II 2HDM+S under the LHC
by: Li, Cheng, et al.
Published: (2026)
by: Li, Cheng, et al.
Published: (2026)
The electroweak precision constraints of the 2HDM+S
by: Li, Cheng, et al.
Published: (2025)
by: Li, Cheng, et al.
Published: (2025)
ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition
by: Liu, Yujie, et al.
Published: (2025)
by: Liu, Yujie, et al.
Published: (2025)
Training-set-free two-stage deep learning for spectroscopic data de-noising
by: Huang, Dongchen, et al.
Published: (2024)
by: Huang, Dongchen, et al.
Published: (2024)
Numerical Approximation Capacity of Neural Networks with Bounded Parameters: Do Limits Exist, and How Can They Be Measured?
by: Liu, Li, et al.
Published: (2024)
by: Liu, Li, et al.
Published: (2024)
Scaling Physical Reasoning with the PHYSICS Dataset
by: Zheng, Shenghe, et al.
Published: (2025)
by: Zheng, Shenghe, et al.
Published: (2025)
Angular $k$-uniformity and the Hyperinvariance of Holographic Codes
by: Cheng, Wanli
Published: (2025)
by: Cheng, Wanli
Published: (2025)
KdV Equation for Theta Functions on Non-commutative Tori
by: Cheng, Wanli
Published: (2024)
by: Cheng, Wanli
Published: (2024)
Similar Items
-
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025) -
Online Test-time Adaptation for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2024) -
Iterative Pretraining Framework for Interatomic Potentials
by: Cui, Taoyong, et al.
Published: (2025) -
Control-R: Towards controllable test-time scaling
by: Zhang, Di, et al.
Published: (2025) -
Chem-R: Learning to Reason as a Chemist
by: Wang, Weida, et al.
Published: (2025)