CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ge, Du, Xinrun, Chen, Bei, Liang, Yiming, Luo, Tongxu, Zheng, Tianyu, Zhu, Kang, Cheng, Yuyang, Xu, Chunpu, Guo, Shuyue, Zhang, Haoran, Qu, Xingwei, Wang, Junjie, Yuan, Ruibin, Li, Yizhi, Wang, Zekun, Liu, Yudong, Tsai, Yu-Hsuan, Zhang, Fengji, Lin, Chenghua, Huang, Wenhao, Fu, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contrastive Learning for Sleep Staging based on Inter Subject Correlation
by: Zhang, Tongxu, et al.
Published: (2023)
by: Zhang, Tongxu, et al.
Published: (2023)
Representation Learning of Point Cloud Upsampling in Global and Local Inputs
by: Zhang, Tongxu, et al.
Published: (2025)
by: Zhang, Tongxu, et al.
Published: (2025)
Med-PU: Point Cloud Upsampling for High-Fidelity 3D Medical Shape Reconstruction
by: Zhang, Tongxu, et al.
Published: (2025)
by: Zhang, Tongxu, et al.
Published: (2025)
Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
by: Zheng, Tianyu, et al.
Published: (2024)
by: Zheng, Tianyu, et al.
Published: (2024)
A Survey of Medical Point Cloud Shape Learning: Registration, Reconstruction and Variation
by: Zhang, Tongxu, et al.
Published: (2025)
by: Zhang, Tongxu, et al.
Published: (2025)
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
by: Qu, Xingwei, et al.
Published: (2024)
by: Qu, Xingwei, et al.
Published: (2024)
I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
by: Liang, Yiming, et al.
Published: (2024)
by: Liang, Yiming, et al.
Published: (2024)
Learning Coarse-to-Fine Osteoarthritis Representations under Noisy Hierarchical Labels
by: Zhang, Tongxu
Published: (2026)
by: Zhang, Tongxu
Published: (2026)
Rethinking Data Input for Point Cloud Upsampling
by: Zhang, Tongxu
Published: (2024)
by: Zhang, Tongxu
Published: (2024)
MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI
by: Yue, Xiang, et al.
Published: (2023)
by: Yue, Xiang, et al.
Published: (2023)
CIF-Bench: A Chinese Instruction-Following Benchmark for Evaluating the Generalizability of Large Language Models
by: LI, Yizhi, et al.
Published: (2024)
by: LI, Yizhi, et al.
Published: (2024)
The Fine Line: Navigating Large Language Model Pretraining with Down-streaming Capability Analysis
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
MIO: A Foundation Model on Multimodal Tokens
by: Wang, Zekun, et al.
Published: (2024)
by: Wang, Zekun, et al.
Published: (2024)
TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling
by: Li, Yizhi, et al.
Published: (2025)
by: Li, Yizhi, et al.
Published: (2025)
OmniBench: Towards The Future of Universal Omni-Language Models
by: Li, Yizhi, et al.
Published: (2024)
by: Li, Yizhi, et al.
Published: (2024)
COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning
by: Bai, Yuelin, et al.
Published: (2024)
by: Bai, Yuelin, et al.
Published: (2024)
MuPT: A Generative Symbolic Music Pretrained Transformer
by: Qu, Xingwei, et al.
Published: (2024)
by: Qu, Xingwei, et al.
Published: (2024)
Observing Micromotives and Macrobehavior of Large Language Models
by: Cheng, Yuyang, et al.
Published: (2024)
by: Cheng, Yuyang, et al.
Published: (2024)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
by: Du, Xinrun, et al.
Published: (2024)
by: Du, Xinrun, et al.
Published: (2024)
Overview of the NLPCC 2025 Shared Task: Gender Bias Mitigation Challenge
by: Li, Yizhi, et al.
Published: (2025)
by: Li, Yizhi, et al.
Published: (2025)
First Return, Entropy-Eliciting Explore
by: Zheng, Tianyu, et al.
Published: (2025)
by: Zheng, Tianyu, et al.
Published: (2025)
Aligning Instruction Tuning with Pre-training
by: Liang, Yiming, et al.
Published: (2025)
by: Liang, Yiming, et al.
Published: (2025)
KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
by: Ma, Kaijing, et al.
Published: (2024)
by: Ma, Kaijing, et al.
Published: (2024)
Understanding DeepResearch via Reports
by: Fan, Tianyu, et al.
Published: (2025)
by: Fan, Tianyu, et al.
Published: (2025)
DeepInnovator: Triggering the Innovative Capabilities of LLMs
by: Fan, Tianyu, et al.
Published: (2026)
by: Fan, Tianyu, et al.
Published: (2026)
MMRA: A Benchmark for Evaluating Multi-Granularity and Multi-Image Relational Association Capabilities in Large Visual Language Models
by: Wu, Siwei, et al.
Published: (2024)
by: Wu, Siwei, et al.
Published: (2024)
Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space
by: Qu, Xingwei, et al.
Published: (2025)
by: Qu, Xingwei, et al.
Published: (2025)
Stop Overvaluing Multi-Agent Debate -- We Must Rethink Evaluation and Embrace Model Heterogeneity
by: Zhang, Hangfan, et al.
Published: (2025)
by: Zhang, Hangfan, et al.
Published: (2025)
COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values
by: P Team, et al.
Published: (2025)
by: P Team, et al.
Published: (2025)
The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents
by: Wang, Xinrun, et al.
Published: (2026)
by: Wang, Xinrun, et al.
Published: (2026)
ING-VP: MLLMs cannot Play Easy Vision-based Games Yet
by: Zhang, Haoran, et al.
Published: (2024)
by: Zhang, Haoran, et al.
Published: (2024)
SciMMIR: Benchmarking Scientific Multi-modal Information Retrieval
by: Wu, Siwei, et al.
Published: (2024)
by: Wu, Siwei, et al.
Published: (2024)
A Novel Self-Evolution Framework for Large Language Models
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
An Uncertainty-Driven Adaptive Self-Alignment Framework for Large Language Models
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
Preference-Aware Memory Update for Long-Term LLM Agents
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
TEGEE: Task dEfinition Guided Expert Ensembling for Generalizable and Few-shot Learning
by: Qu, Xingwei, et al.
Published: (2024)
by: Qu, Xingwei, et al.
Published: (2024)
From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution
by: Wang, Junjie, et al.
Published: (2026)
by: Wang, Junjie, et al.
Published: (2026)
Pure Core Sets of $n \times n$ Matrices over Finite Fields
by: Wang, Hongyu, et al.
Published: (2025)
by: Wang, Hongyu, et al.
Published: (2025)
Language Model as an Annotator: Unsupervised Context-aware Quality Phrase Generation
by: Zhang, Zhihao, et al.
Published: (2023)
by: Zhang, Zhihao, et al.
Published: (2023)
Harnessing Machine Learning for Discerning AI-Generated Synthetic Images
by: Wang, Yuyang, et al.
Published: (2024)
by: Wang, Yuyang, et al.
Published: (2024)
Similar Items
-
Contrastive Learning for Sleep Staging based on Inter Subject Correlation
by: Zhang, Tongxu, et al.
Published: (2023) -
Representation Learning of Point Cloud Upsampling in Global and Local Inputs
by: Zhang, Tongxu, et al.
Published: (2025) -
Med-PU: Point Cloud Upsampling for High-Fidelity 3D Medical Shape Reconstruction
by: Zhang, Tongxu, et al.
Published: (2025) -
Kun: Answer Polishment for Chinese Self-Alignment with Instruction Back-Translation
by: Zheng, Tianyu, et al.
Published: (2024) -
A Survey of Medical Point Cloud Shape Learning: Registration, Reconstruction and Variation
by: Zhang, Tongxu, et al.
Published: (2025)