VisioMath: Benchmarking Figure-based Mathematical Reasoning in LMMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Can, Liu, Ying, Zhang, Ting, Wang, Mei, Huang, Hua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models
von: Zou, Chengke, et al.
Veröffentlicht: (2024)
von: Zou, Chengke, et al.
Veröffentlicht: (2024)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
von: Lu, Zimu, et al.
Veröffentlicht: (2024)
MathDoc: Benchmarking Structured Extraction and Active Refusal on Noisy Mathematics Exam Papers
von: Zhou, Chenyue, et al.
Veröffentlicht: (2026)
von: Zhou, Chenyue, et al.
Veröffentlicht: (2026)
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)
MOAT: Evaluating LMMs for Capability Integration and Instruction Grounding
von: Ye, Zhoutong, et al.
Veröffentlicht: (2025)
von: Ye, Zhoutong, et al.
Veröffentlicht: (2025)
MathCoder-VL: Bridging Vision and Code for Enhanced Multimodal Mathematical Reasoning
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
A-Bench: Are LMMs Masters at Evaluating AI-generated Images?
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
von: Lu, Pan, et al.
Veröffentlicht: (2023)
von: Lu, Pan, et al.
Veröffentlicht: (2023)
Exploring the Potential of Encoder-free Architectures in 3D LMMs
von: Tang, Yiwen, et al.
Veröffentlicht: (2025)
von: Tang, Yiwen, et al.
Veröffentlicht: (2025)
MIBench: Evaluating LMMs on Multimodal Interaction
von: Miao, Yu, et al.
Veröffentlicht: (2026)
von: Miao, Yu, et al.
Veröffentlicht: (2026)
VisioFirm: Cross-Platform AI-assisted Annotation Tool for Computer Vision
von: Ghazouali, Safouane El, et al.
Veröffentlicht: (2025)
von: Ghazouali, Safouane El, et al.
Veröffentlicht: (2025)
SciVerse: Unveiling the Knowledge Comprehension and Visual Reasoning of LMMs on Multi-modal Scientific Problems
von: Guo, Ziyu, et al.
Veröffentlicht: (2025)
von: Guo, Ziyu, et al.
Veröffentlicht: (2025)
Exploring the Spectrum of Visio-Linguistic Compositionality and Recognition
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
MathNet: A Data-Centric Approach for Printed Mathematical Expression Recognition
von: Schmitt-Koopmann, Felix M., et al.
Veröffentlicht: (2024)
von: Schmitt-Koopmann, Felix M., et al.
Veröffentlicht: (2024)
On Multi-Step Theorem Prediction via Non-Parametric Structural Priors
von: Zhao, Junbo, et al.
Veröffentlicht: (2026)
von: Zhao, Junbo, et al.
Veröffentlicht: (2026)
Multimodal Mathematical Reasoning Embedded in Aerial Vehicle Imagery: Benchmarking, Analysis, and Exploration
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
CounterCurate: Enhancing Physical and Semantic Visio-Linguistic Compositional Reasoning via Counterfactual Examples
von: Zhang, Jianrui, et al.
Veröffentlicht: (2024)
von: Zhang, Jianrui, et al.
Veröffentlicht: (2024)
GaussianProperty: Integrating Physical Properties to 3D Gaussians with LMMs
von: Xu, Xinli, et al.
Veröffentlicht: (2024)
von: Xu, Xinli, et al.
Veröffentlicht: (2024)
Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation
von: Fu, Kang, et al.
Veröffentlicht: (2026)
von: Fu, Kang, et al.
Veröffentlicht: (2026)
Vinoground: Scrutinizing LMMs over Dense Temporal Reasoning with Short Videos
von: Zhang, Jianrui, et al.
Veröffentlicht: (2024)
von: Zhang, Jianrui, et al.
Veröffentlicht: (2024)
VITAL: Vision-Encoder-centered Pre-training for LMMs in Visual Quality Assessment
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
von: Duan, Chengqi, et al.
Veröffentlicht: (2025)
OBI-Bench: Can LMMs Aid in Study of Ancient Script on Oracle Bones?
von: Chen, Zijian, et al.
Veröffentlicht: (2024)
von: Chen, Zijian, et al.
Veröffentlicht: (2024)
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
SketchJudge: A Diagnostic Benchmark for Grading Hand-drawn Diagrams with Multimodal Large Language Models
von: Su, Yuhang, et al.
Veröffentlicht: (2026)
von: Su, Yuhang, et al.
Veröffentlicht: (2026)
The Role of Visual Modality in Multimodal Mathematical Reasoning: Challenges and Insights
von: Liu, Yufang, et al.
Veröffentlicht: (2025)
von: Liu, Yufang, et al.
Veröffentlicht: (2025)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
von: Jiao, Pengkun, et al.
Veröffentlicht: (2025)
von: Jiao, Pengkun, et al.
Veröffentlicht: (2025)
Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning
von: Bai, Hongbo, et al.
Veröffentlicht: (2026)
von: Bai, Hongbo, et al.
Veröffentlicht: (2026)
CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos
von: Li, Xuchen, et al.
Veröffentlicht: (2025)
von: Li, Xuchen, et al.
Veröffentlicht: (2025)
OpenGround: Active Cognition-based Reasoning for Open-World 3D Visual Grounding
von: Huang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenyuan, et al.
Veröffentlicht: (2025)
From Prompts to Pavement: LMMs-based Agentic Behavior-Tree Generation Framework for Autonomous Vehicles
von: Goba, Omar Y., et al.
Veröffentlicht: (2026)
von: Goba, Omar Y., et al.
Veröffentlicht: (2026)
We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
von: Qiao, Runqi, et al.
Veröffentlicht: (2024)
von: Qiao, Runqi, et al.
Veröffentlicht: (2024)
OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning
von: Fu, Ling, et al.
Veröffentlicht: (2024)
von: Fu, Ling, et al.
Veröffentlicht: (2024)
V-Math: An Agentic Approach to the Vietnamese National High School Graduation Mathematics Exams
von: Nguyen, Duong Q., et al.
Veröffentlicht: (2025)
von: Nguyen, Duong Q., et al.
Veröffentlicht: (2025)
UR-Bench: A Benchmark for Multi-Hop Reasoning over Ultra-High-Resolution Images
von: Li, Siqi, et al.
Veröffentlicht: (2025)
von: Li, Siqi, et al.
Veröffentlicht: (2025)
MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision
von: Du, Lingxiao, et al.
Veröffentlicht: (2025)
von: Du, Lingxiao, et al.
Veröffentlicht: (2025)
Beyond Static Artifacts: A Forensic Benchmark for Video Deepfake Reasoning in Vision Language Models
von: Gu, Zheyuan, et al.
Veröffentlicht: (2026)
von: Gu, Zheyuan, et al.
Veröffentlicht: (2026)
Revisiting Reliability in the Reasoning-based Pose Estimation Benchmark
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning
von: Liu, Chengwen, et al.
Veröffentlicht: (2026)
von: Liu, Chengwen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024) -
DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models
von: Zou, Chengke, et al.
Veröffentlicht: (2024) -
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
von: Lu, Zimu, et al.
Veröffentlicht: (2024) -
MathDoc: Benchmarking Structured Extraction and Active Refusal on Noisy Mathematics Exam Papers
von: Zhou, Chenyue, et al.
Veröffentlicht: (2026) -
We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning
von: Qiao, Runqi, et al.
Veröffentlicht: (2025)