Scientists' First Exam: Probing Cognitive Abilities of MLLM via Perception, Understanding, and Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Yuhao, Wang, Yiheng, He, Xuming, Shen, Ao, Xiao, Ruoyao, Li, Zhiwei, Feng, Qiantai, Guo, Zijie, Yang, Yuejin, Wu, Hao, Huang, Wenxuan, Wei, Jiaqi, Si, Dan, Yao, Xiuqi, Bu, Jia, Huang, Haiwen, Wang, Manning, Fu, Tianfan, Tang, Shixiang, Fei, Ben, Zhou, Dongzhan, Ling, Fenghua, Lu, Yan, Sun, Siqi, Li, Chenhui, Zheng, Guanjie, Lv, Jiancheng, Zhang, Wenlong, Bai, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ChemMLLM: Chemical Multimodal Large Language Model
von: Tan, Qian, et al.
Veröffentlicht: (2025)
von: Tan, Qian, et al.
Veröffentlicht: (2025)
DeepResearch Arena: The First Exam of LLMs' Research Abilities via Seminar-Grounded Tasks
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025)
DetToolChain: A New Prompting Paradigm to Unleash Detection Ability of MLLM
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
von: Wu, Yixuan, et al.
Veröffentlicht: (2024)
SUPQA: LLM‐based Geo‐Visualization for Subjective Urban Performance Question‐Answering
von: Haiwen Huang, et al.
Veröffentlicht: (2025)
von: Haiwen Huang, et al.
Veröffentlicht: (2025)
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
von: Xu, Wanghan, et al.
Veröffentlicht: (2025)
von: Xu, Wanghan, et al.
Veröffentlicht: (2025)
Lend a Hand: Semi Training-Free Cued Speech Recognition via MLLM-Driven Hand Modeling for Barrier-free Communication
von: Huang, Guanjie, et al.
Veröffentlicht: (2025)
von: Huang, Guanjie, et al.
Veröffentlicht: (2025)
Institutionalisation and Institutional Evolution: A Model of Selecting Government Officials in Ancient China
von: Haiwen Zhou
Veröffentlicht: (2025)
von: Haiwen Zhou
Veröffentlicht: (2025)
RADAR: Revealing Asymmetric Development of Abilities in MLLM Pre-training
von: Nie, Yunshuang, et al.
Veröffentlicht: (2026)
von: Nie, Yunshuang, et al.
Veröffentlicht: (2026)
Exponential integrator Fourier Galerkin methods for semilinear parabolic equations
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
A Parareal exponential integrator finite element method for linear parabolic equations
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
von: Huang, Jianguo, et al.
Veröffentlicht: (2024)
How Far Are AI Scientists from Changing the World?
von: Xie, Qiujie, et al.
Veröffentlicht: (2025)
von: Xie, Qiujie, et al.
Veröffentlicht: (2025)
Automated Social Science: Language Models as Scientist and Subjects
von: Manning, Benjamin S., et al.
Veröffentlicht: (2024)
von: Manning, Benjamin S., et al.
Veröffentlicht: (2024)
IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering
von: Wang, Yiyao, et al.
Veröffentlicht: (2025)
von: Wang, Yiyao, et al.
Veröffentlicht: (2025)
anyECG-chat: A Generalist ECG-MLLM for Flexible ECG Input and Multi-Task Understanding
von: Li, Haitao, et al.
Veröffentlicht: (2025)
von: Li, Haitao, et al.
Veröffentlicht: (2025)
Chem3DLLM: 3D Multimodal Large Language Models for Chemistry
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wei, Jiaqi, et al.
Veröffentlicht: (2025)
SalienTime: User-driven Selection of Salient Time Steps for Large-Scale Geospatial Data Visualization
von: Chen, Juntong, et al.
Veröffentlicht: (2024)
von: Chen, Juntong, et al.
Veröffentlicht: (2024)
Calibrating Decision Robustness via Inverse Conformal Risk Control
von: Zhou, Wenbin, et al.
Veröffentlicht: (2025)
von: Zhou, Wenbin, et al.
Veröffentlicht: (2025)
Hierarchical Probabilistic Conformal Prediction for Distributed Energy Resources Adoption
von: Zhou, Wenbin, et al.
Veröffentlicht: (2024)
von: Zhou, Wenbin, et al.
Veröffentlicht: (2024)
Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
CMT: A Cascade MAR with Topology Predictor for Multimodal Conditional CAD Generation
von: Wu, Jianyu, et al.
Veröffentlicht: (2025)
von: Wu, Jianyu, et al.
Veröffentlicht: (2025)
MLLM-based Discovery of Intrinsic Coordinates and Governing Equations from High-Dimensional Data
von: Li, Ruikun, et al.
Veröffentlicht: (2025)
von: Li, Ruikun, et al.
Veröffentlicht: (2025)
A parareal exponential integrator finite element method for semilinear parabolic equations
von: Jianguo Huang, et al.
Veröffentlicht: (2024)
von: Jianguo Huang, et al.
Veröffentlicht: (2024)
$Δ$-AttnMask: Attention-Guided Masked Hidden States for Efficient Data Selection and Augmentation
von: Hu, Jucheng, et al.
Veröffentlicht: (2025)
von: Hu, Jucheng, et al.
Veröffentlicht: (2025)
SketchThinker-R1: Towards Efficient Sketch-Style Reasoning in Large Multimodal Models
von: Zhang, Ruiyang, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiyang, et al.
Veröffentlicht: (2026)
AiraXiv: An AI-Driven Open-Access Platform for Human and AI Scientists
von: Pan, Junshu, et al.
Veröffentlicht: (2026)
von: Pan, Junshu, et al.
Veröffentlicht: (2026)
Moment of Derivatives of Quadratic Twists of Modular $L$-Functions
von: Zhou, Zijie
Veröffentlicht: (2025)
von: Zhou, Zijie
Veröffentlicht: (2025)
Position: LLM Serving Needs Mathematical Optimization and Algorithmic Foundations, Not Just Heuristics
von: Zhou, Zijie
Veröffentlicht: (2026)
von: Zhou, Zijie
Veröffentlicht: (2026)
A generalization of Kadell's orthogonality ex-conjecture
von: Huang, Zihao, et al.
Veröffentlicht: (2026)
von: Huang, Zihao, et al.
Veröffentlicht: (2026)
EWE: An Agentic Framework for Extreme Weather Analysis
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
von: Jiang, Zhe, et al.
Veröffentlicht: (2025)
DMESR: Dual-view MLLM-based Enhancing Framework for Multimodal Sequential Recommendation
von: Huang, Mingyao, et al.
Veröffentlicht: (2026)
von: Huang, Mingyao, et al.
Veröffentlicht: (2026)
Optimal error estimates of the diffuse domain method for semilinear parabolic equations
von: Xu, Yuejin
Veröffentlicht: (2025)
von: Xu, Yuejin
Veröffentlicht: (2025)
APP ‐ HNTs Polyurea Composites Towards Mechanical Reinforcement, Flame‐Retardancy and Shock‐Wave Mitigation
von: Chenhui Zhu, et al.
Veröffentlicht: (2025)
von: Chenhui Zhu, et al.
Veröffentlicht: (2025)
Decoding by Perturbation: Mitigating MLLM Hallucinations via Dynamic Textual Perturbation
von: Jia, Sihang, et al.
Veröffentlicht: (2026)
von: Jia, Sihang, et al.
Veröffentlicht: (2026)
DA-DPO: Cost-efficient Difficulty-aware Preference Optimization for Reducing MLLM Hallucinations
von: Qiu, Longtian, et al.
Veröffentlicht: (2026)
von: Qiu, Longtian, et al.
Veröffentlicht: (2026)
Sel3DCraft: Interactive Visual Prompts for User-Friendly Text-to-3D Generation
von: Xiang, Nan, et al.
Veröffentlicht: (2025)
von: Xiang, Nan, et al.
Veröffentlicht: (2025)
A Survey of AI Scientists
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
von: Tie, Guiyao, et al.
Veröffentlicht: (2025)
Can MLLMs Absorb Math Reasoning Abilities from LLMs as Free Lunch?
von: Hu, Yijie, et al.
Veröffentlicht: (2025)
von: Hu, Yijie, et al.
Veröffentlicht: (2025)
UI-UG: A Unified MLLM for UI Understanding and Generation
von: Yang, Hao, et al.
Veröffentlicht: (2025)
von: Yang, Hao, et al.
Veröffentlicht: (2025)
Bootstrapping MLLM for Weakly-Supervised Class-Agnostic Object Counting
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaowen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ChemMLLM: Chemical Multimodal Large Language Model
von: Tan, Qian, et al.
Veröffentlicht: (2025) -
DeepResearch Arena: The First Exam of LLMs' Research Abilities via Seminar-Grounded Tasks
von: Wan, Haiyuan, et al.
Veröffentlicht: (2025) -
DetToolChain: A New Prompting Paradigm to Unleash Detection Ability of MLLM
von: Wu, Yixuan, et al.
Veröffentlicht: (2024) -
SUPQA: LLM‐based Geo‐Visualization for Subjective Urban Performance Question‐Answering
von: Haiwen Huang, et al.
Veröffentlicht: (2025) -
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
von: Xu, Wanghan, et al.
Veröffentlicht: (2025)