Agentic-MME: What Agentic Capability Really Brings to Multimodal Intelligence?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Qianshan, Yang, Yishan, Wang, Siyi, Chen, Jinglin, Wang, Binyu, Wang, Jiaming, Chen, Shuang, Li, Zechen, Shi, Yang, Tang, Yuqi, Wang, Weining, Yu, Yi, Fu, Chaoyou, Li, Qi, Zhang, Yi-Fan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios
von: Shi, Yang, et al.
Veröffentlicht: (2025)
von: Shi, Yang, et al.
Veröffentlicht: (2025)
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
von: Xie, Wulin, et al.
Veröffentlicht: (2025)
von: Xie, Wulin, et al.
Veröffentlicht: (2025)
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
Unified Ultrasound Intelligence Toward an End-to-End Agentic System
von: Ma, Chen, et al.
Veröffentlicht: (2026)
von: Ma, Chen, et al.
Veröffentlicht: (2026)
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
von: Yi, Dongyi, et al.
Veröffentlicht: (2025)
von: Yi, Dongyi, et al.
Veröffentlicht: (2025)
MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2024)
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
von: Fu, Chaoyou, et al.
Veröffentlicht: (2023)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2023)
MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
von: Jiang, Dongzhi, et al.
Veröffentlicht: (2025)
Octopus: Agentic Multimodal Reasoning with Six-Capability Orchestration
von: Guo, Yifu, et al.
Veröffentlicht: (2025)
von: Guo, Yifu, et al.
Veröffentlicht: (2025)
MME-Emotion: A Holistic Evaluation Benchmark for Emotional Intelligence in Multimodal Large Language Models
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Fan, et al.
Veröffentlicht: (2025)
MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning
von: Gan, Ziliang, et al.
Veröffentlicht: (2024)
von: Gan, Ziliang, et al.
Veröffentlicht: (2024)
A Unified Framework for the Evaluation of LLM Agentic Capabilities
von: Zhu, Pengyu, et al.
Veröffentlicht: (2026)
von: Zhu, Pengyu, et al.
Veröffentlicht: (2026)
HumanVideo-MME: Benchmarking MLLMs for Human-Centric Video Understanding
von: Cai, Yuxuan, et al.
Veröffentlicht: (2025)
von: Cai, Yuxuan, et al.
Veröffentlicht: (2025)
AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
von: Wang, Yixu, et al.
Veröffentlicht: (2025)
Kimi K2.5: Visual Agentic Intelligence
von: Kimi Team, et al.
Veröffentlicht: (2026)
von: Kimi Team, et al.
Veröffentlicht: (2026)
Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models
von: Liu, Yuansen, et al.
Veröffentlicht: (2025)
von: Liu, Yuansen, et al.
Veröffentlicht: (2025)
Kimi K2: Open Agentic Intelligence
von: Kimi Team, et al.
Veröffentlicht: (2025)
von: Kimi Team, et al.
Veröffentlicht: (2025)
KnowCoder-A1: Incentivizing Agentic Reasoning Capability with Outcome Supervision for KBQA
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
von: Yuan, Jiakang, et al.
Veröffentlicht: (2025)
Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
von: Guo, Garvin, et al.
Veröffentlicht: (2026)
von: Guo, Garvin, et al.
Veröffentlicht: (2026)
ManiAgent: An Agentic Framework for General Robotic Manipulation
von: Yang, Yi, et al.
Veröffentlicht: (2025)
von: Yang, Yi, et al.
Veröffentlicht: (2025)
MME-CC: A Challenging Multi-Modal Evaluation Benchmark of Cognitive Capacity
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Kaiyuan, et al.
Veröffentlicht: (2025)
Toward Autonomous Computational Catalysis Research via Agentic Systems
von: Chen, Honghao, et al.
Veröffentlicht: (2026)
von: Chen, Honghao, et al.
Veröffentlicht: (2026)
Preacher: Paper-to-Video Agentic System
von: Liu, Jingwei, et al.
Veröffentlicht: (2025)
von: Liu, Jingwei, et al.
Veröffentlicht: (2025)
EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis
von: Tang, Yi, et al.
Veröffentlicht: (2025)
von: Tang, Yi, et al.
Veröffentlicht: (2025)
From Web to Pixels: Bringing Agentic Search into Visual Perception
von: Yang, Bokang, et al.
Veröffentlicht: (2026)
von: Yang, Bokang, et al.
Veröffentlicht: (2026)
Agentic Workflows for Economic Research: Design and Implementation
von: Dawid, Herbert, et al.
Veröffentlicht: (2025)
von: Dawid, Herbert, et al.
Veröffentlicht: (2025)
AgenticGEO: A Self-Evolving Agentic System for Generative Engine Optimization
von: Yuan, Jiaqi, et al.
Veröffentlicht: (2026)
von: Yuan, Jiaqi, et al.
Veröffentlicht: (2026)
Sports Intelligence: Assessing the Sports Understanding Capabilities of Language Models through Question Answering from Text to Video
von: Yang, Zhengbang, et al.
Veröffentlicht: (2024)
von: Yang, Zhengbang, et al.
Veröffentlicht: (2024)
Agentic Episodic Control
von: Yang, Xidong, et al.
Veröffentlicht: (2025)
von: Yang, Xidong, et al.
Veröffentlicht: (2025)
AdapTools: Adaptive Tool-based Indirect Prompt Injection Attacks on Agentic LLMs
von: Wang, Che, et al.
Veröffentlicht: (2026)
von: Wang, Che, et al.
Veröffentlicht: (2026)
AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information?
von: Gong, Kaixiong, et al.
Veröffentlicht: (2024)
von: Gong, Kaixiong, et al.
Veröffentlicht: (2024)
Verifiable Process Rewards for Agentic Reasoning
von: Yuan, Huining, et al.
Veröffentlicht: (2026)
von: Yuan, Huining, et al.
Veröffentlicht: (2026)
Lita: Light Agent Uncovers the Agentic Coding Capabilities of LLMs
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
von: Dai, Hankun, et al.
Veröffentlicht: (2025)
Agentic AI-Empowered Conversational Embodied Intelligence Networks in 6G
von: Chen, Mingkai, et al.
Veröffentlicht: (2025)
von: Chen, Mingkai, et al.
Veröffentlicht: (2025)
Beyond LLaVA-HD: Diving into High-Resolution Large Multimodal Models
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2024)
Knowdit: Agentic Smart Contract Vulnerability Detection with Auditing Knowledge Summarization
von: Kong, Ziqiao, et al.
Veröffentlicht: (2026)
von: Kong, Ziqiao, et al.
Veröffentlicht: (2026)
Skywork-R1V4: Toward Agentic Multimodal Intelligence through Interleaved Thinking with Images and DeepResearch
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
ContextNav: Towards Agentic Multimodal In-Context Learning
von: Fu, Honghao, et al.
Veröffentlicht: (2025)
von: Fu, Honghao, et al.
Veröffentlicht: (2025)
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios
von: Shi, Yang, et al.
Veröffentlicht: (2025) -
MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models
von: Xie, Wulin, et al.
Veröffentlicht: (2025) -
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024) -
Unified Ultrasound Intelligence Toward an End-to-End Agentic System
von: Ma, Chen, et al.
Veröffentlicht: (2026) -
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
von: Yi, Dongyi, et al.
Veröffentlicht: (2025)