UPME: An Unsupervised Peer Review Framework for Multimodal Large Language Model Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Qihui, Ning, Munan, Liu, Zheyuan, Wang, Yanbo, Ye, Jiayi, Huang, Yue, Yang, Shuo, Chen, Xiao, Song, Yibing, Yuan, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
von: Lin, Bin, et al.
Veröffentlicht: (2024)
von: Lin, Bin, et al.
Veröffentlicht: (2024)
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
von: Bao, Han, et al.
Veröffentlicht: (2024)
von: Bao, Han, et al.
Veröffentlicht: (2024)
Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
von: Lin, Bin, et al.
Veröffentlicht: (2023)
von: Lin, Bin, et al.
Veröffentlicht: (2023)
PiCO: Peer Review in LLMs based on the Consistency Optimization
von: Ning, Kun-Peng, et al.
Veröffentlicht: (2024)
von: Ning, Kun-Peng, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection Using Peer-Class Generated by Large Language Model
von: Huang, K, et al.
Veröffentlicht: (2024)
von: Huang, K, et al.
Veröffentlicht: (2024)
AdaMMS: Model Merging for Heterogeneous Multimodal Large Language Models with Unsupervised Coefficient Optimization
von: Du, Yiyang, et al.
Veröffentlicht: (2025)
von: Du, Yiyang, et al.
Veröffentlicht: (2025)
V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models
von: Zheng, Xiangxi, et al.
Veröffentlicht: (2025)
von: Zheng, Xiangxi, et al.
Veröffentlicht: (2025)
MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
SDEval: Safety Dynamic Evaluation for Multimodal Large Language Models
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2023)
von: Yuan, Zhengqing, et al.
Veröffentlicht: (2023)
MIBench: Evaluating Multimodal Large Language Models over Multiple Images
von: Liu, Haowei, et al.
Veröffentlicht: (2024)
von: Liu, Haowei, et al.
Veröffentlicht: (2024)
A Survey on Evaluation of Multimodal Large Language Models
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
PRISMM-Bench: A Benchmark of Peer-Review Grounded Multimodal Inconsistencies
von: Selch, Lukas, et al.
Veröffentlicht: (2025)
von: Selch, Lukas, et al.
Veröffentlicht: (2025)
EmoLLM: Multimodal Emotional Understanding Meets Large Language Models
von: Yang, Qu, et al.
Veröffentlicht: (2024)
von: Yang, Qu, et al.
Veröffentlicht: (2024)
AvatarArtist: Open-Domain 4D Avatarization
von: Liu, Hongyu, et al.
Veröffentlicht: (2025)
von: Liu, Hongyu, et al.
Veröffentlicht: (2025)
EgoEMG: A Multimodal Egocentric Dataset with Bilateral EMG and Vision for Hand Pose Estimation
von: Xi, Ziheng, et al.
Veröffentlicht: (2026)
von: Xi, Ziheng, et al.
Veröffentlicht: (2026)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
Evaluating and Analyzing Relationship Hallucinations in Large Vision-Language Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
LoRASculpt: Sculpting LoRA for Harmonizing General and Specialized Knowledge in Multimodal Large Language Models
von: Liang, Jian, et al.
Veröffentlicht: (2025)
von: Liang, Jian, et al.
Veröffentlicht: (2025)
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability
von: Yang, Haiqi, et al.
Veröffentlicht: (2025)
von: Yang, Haiqi, et al.
Veröffentlicht: (2025)
Scene-Agnostic Traversability Labeling and Estimation via a Multimodal Self-supervised Framework
von: Fang, Zipeng, et al.
Veröffentlicht: (2025)
von: Fang, Zipeng, et al.
Veröffentlicht: (2025)
Survey of Multimodal Geospatial Foundation Models: Techniques, Applications, and Challenges
von: Yang, Liling, et al.
Veröffentlicht: (2025)
von: Yang, Liling, et al.
Veröffentlicht: (2025)
Proactive Reasoning-with-Retrieval Framework for Medical Multimodal Large Language Models
von: Wang, Lehan, et al.
Veröffentlicht: (2025)
von: Wang, Lehan, et al.
Veröffentlicht: (2025)
Hallucination Augmented Contrastive Learning for Multimodal Large Language Model
von: Jiang, Chaoya, et al.
Veröffentlicht: (2023)
von: Jiang, Chaoya, et al.
Veröffentlicht: (2023)
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
MMaDA: Multimodal Large Diffusion Language Models
von: Yang, Ling, et al.
Veröffentlicht: (2025)
von: Yang, Ling, et al.
Veröffentlicht: (2025)
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
von: Liu, Daixian, et al.
Veröffentlicht: (2026)
Graph-based Unsupervised Disentangled Representation Learning via Multimodal Large Language Models
von: Xie, Baao, et al.
Veröffentlicht: (2024)
von: Xie, Baao, et al.
Veröffentlicht: (2024)
GSVA: Generalized Segmentation via Multimodal Large Language Models
von: Xia, Zhuofan, et al.
Veröffentlicht: (2023)
von: Xia, Zhuofan, et al.
Veröffentlicht: (2023)
LENS: Multi-level Evaluation of Multimodal Reasoning with Large Language Models
von: Yao, Ruilin, et al.
Veröffentlicht: (2025)
von: Yao, Ruilin, et al.
Veröffentlicht: (2025)
Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback
von: Li, Zongjian, et al.
Veröffentlicht: (2025)
von: Li, Zongjian, et al.
Veröffentlicht: (2025)
Dynamic Multimodal Evaluation with Flexible Complexity by Vision-Language Bootstrapping
von: Yang, Yue, et al.
Veröffentlicht: (2024)
von: Yang, Yue, et al.
Veröffentlicht: (2024)
Test-Time Computing for Referring Multimodal Large Language Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2026)
von: Wu, Mingrui, et al.
Veröffentlicht: (2026)
Large Language Models for Multimodal Deformable Image Registration
von: Ma, Mingrui, et al.
Veröffentlicht: (2024)
von: Ma, Mingrui, et al.
Veröffentlicht: (2024)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
A Refer-and-Ground Multimodal Large Language Model for Biomedicine
von: Huang, Xiaoshuang, et al.
Veröffentlicht: (2024)
von: Huang, Xiaoshuang, et al.
Veröffentlicht: (2024)
ControlMLLM: Training-Free Visual Prompt Learning for Multimodal Large Language Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
ConvLLaVA: Hierarchical Backbones as Visual Encoder for Large Multimodal Models
von: Ge, Chunjiang, et al.
Veröffentlicht: (2024)
von: Ge, Chunjiang, et al.
Veröffentlicht: (2024)
Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025) -
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
von: Lin, Bin, et al.
Veröffentlicht: (2024) -
AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
von: Bao, Han, et al.
Veröffentlicht: (2024) -
Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
von: Lin, Bin, et al.
Veröffentlicht: (2023) -
PiCO: Peer Review in LLMs based on the Consistency Optimization
von: Ning, Kun-Peng, et al.
Veröffentlicht: (2024)