Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yifei, Song, James, Gu, Siyi, Jiang, Tianxu, Pan, Bo, Bai, Guangji, Zhao, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MEGL: Multimodal Explanation-Guided Learning
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
Visual Attention Prompted Prediction and Learning
von: Zhang, Yifei, et al.
Veröffentlicht: (2023)
von: Zhang, Yifei, et al.
Veröffentlicht: (2023)
DUE: Dynamic Uncertainty-Aware Explanation Supervision via 3D Imputation
von: Zhao, Qilong, et al.
Veröffentlicht: (2024)
von: Zhao, Qilong, et al.
Veröffentlicht: (2024)
Graphical Perception of Saliency-based Model Explanations
von: Zhao, Yayan, et al.
Veröffentlicht: (2024)
von: Zhao, Yayan, et al.
Veröffentlicht: (2024)
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
Towards A Comprehensive Visual Saliency Explanation Framework for AI-based Face Recognition Systems
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
ICE-Bench: A Unified and Comprehensive Benchmark for Image Creating and Editing
von: Pan, Yulin, et al.
Veröffentlicht: (2025)
von: Pan, Yulin, et al.
Veröffentlicht: (2025)
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
von: Zhu, Mengdan, et al.
Veröffentlicht: (2025)
von: Zhu, Mengdan, et al.
Veröffentlicht: (2025)
OCRGenBench: A Comprehensive Benchmark for Evaluating OCR Generative Capabilities
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
von: Xu, Mingjie, et al.
Veröffentlicht: (2025)
EventBench: Towards Comprehensive Benchmarking of Event-based MLLMs
von: Liu, Shaoyu, et al.
Veröffentlicht: (2025)
von: Liu, Shaoyu, et al.
Veröffentlicht: (2025)
LongInsightBench: A Comprehensive Benchmark for Evaluating Omni-Modal Models on Human-Centric Long-Video Understanding
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
IllusionBench+: A Large-scale and Comprehensive Benchmark for Visual Illusion Understanding in Vision-Language Models
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation
von: Chen, Yubin, et al.
Veröffentlicht: (2025)
von: Chen, Yubin, et al.
Veröffentlicht: (2025)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
von: Zhang, Tao, et al.
Veröffentlicht: (2025)
TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
XBench: A Comprehensive Benchmark for Visual-Language Explanations in Chest Radiography
von: Luo, Haozhe, et al.
Veröffentlicht: (2025)
von: Luo, Haozhe, et al.
Veröffentlicht: (2025)
VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On
von: Liang, Xiaoye, et al.
Veröffentlicht: (2026)
von: Liang, Xiaoye, et al.
Veröffentlicht: (2026)
SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension
von: Li, Bohao, et al.
Veröffentlicht: (2024)
von: Li, Bohao, et al.
Veröffentlicht: (2024)
Relevance-guided Audio Visual Fusion for Video Saliency Prediction
von: Yu, Li, et al.
Veröffentlicht: (2024)
von: Yu, Li, et al.
Veröffentlicht: (2024)
VERIFY: A Benchmark of Visual Explanation and Reasoning for Investigating Multimodal Reasoning Fidelity
von: Bi, Jing, et al.
Veröffentlicht: (2025)
von: Bi, Jing, et al.
Veröffentlicht: (2025)
TokBench: Evaluating Your Visual Tokenizer before Visual Generation
von: Wu, Junfeng, et al.
Veröffentlicht: (2025)
von: Wu, Junfeng, et al.
Veröffentlicht: (2025)
GeoR-Bench: Evaluating Geoscience Visual Reasoning
von: Zheng, Yushuo, et al.
Veröffentlicht: (2026)
von: Zheng, Yushuo, et al.
Veröffentlicht: (2026)
FinChart-Bench: Benchmarking Financial Chart Comprehension in Vision-Language Models
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
On the Evaluation Consistency of Attribution-based Explanations
von: Duan, Jiarui, et al.
Veröffentlicht: (2024)
von: Duan, Jiarui, et al.
Veröffentlicht: (2024)
FineState-Bench: A Comprehensive Benchmark for Fine-Grained State Control in GUI Agents
von: Ji, Fengxian, et al.
Veröffentlicht: (2025)
von: Ji, Fengxian, et al.
Veröffentlicht: (2025)
UrBench: A Comprehensive Benchmark for Evaluating Large Multimodal Models in Multi-View Urban Scenarios
von: Zhou, Baichuan, et al.
Veröffentlicht: (2024)
von: Zhou, Baichuan, et al.
Veröffentlicht: (2024)
VST++: Efficient and Stronger Visual Saliency Transformer
von: Liu, Nian, et al.
Veröffentlicht: (2023)
von: Liu, Nian, et al.
Veröffentlicht: (2023)
VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2026)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2026)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
von: Ying, Kaining, et al.
Veröffentlicht: (2024)
ConvBench: A Comprehensive Benchmark for 2D Convolution Primitive Evaluation
von: Alvarenga, Lucas, et al.
Veröffentlicht: (2024)
von: Alvarenga, Lucas, et al.
Veröffentlicht: (2024)
PAI-Bench: A Comprehensive Benchmark For Physical AI
von: Zhou, Fengzhe, et al.
Veröffentlicht: (2025)
von: Zhou, Fengzhe, et al.
Veröffentlicht: (2025)
PhenoBench: A Comprehensive Benchmark for Cell Phenotyping
von: Winklmayr, Claudia, et al.
Veröffentlicht: (2025)
von: Winklmayr, Claudia, et al.
Veröffentlicht: (2025)
MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
von: Ma, Yubo, et al.
Veröffentlicht: (2024)
Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments
von: Ali, Muhammad, et al.
Veröffentlicht: (2025)
von: Ali, Muhammad, et al.
Veröffentlicht: (2025)
M-ErasureBench: A Comprehensive Multimodal Evaluation Benchmark for Concept Erasure in Diffusion Models
von: Weng, Ju-Hsuan, et al.
Veröffentlicht: (2025)
von: Weng, Ju-Hsuan, et al.
Veröffentlicht: (2025)
TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos
von: Feng, Hengyi, et al.
Veröffentlicht: (2026)
von: Feng, Hengyi, et al.
Veröffentlicht: (2026)
GT23D-Bench: A Comprehensive General Text-to-3D Generation Benchmark
von: Cai, Xiao, et al.
Veröffentlicht: (2024)
von: Cai, Xiao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MEGL: Multimodal Explanation-Guided Learning
von: Zhang, Yifei, et al.
Veröffentlicht: (2024) -
Visual Attention Prompted Prediction and Learning
von: Zhang, Yifei, et al.
Veröffentlicht: (2023) -
DUE: Dynamic Uncertainty-Aware Explanation Supervision via 3D Imputation
von: Zhao, Qilong, et al.
Veröffentlicht: (2024) -
Graphical Perception of Saliency-based Model Explanations
von: Zhao, Yayan, et al.
Veröffentlicht: (2024) -
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)