Saved in:
| Main Authors: | Cedro, Mateusz, Chlebus, Marcin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.10142 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond the Black Box: Do More Complex Deep Learning Models Provide Superior XAI Explanations?
by: Cedro, Mateusz, et al.
Published: (2024)
by: Cedro, Mateusz, et al.
Published: (2024)
Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
by: Ding, Sihao, et al.
Published: (2025)
by: Ding, Sihao, et al.
Published: (2025)
Local-to-Global Logical Explanations for Deep Vision Models
by: Vasu, Bhavan, et al.
Published: (2026)
by: Vasu, Bhavan, et al.
Published: (2026)
Adapting Vision-Language Models for E-commerce Understanding at Scale
by: Nulli, Matteo, et al.
Published: (2026)
by: Nulli, Matteo, et al.
Published: (2026)
Quantifying Explanation Consistency: The C-Score Metric for CAM-Based Explainability in Medical Image Classification
by: Elangovan, Kabilan, et al.
Published: (2026)
by: Elangovan, Kabilan, et al.
Published: (2026)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
by: Zhao, Ziwei, et al.
Published: (2024)
by: Zhao, Ziwei, et al.
Published: (2024)
MambaLoc: Efficient Camera Localisation via State Space Model
by: Wang, Jialu, et al.
Published: (2024)
by: Wang, Jialu, et al.
Published: (2024)
Towards Self-Refinement of Vision-Language Models with Triangular Consistency
by: Deng, Yunlong, et al.
Published: (2025)
by: Deng, Yunlong, et al.
Published: (2025)
Vision-Language Consistency Guided Multi-modal Prompt Learning for Blind AI Generated Image Quality Assessment
by: Fu, Jun, et al.
Published: (2024)
by: Fu, Jun, et al.
Published: (2024)
Improving Deep Learning-based Automatic Cranial Defect Reconstruction by Heavy Data Augmentation: From Image Registration to Latent Diffusion Models
by: Wodzinski, Marek, et al.
Published: (2024)
by: Wodzinski, Marek, et al.
Published: (2024)
On the Importance and Evaluation of Narrativity in Natural Language AI Explanations
by: Cedro, Mateusz, et al.
Published: (2026)
by: Cedro, Mateusz, et al.
Published: (2026)
PEnG: Pose-Enhanced Geo-Localisation
by: Shore, Tavis, et al.
Published: (2024)
by: Shore, Tavis, et al.
Published: (2024)
Improving Prototypical Parts Abstraction for Case-Based Reasoning Explanations Designed for the Kidney Stone Type Recognition
by: Flores-Araiza, Daniel, et al.
Published: (2024)
by: Flores-Araiza, Daniel, et al.
Published: (2024)
Improving Medical Diagnostics with Vision-Language Models: Convex Hull-Based Uncertainty Analysis
by: Catak, Ferhat Ozgur, et al.
Published: (2024)
by: Catak, Ferhat Ozgur, et al.
Published: (2024)
Vision Inference Former: Sustaining Visual Consistency in Multimodal Large Language Models
by: Dong, Xinpeng, et al.
Published: (2026)
by: Dong, Xinpeng, et al.
Published: (2026)
The Solution for Temporal Action Localisation Task of Perception Test Challenge 2024
by: Han, Yinan, et al.
Published: (2024)
by: Han, Yinan, et al.
Published: (2024)
Xray-Visual Models: Scaling Vision models on Industry Scale Data
by: Mishra, Shlok, et al.
Published: (2026)
by: Mishra, Shlok, et al.
Published: (2026)
LINE: LLM-based Iterative Neuron Explanations for Vision Models
by: Zaigrajew, Vladimir, et al.
Published: (2026)
by: Zaigrajew, Vladimir, et al.
Published: (2026)
Consist-Retinex: One-Step Noise-Emphasized Consistency Training Accelerates High-Quality Retinex Enhancement
by: Xu, Jian, et al.
Published: (2025)
by: Xu, Jian, et al.
Published: (2025)
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency
by: Liu, Junming, et al.
Published: (2026)
by: Liu, Junming, et al.
Published: (2026)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
by: Leem, Saebom, et al.
Published: (2024)
by: Leem, Saebom, et al.
Published: (2024)
Musketeer: Joint Training for Multi-task Vision Language Model with Task Explanation Prompts
by: Zhang, Zhaoyang, et al.
Published: (2023)
by: Zhang, Zhaoyang, et al.
Published: (2023)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
by: Pach, Mateusz, et al.
Published: (2025)
by: Pach, Mateusz, et al.
Published: (2025)
ALAM: Algebraically Consistent Latent Action Model for Vision-Language-Action Models
by: Tang, Zuojin, et al.
Published: (2026)
by: Tang, Zuojin, et al.
Published: (2026)
ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models
by: Zhang, Jieyu, et al.
Published: (2024)
by: Zhang, Jieyu, et al.
Published: (2024)
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
by: Chen, Zhihao, et al.
Published: (2023)
by: Chen, Zhihao, et al.
Published: (2023)
From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models
by: Khokhar, Pir Bakhsh, et al.
Published: (2026)
by: Khokhar, Pir Bakhsh, et al.
Published: (2026)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
Improving Consistency Models with Generator-Augmented Flows
by: Issenhuth, Thibaut, et al.
Published: (2024)
by: Issenhuth, Thibaut, et al.
Published: (2024)
CTS: A Consistency-Based Medical Image Segmentation Model
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?
by: He, Jingtao, et al.
Published: (2026)
by: He, Jingtao, et al.
Published: (2026)
In-Context Learning Improves Compositional Understanding of Vision-Language Models
by: Nulli, Matteo, et al.
Published: (2024)
by: Nulli, Matteo, et al.
Published: (2024)
Margin and Consistency Supervision for Calibrated and Robust Vision Models
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
Generating Part-Based Global Explanations Via Correspondence
by: Rathore, Kunal, et al.
Published: (2025)
by: Rathore, Kunal, et al.
Published: (2025)
A Circular Argument : Does RoPE need to be Equivariant for Vision?
by: van de Geijn, Chase, et al.
Published: (2025)
by: van de Geijn, Chase, et al.
Published: (2025)
Leveraging Local Structure for Improving Model Explanations: An Information Propagation Approach
by: Yang, Ruo, et al.
Published: (2024)
by: Yang, Ruo, et al.
Published: (2024)
Self-Consistency as a Free Lunch: Reducing Hallucinations in Vision-Language Models via Self-Reflection
by: Han, Mingfei, et al.
Published: (2025)
by: Han, Mingfei, et al.
Published: (2025)
Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Models
by: Yan, Bei, et al.
Published: (2024)
by: Yan, Bei, et al.
Published: (2024)
Vision Bridge Transformer at Scale
by: Tan, Zhenxiong, et al.
Published: (2025)
by: Tan, Zhenxiong, et al.
Published: (2025)
Similar Items
-
Beyond the Black Box: Do More Complex Deep Learning Models Provide Superior XAI Explanations?
by: Cedro, Mateusz, et al.
Published: (2024) -
Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
by: Ding, Sihao, et al.
Published: (2025) -
Local-to-Global Logical Explanations for Deep Vision Models
by: Vasu, Bhavan, et al.
Published: (2026) -
Adapting Vision-Language Models for E-commerce Understanding at Scale
by: Nulli, Matteo, et al.
Published: (2026) -
Quantifying Explanation Consistency: The C-Score Metric for CAM-Based Explainability in Medical Image Classification
by: Elangovan, Kabilan, et al.
Published: (2026)