Local-to-Global Logical Explanations for Deep Vision Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vasu, Bhavan, Raffa, Giuseppe, Tadepalli, Prasad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generating Part-Based Global Explanations Via Correspondence
von: Rathore, Kunal, et al.
Veröffentlicht: (2025)
von: Rathore, Kunal, et al.
Veröffentlicht: (2025)
Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
von: Ding, Sihao, et al.
Veröffentlicht: (2025)
von: Ding, Sihao, et al.
Veröffentlicht: (2025)
LogicQA: Logical Anomaly Detection with Vision Language Model Generated Questions
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
von: Kwon, Yejin, et al.
Veröffentlicht: (2025)
KPCA-CAM: Visual Explainability of Deep Computer Vision Models using Kernel PCA
von: Karmani, Sachin, et al.
Veröffentlicht: (2024)
von: Karmani, Sachin, et al.
Veröffentlicht: (2024)
Diffexplainer: Towards Cross-modal Global Explanations with Diffusion Models
von: Pennisi, Matteo, et al.
Veröffentlicht: (2024)
von: Pennisi, Matteo, et al.
Veröffentlicht: (2024)
From Local Cues to Global Percepts: Emergent Gestalt Organization in Self-Supervised Vision Models
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
Scaling Vision Models Does Not Consistently Improve Localisation-Based Explanation Quality
von: Cedro, Mateusz, et al.
Veröffentlicht: (2026)
von: Cedro, Mateusz, et al.
Veröffentlicht: (2026)
DLaVA: Document Language and Vision Assistant for Answer Localization with Enhanced Interpretability and Trustworthiness
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
von: Mohammadshirazi, Ahmad, et al.
Veröffentlicht: (2024)
Weakly Supervised Pneumonia Localization from Chest X-Rays Using Deep Neural Network and Grad-CAM Explanations
von: Shahi, Kiran, et al.
Veröffentlicht: (2025)
von: Shahi, Kiran, et al.
Veröffentlicht: (2025)
SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal Model
von: Huang, Zhenglin, et al.
Veröffentlicht: (2024)
von: Huang, Zhenglin, et al.
Veröffentlicht: (2024)
VALUED -- Vision and Logical Understanding Evaluation Dataset
von: Saha, Soumadeep, et al.
Veröffentlicht: (2023)
von: Saha, Soumadeep, et al.
Veröffentlicht: (2023)
Understanding the (Extra-)Ordinary: Validating Deep Model Decisions with Prototypical Concept-based Explanations
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2023)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2023)
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
von: An, Wenbin, et al.
Veröffentlicht: (2024)
von: An, Wenbin, et al.
Veröffentlicht: (2024)
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency
von: Liu, Junming, et al.
Veröffentlicht: (2026)
von: Liu, Junming, et al.
Veröffentlicht: (2026)
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
von: Kaduri, Omri, et al.
Veröffentlicht: (2024)
von: Kaduri, Omri, et al.
Veröffentlicht: (2024)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
Landmark-based Localization using Stereo Vision and Deep Learning in GPS-Denied Battlefield Environment
von: Sapkota, Ganesh, et al.
Veröffentlicht: (2024)
von: Sapkota, Ganesh, et al.
Veröffentlicht: (2024)
From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models
von: Khokhar, Pir Bakhsh, et al.
Veröffentlicht: (2026)
von: Khokhar, Pir Bakhsh, et al.
Veröffentlicht: (2026)
LINE: LLM-based Iterative Neuron Explanations for Vision Models
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2026)
von: Zaigrajew, Vladimir, et al.
Veröffentlicht: (2026)
Understanding Trade offs When Conditioning Synthetic Data
von: Trabucco, Brandon, et al.
Veröffentlicht: (2025)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2025)
Musketeer: Joint Training for Multi-task Vision Language Model with Task Explanation Prompts
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2023)
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2023)
Composed Vision-Language Retrieval for Skin Cancer Case Search via Joint Alignment of Global and Local Representations
von: Wang, Yuheng, et al.
Veröffentlicht: (2026)
von: Wang, Yuheng, et al.
Veröffentlicht: (2026)
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
von: Liu, Zhou, et al.
Veröffentlicht: (2026)
BiPVL-Seg: Bidirectional Progressive Vision-Language Fusion with Global-Local Alignment for Medical Image Segmentation
von: Sultan, Rafi Ibn, et al.
Veröffentlicht: (2025)
von: Sultan, Rafi Ibn, et al.
Veröffentlicht: (2025)
Vision-DeepResearch: Incentivizing DeepResearch Capability in Multimodal Large Language Models
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2026)
Neural Residual Diffusion Models for Deep Scalable Vision Generation
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2024)
FastVLM: Efficient Vision Encoding for Vision Language Models
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
Local vs. Global: Local Land-Use and Land-Cover Models Deliver Higher Quality Maps
von: Tadesse, Girmaw Abebe, et al.
Veröffentlicht: (2024)
von: Tadesse, Girmaw Abebe, et al.
Veröffentlicht: (2024)
Global Geometry Is Not Enough for Vision Representations
von: Chung, Jiwan, et al.
Veröffentlicht: (2026)
von: Chung, Jiwan, et al.
Veröffentlicht: (2026)
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
von: Zhao, Chenchen, et al.
Veröffentlicht: (2026)
von: Zhao, Chenchen, et al.
Veröffentlicht: (2026)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
von: Wang, Hengyi, et al.
Veröffentlicht: (2024)
Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Efficient Robotic Manipulation
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
von: Li, Zaijing, et al.
Veröffentlicht: (2026)
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
von: Zeng, Haoxi, et al.
Veröffentlicht: (2025)
von: Zeng, Haoxi, et al.
Veröffentlicht: (2025)
Leveraging Local Structure for Improving Model Explanations: An Information Propagation Approach
von: Yang, Ruo, et al.
Veröffentlicht: (2024)
von: Yang, Ruo, et al.
Veröffentlicht: (2024)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
von: Zhao, Ziwei, et al.
Veröffentlicht: (2024)
von: Zhao, Ziwei, et al.
Veröffentlicht: (2024)
Locatability-Guided Adaptive Reasoning for Image Geo-Localization with Vision-Language Models
von: Yu, Bo, et al.
Veröffentlicht: (2026)
von: Yu, Bo, et al.
Veröffentlicht: (2026)
Balanced Token Pruning: Accelerating Vision Language Models Beyond Local Optimization
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
Non-Contrastive Vision-Language Learning with Predictive Embedding Alignment
von: Kuhn, Lukas, et al.
Veröffentlicht: (2026)
von: Kuhn, Lukas, et al.
Veröffentlicht: (2026)
Fuzzy-Logic and Deep Learning for Environmental Condition-Aware Road Surface Classification
von: Demetgul, Mustafa, et al.
Veröffentlicht: (2025)
von: Demetgul, Mustafa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Generating Part-Based Global Explanations Via Correspondence
von: Rathore, Kunal, et al.
Veröffentlicht: (2025) -
Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
von: Ding, Sihao, et al.
Veröffentlicht: (2025) -
LogicQA: Logical Anomaly Detection with Vision Language Model Generated Questions
von: Kwon, Yejin, et al.
Veröffentlicht: (2025) -
KPCA-CAM: Visual Explainability of Deep Computer Vision Models using Kernel PCA
von: Karmani, Sachin, et al.
Veröffentlicht: (2024) -
Diffexplainer: Towards Cross-modal Global Explanations with Diffusion Models
von: Pennisi, Matteo, et al.
Veröffentlicht: (2024)