Explaining Human Comparisons using Alignment-Importance Heatmaps
Fuente:
arXiv
Saved in:
| Main Authors: | Truong, Nhut, Pesenti, Dario, Hasson, Uri |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Heatmap Regression without Soft-Argmax for Facial Landmark Detection
by: Yang, Chiao-An, et al.
Published: (2025)
by: Yang, Chiao-An, et al.
Published: (2025)
Part-based Quantitative Analysis for Heatmaps
by: Tursun, Osman, et al.
Published: (2024)
by: Tursun, Osman, et al.
Published: (2024)
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
by: Maldonado, Gabriel, et al.
Published: (2025)
by: Maldonado, Gabriel, et al.
Published: (2025)
Heatmap Guided Query Transformers for Robust Astrocyte Detection across Immunostains and Resolutions
by: Zhang, Xizhe, et al.
Published: (2025)
by: Zhang, Xizhe, et al.
Published: (2025)
Efficient Heatmap-Guided 6-Dof Grasp Detection in Cluttered Scenes
by: Chen, Siang, et al.
Published: (2024)
by: Chen, Siang, et al.
Published: (2024)
Language Models Can Explain Visual Features via Steering
by: Ferrando, Javier, et al.
Published: (2026)
by: Ferrando, Javier, et al.
Published: (2026)
Score-Regularized Joint Sampling with Importance Weights for Flow Matching
by: Liu, Xinshuang, et al.
Published: (2025)
by: Liu, Xinshuang, et al.
Published: (2025)
Selective Social-Interaction via Individual Importance for Fast Human Trajectory Prediction
by: Urano, Yota, et al.
Published: (2025)
by: Urano, Yota, et al.
Published: (2025)
Think First, Assign Next (ThiFAN-VQA): A Two-stage Chain-of-Thought Framework for Post-Disaster Damage Assessment
by: Karimi, Ehsan, et al.
Published: (2025)
by: Karimi, Ehsan, et al.
Published: (2025)
Implicit Preference Alignment for Human Image Animation
by: Wang, Yuanzhi, et al.
Published: (2026)
by: Wang, Yuanzhi, et al.
Published: (2026)
Not Too Generative, Not Too Discriminative: The Human Alignment Sweet Spot
by: Ortega, Jorge Chang, et al.
Published: (2026)
by: Ortega, Jorge Chang, et al.
Published: (2026)
Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition
by: Luo, Róisín, et al.
Published: (2024)
by: Luo, Róisín, et al.
Published: (2024)
XEdgeAI: A Human-centered Industrial Inspection Framework with Data-centric Explainable Edge AI Approach
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
RT-DETRv2 Explained in 8 Illustrations
by: Chua, Ethan Qi Yang, et al.
Published: (2025)
by: Chua, Ethan Qi Yang, et al.
Published: (2025)
Beyond Human-prompting: Adaptive Prompt Tuning with Semantic Alignment for Anomaly Detection
by: Chen, Pi-Wei, et al.
Published: (2025)
by: Chen, Pi-Wei, et al.
Published: (2025)
Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers
by: Knights, Ethan
Published: (2026)
by: Knights, Ethan
Published: (2026)
A Novel Framework for Automated Explain Vision Model Using Vision-Language Models
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
by: Nguyen, Phu-Vinh, et al.
Published: (2025)
IE-Bench: Advancing the Measurement of Text-Driven Image Editing for Human Perception Alignment
by: Sun, Shangkun, et al.
Published: (2025)
by: Sun, Shangkun, et al.
Published: (2025)
On Explaining Visual Captioning with Hybrid Markov Logic Networks
by: Shah, Monika, et al.
Published: (2025)
by: Shah, Monika, et al.
Published: (2025)
Explaining 3D Computed Tomography Classifiers with Counterfactuals
by: Cohen, Joseph Paul, et al.
Published: (2025)
by: Cohen, Joseph Paul, et al.
Published: (2025)
Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments
by: Depaolini, Icaro Re, et al.
Published: (2026)
by: Depaolini, Icaro Re, et al.
Published: (2026)
Med-CAM: Minimal Evidence for Explaining Medical Decision Making
by: Suhail, Pirzada, et al.
Published: (2026)
by: Suhail, Pirzada, et al.
Published: (2026)
VerLM: Explaining Face Verification Using Natural Language
by: Hannan, Syed Abdul, et al.
Published: (2026)
by: Hannan, Syed Abdul, et al.
Published: (2026)
Sparks of Explainability: Recent Advancements in Explaining Large Vision Models
by: Fel, Thomas
Published: (2025)
by: Fel, Thomas
Published: (2025)
P-TAME: Explain Any Image Classifier with Trained Perturbations
by: Ntrougkas, Mariano V., et al.
Published: (2025)
by: Ntrougkas, Mariano V., et al.
Published: (2025)
Explaining multimodal LLMs via intra-modal token interactions
by: Liang, Jiawei, et al.
Published: (2025)
by: Liang, Jiawei, et al.
Published: (2025)
Explaining How Visual, Textual and Multimodal Encoders Share Concepts
by: Cornet, Clément, et al.
Published: (2025)
by: Cornet, Clément, et al.
Published: (2025)
Explaining raw data complexity to improve satellite onboard processing
by: Dorise, Adrien, et al.
Published: (2025)
by: Dorise, Adrien, et al.
Published: (2025)
Trends, Applications, and Challenges in Human Attention Modelling
by: Cartella, Giuseppe, et al.
Published: (2024)
by: Cartella, Giuseppe, et al.
Published: (2024)
Policy Optimized Text-to-Image Pipeline Design
by: Gadot, Uri, et al.
Published: (2025)
by: Gadot, Uri, et al.
Published: (2025)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
by: Truong, Khang, et al.
Published: (2025)
by: Truong, Khang, et al.
Published: (2025)
Explaining Multi-modal Large Language Models by Analyzing their Vision Perception
by: Giulivi, Loris, et al.
Published: (2024)
by: Giulivi, Loris, et al.
Published: (2024)
Improving the Explain-Any-Concept by Introducing Nonlinearity to the Trainable Surrogate Model
by: Zaval, Mounes, et al.
Published: (2024)
by: Zaval, Mounes, et al.
Published: (2024)
Explain Before You Answer: A Survey on Compositional Visual Reasoning
by: Ke, Fucai, et al.
Published: (2025)
by: Ke, Fucai, et al.
Published: (2025)
DEPICT: Diffusion-Enabled Permutation Importance for Image Classification Tasks
by: Jabbour, Sarah, et al.
Published: (2024)
by: Jabbour, Sarah, et al.
Published: (2024)
IDPruner: Harmonizing Importance and Diversity in Visual Token Pruning for MLLMs
by: Tan, Yifan, et al.
Published: (2026)
by: Tan, Yifan, et al.
Published: (2026)
Surveillance Video-Based Traffic Accident Detection Using Transformer Architecture
by: Singh, Tanu, et al.
Published: (2025)
by: Singh, Tanu, et al.
Published: (2025)
SegMaFormer: A Hybrid State-Space and Transformer Model for Efficient Segmentation
by: Nguyen, Duy D., et al.
Published: (2026)
by: Nguyen, Duy D., et al.
Published: (2026)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
FPANet: Frequency-based Video Demoireing using Frame-level Post Alignment
by: Oh, Gyeongrok, et al.
Published: (2023)
by: Oh, Gyeongrok, et al.
Published: (2023)
Similar Items
-
Heatmap Regression without Soft-Argmax for Facial Landmark Detection
by: Yang, Chiao-An, et al.
Published: (2025) -
Part-based Quantitative Analysis for Heatmaps
by: Tursun, Osman, et al.
Published: (2024) -
Adversarially-Refined VQ-GAN with Dense Motion Tokenization for Spatio-Temporal Heatmaps
by: Maldonado, Gabriel, et al.
Published: (2025) -
Heatmap Guided Query Transformers for Robust Astrocyte Detection across Immunostains and Resolutions
by: Zhang, Xizhe, et al.
Published: (2025) -
Efficient Heatmap-Guided 6-Dof Grasp Detection in Cluttered Scenes
by: Chen, Siang, et al.
Published: (2024)