Feature CAM: Interpretable AI in Image Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Clement, Frincy, Yang, Ji, Cheng, Irene |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Harnessing Self-Supervised Features for Art Classification
by: Melis, Federico, et al.
Published: (2026)
by: Melis, Federico, et al.
Published: (2026)
Robust Latent Representation Tuning for Image-text Classification
by: Sun, Hao, et al.
Published: (2024)
by: Sun, Hao, et al.
Published: (2024)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
by: Huang, Victor Shea-Jay, et al.
Published: (2025)
by: Huang, Victor Shea-Jay, et al.
Published: (2025)
PFB-Diff: Progressive Feature Blending Diffusion for Text-driven Image Editing
by: Huang, Wenjing, et al.
Published: (2023)
by: Huang, Wenjing, et al.
Published: (2023)
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
by: Ouyang, Shuyi, et al.
Published: (2024)
by: Ouyang, Shuyi, et al.
Published: (2024)
JPEG AI Image Compression Visual Artifacts: Detection Methods and Dataset
by: Tsereh, Daria, et al.
Published: (2024)
by: Tsereh, Daria, et al.
Published: (2024)
ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images
by: Li, Xinyue, et al.
Published: (2026)
by: Li, Xinyue, et al.
Published: (2026)
D-Judge: How Far Are We? Assessing the Discrepancies Between AI-synthesized and Natural Images through Multimodal Guidance
by: Liu, Renyang, et al.
Published: (2024)
by: Liu, Renyang, et al.
Published: (2024)
BRITE: A Benchmark for Reliable and Interpretable T2V Evaluation on Implausible Scenarios
by: Tilak, Advait, et al.
Published: (2026)
by: Tilak, Advait, et al.
Published: (2026)
WILD: a new in-the-Wild Image Linkage Dataset for synthetic image attribution
by: Bongini, Pietro, et al.
Published: (2025)
by: Bongini, Pietro, et al.
Published: (2025)
LLM-based Fusion of Multi-modal Features for Commercial Memorability Prediction
by: Pramov, Aleksandar
Published: (2025)
by: Pramov, Aleksandar
Published: (2025)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
by: Cao, Pu, et al.
Published: (2023)
by: Cao, Pu, et al.
Published: (2023)
MRD: Multi-resolution Retrieval-Detection Fusion for High-Resolution Image Understanding
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents
by: Yi, Hongzhu, et al.
Published: (2026)
by: Yi, Hongzhu, et al.
Published: (2026)
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
by: Wu, Xiongwei, et al.
Published: (2024)
by: Wu, Xiongwei, et al.
Published: (2024)
Grounding is All You Need? Dual Temporal Grounding for Video Dialog
by: Qin, You, et al.
Published: (2024)
by: Qin, You, et al.
Published: (2024)
SAFIRE: Segment Any Forged Image Region
by: Kwon, Myung-Joon, et al.
Published: (2024)
by: Kwon, Myung-Joon, et al.
Published: (2024)
Localization of Synthetic Manipulations in Western Blot Images
by: Manjunath, Anmol, et al.
Published: (2024)
by: Manjunath, Anmol, et al.
Published: (2024)
ENCLIP: Ensembling and Clustering-Based Contrastive Language-Image Pretraining for Fashion Multimodal Search with Limited Data and Low-Quality Images
by: Naik, Prithviraj Purushottam, et al.
Published: (2024)
by: Naik, Prithviraj Purushottam, et al.
Published: (2024)
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
by: Zhou, Pengyuan, et al.
Published: (2024)
by: Zhou, Pengyuan, et al.
Published: (2024)
FakeParts: a New Family of AI-Generated DeepFakes
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
A Rate-Distortion-Classification Approach for Lossy Image Compression
by: Zhang, Yuefeng
Published: (2024)
by: Zhang, Yuefeng
Published: (2024)
Image Conductor: Precision Control for Interactive Video Synthesis
by: Li, Yaowei, et al.
Published: (2024)
by: Li, Yaowei, et al.
Published: (2024)
Moiré Video Authentication: A Physical Signature Against AI Video Generation
by: Qing, Yuan, et al.
Published: (2026)
by: Qing, Yuan, et al.
Published: (2026)
Kandinsky 3: Text-to-Image Synthesis for Multifunctional Generative Framework
by: Arkhipkin, Vladimir, et al.
Published: (2024)
by: Arkhipkin, Vladimir, et al.
Published: (2024)
End-to-End Optimized Image Compression with the Frequency-Oriented Transform
by: Zhang, Yuefeng, et al.
Published: (2024)
by: Zhang, Yuefeng, et al.
Published: (2024)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
by: Li, Yongqi, et al.
Published: (2024)
by: Li, Yongqi, et al.
Published: (2024)
Cross Modification Attention Based Deliberation Model for Image Captioning
by: Lian, Zheng, et al.
Published: (2021)
by: Lian, Zheng, et al.
Published: (2021)
InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing
by: Lin, Kun-Hsiang, et al.
Published: (2025)
by: Lin, Kun-Hsiang, et al.
Published: (2025)
VidCtx: Context-aware Video Question Answering with Image Models
by: Goulas, Andreas, et al.
Published: (2024)
by: Goulas, Andreas, et al.
Published: (2024)
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
by: Mandelli, Sara, et al.
Published: (2024)
by: Mandelli, Sara, et al.
Published: (2024)
Concept Conductor: Orchestrating Multiple Personalized Concepts in Text-to-Image Synthesis
by: Yao, Zebin, et al.
Published: (2024)
by: Yao, Zebin, et al.
Published: (2024)
Diversify, Contextualize, and Adapt: Efficient Entropy Modeling for Neural Image Codec
by: Kim, Jun-Hyuk, et al.
Published: (2024)
by: Kim, Jun-Hyuk, et al.
Published: (2024)
Parents and Children: Distinguishing Multimodal DeepFakes from Natural Images
by: Amoroso, Roberto, et al.
Published: (2023)
by: Amoroso, Roberto, et al.
Published: (2023)
BlobCtrl: Taming Controllable Blob for Element-level Image Editing
by: Li, Yaowei, et al.
Published: (2025)
by: Li, Yaowei, et al.
Published: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
by: Qu, Qiang, et al.
Published: (2025)
by: Qu, Qiang, et al.
Published: (2025)
Composing Concepts from Images and Videos via Concept-prompt Binding
by: Kong, Xianghao, et al.
Published: (2025)
by: Kong, Xianghao, et al.
Published: (2025)
Rethink Predicting the Optical Flow with the Kinetics Perspective
by: Cheng, Yuhao, et al.
Published: (2024)
by: Cheng, Yuhao, et al.
Published: (2024)
Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective
by: Mao, Yuxin, et al.
Published: (2025)
by: Mao, Yuxin, et al.
Published: (2025)
Omni-Dish: Photorealistic and Faithful Image Generation and Editing for Arbitrary Chinese Dishes
by: Liu, Huijie, et al.
Published: (2025)
by: Liu, Huijie, et al.
Published: (2025)
Similar Items
-
Harnessing Self-Supervised Features for Art Classification
by: Melis, Federico, et al.
Published: (2026) -
Robust Latent Representation Tuning for Image-text Classification
by: Sun, Hao, et al.
Published: (2024) -
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
by: Huang, Victor Shea-Jay, et al.
Published: (2025) -
PFB-Diff: Progressive Feature Blending Diffusion for Text-driven Image Editing
by: Huang, Wenjing, et al.
Published: (2023) -
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
by: Ouyang, Shuyi, et al.
Published: (2024)