CA-Stream: Attention-based pooling for interpretable image recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Torres, Felipe, Zhang, Hanwei, Sicre, Ronan, Ayache, Stéphane, Avrithis, Yannis |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Opti-CAM: Optimizing saliency maps for interpretability
by: Zhang, Hanwei, et al.
Published: (2023)
by: Zhang, Hanwei, et al.
Published: (2023)
A Learning Paradigm for Interpretable Gradients
by: Figueroa, Felipe Torres, et al.
Published: (2024)
by: Figueroa, Felipe Torres, et al.
Published: (2024)
DP-Net: Learning Discriminative Parts for image recognition
by: Sicre, Ronan, et al.
Published: (2024)
by: Sicre, Ronan, et al.
Published: (2024)
Multi-Target Unsupervised Domain Adaptation for Semantic Segmentation without External Data
by: Xu, Yonghao, et al.
Published: (2024)
by: Xu, Yonghao, et al.
Published: (2024)
Is ImageNet worth 1 video? Learning strong image encoders from 1 long unlabelled video
by: Venkataramanan, Shashanka, et al.
Published: (2023)
by: Venkataramanan, Shashanka, et al.
Published: (2023)
Attention, Please! Revisiting Attentive Probing Through the Lens of Efficiency
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
CAM-Based Methods Can See through Walls
by: Taimeskhanov, Magamed, et al.
Published: (2024)
by: Taimeskhanov, Magamed, et al.
Published: (2024)
Spatial regularisation for improved accuracy and interpretability in keypoint-based registration
by: Billot, Benjamin, et al.
Published: (2025)
by: Billot, Benjamin, et al.
Published: (2025)
Eidos: Efficient, Imperceptible Adversarial 3D Point Clouds
by: Zhang, Hanwei, et al.
Published: (2024)
by: Zhang, Hanwei, et al.
Published: (2024)
Composed Image Retrieval for Training-Free Domain Conversion
by: Efthymiadis, Nikos, et al.
Published: (2024)
by: Efthymiadis, Nikos, et al.
Published: (2024)
Composed Image Retrieval for Remote Sensing
by: Psomas, Bill, et al.
Published: (2024)
by: Psomas, Bill, et al.
Published: (2024)
Revisiting Transferable Adversarial Images: Systemization, Evaluation, and New Insights
by: Zhao, Zhengyu, et al.
Published: (2023)
by: Zhao, Zhengyu, et al.
Published: (2023)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
On Train-Test Class Overlap and Detection for Image Retrieval
by: Song, Chull Hwan, et al.
Published: (2024)
by: Song, Chull Hwan, et al.
Published: (2024)
CA-YOLO: Cross Attention Empowered YOLO for Biomimetic Localization
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Benchmarking Composed Image Retrieval for Applied Earth Observation
by: Psomas, Bill, et al.
Published: (2026)
by: Psomas, Bill, et al.
Published: (2026)
CarGait: Cross-Attention based Re-ranking for Gait recognition
by: Habib, Gavriel, et al.
Published: (2025)
by: Habib, Gavriel, et al.
Published: (2025)
EchoVQA: Enabling Conversational Assistance for Point-of-Care Cardiac Ultrasound
by: Bellos, Filippos, et al.
Published: (2026)
by: Bellos, Filippos, et al.
Published: (2026)
HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction
by: Cheng, Chong, et al.
Published: (2026)
by: Cheng, Chong, et al.
Published: (2026)
Dense outlier detection and open-set recognition based on training with noisy negative images
by: Bevandić, Petra, et al.
Published: (2021)
by: Bevandić, Petra, et al.
Published: (2021)
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
by: Liu, Wenjie, et al.
Published: (2026)
by: Liu, Wenjie, et al.
Published: (2026)
StochCA: A Novel Approach for Exploiting Pretrained Models with Cross-Attention
by: Seo, Seungwon, et al.
Published: (2024)
by: Seo, Seungwon, et al.
Published: (2024)
MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation
by: Li, Zhiqi, et al.
Published: (2025)
by: Li, Zhiqi, et al.
Published: (2025)
Generative Medical Image Anonymization Based on Latent Code Projection and Optimization
by: Li, Huiyu, et al.
Published: (2025)
by: Li, Huiyu, et al.
Published: (2025)
Research on gesture recognition method based on SEDCNN-SVM
by: Zhang, Mingjin, et al.
Published: (2024)
by: Zhang, Mingjin, et al.
Published: (2024)
Indoor scene recognition from images under visual corruptions
by: Costa, Willams de Lima, et al.
Published: (2024)
by: Costa, Willams de Lima, et al.
Published: (2024)
MoCA: Identity-Preserving Text-to-Video Generation via Mixture of Cross Attention
by: Xie, Qi, et al.
Published: (2025)
by: Xie, Qi, et al.
Published: (2025)
CA^2ST: Cross-Attention in Audio, Space, and Time for Holistic Video Recognition
by: Lee, Jongseo, et al.
Published: (2025)
by: Lee, Jongseo, et al.
Published: (2025)
Predictive Temporal Attention on Event-based Video Stream for Energy-efficient Situation Awareness
by: Bu, Yiming, et al.
Published: (2024)
by: Bu, Yiming, et al.
Published: (2024)
GATE-AD: Graph Attention Network Encoding For Few-Shot Industrial Visual Anomaly Detection
by: Psiris, Aggelos, et al.
Published: (2026)
by: Psiris, Aggelos, et al.
Published: (2026)
Revisiting Physically Realizable Adversarial Object Attack against LiDAR-based Detection: Clarifying Problem Formulation and Experimental Protocols
by: Cheng, Luo, et al.
Published: (2025)
by: Cheng, Luo, et al.
Published: (2025)
DeformStream: Deformation-based Adaptive Volumetric Video Streaming
by: Li, Boyan, et al.
Published: (2024)
by: Li, Boyan, et al.
Published: (2024)
Prior-based Objective Inference Mining Potential Uncertainty for Facial Expression Recognition
by: Liu, Hanwei, et al.
Published: (2024)
by: Liu, Hanwei, et al.
Published: (2024)
Y-CA-Net: A Convolutional Attention Based Network for Volumetric Medical Image Segmentation
by: Sharif, Muhammad Hamza, et al.
Published: (2024)
by: Sharif, Muhammad Hamza, et al.
Published: (2024)
CA-IDD: Cross-Attention Guided Identity-Conditional Diffusion for Identity-Consistent Face Swapping
by: Rana, Md Shohel, et al.
Published: (2026)
by: Rana, Md Shohel, et al.
Published: (2026)
Evaluating the plausibility of synthetic images for improving automated endoscopic stone recognition
by: Gonzalez-Perez, Ruben, et al.
Published: (2024)
by: Gonzalez-Perez, Ruben, et al.
Published: (2024)
Intelligent recognition of GPR road hidden defect images based on feature fusion and attention mechanism
by: Lv, Haotian, et al.
Published: (2025)
by: Lv, Haotian, et al.
Published: (2025)
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge
by: Lagani, Gabriele, et al.
Published: (2025)
by: Lagani, Gabriele, et al.
Published: (2025)
BCFPL: Binary classification ConvNet based Fast Parking space recognition with Low resolution image
by: Zhang, Shuo, et al.
Published: (2024)
by: Zhang, Shuo, et al.
Published: (2024)
Disability Representations: Finding Biases in Automatic Image Generation
by: Tevissen, Yannis
Published: (2024)
by: Tevissen, Yannis
Published: (2024)
Similar Items
-
Opti-CAM: Optimizing saliency maps for interpretability
by: Zhang, Hanwei, et al.
Published: (2023) -
A Learning Paradigm for Interpretable Gradients
by: Figueroa, Felipe Torres, et al.
Published: (2024) -
DP-Net: Learning Discriminative Parts for image recognition
by: Sicre, Ronan, et al.
Published: (2024) -
Multi-Target Unsupervised Domain Adaptation for Semantic Segmentation without External Data
by: Xu, Yonghao, et al.
Published: (2024) -
Is ImageNet worth 1 video? Learning strong image encoders from 1 long unlabelled video
by: Venkataramanan, Shashanka, et al.
Published: (2023)