Explainable Image Recognition via Enhanced Slot-attention Based Classifier
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Bowen, Li, Liangzhi, Zhang, Jiahao, Nakashima, Yuta, Nagahara, Hajime |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
Enhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025)
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025)
MIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025)
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025)
Point-Supervised Facial Expression Spotting with Gaussian-Based Instance-Adaptive Intensity Modeling
von: Deng, Yicheng, et al.
Veröffentlicht: (2025)
von: Deng, Yicheng, et al.
Veröffentlicht: (2025)
ReLayout: Towards Real-World Document Understanding via Layout-enhanced Pre-training
von: Jiang, Zhouqiang, et al.
Veröffentlicht: (2024)
von: Jiang, Zhouqiang, et al.
Veröffentlicht: (2024)
Deep Polarization Cues for Single-shot Shape and Subsurface Scattering Estimation
von: Li, Chenhao, et al.
Veröffentlicht: (2024)
von: Li, Chenhao, et al.
Veröffentlicht: (2024)
SpotFormer: Multi-Scale Spatio-Temporal Transformer for Facial Expression Spotting
von: Deng, Yicheng, et al.
Veröffentlicht: (2024)
von: Deng, Yicheng, et al.
Veröffentlicht: (2024)
VASCAR: Content-Aware Layout Generation via Visual-Aware Self-Correction
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2024)
Coded-E2LF: Coded Aperture Light Field Imaging from Events
von: Tsuchida, Tomoya, et al.
Veröffentlicht: (2026)
von: Tsuchida, Tomoya, et al.
Veröffentlicht: (2026)
Multi-Scale Spatio-Temporal Graph Convolutional Network for Facial Expression Spotting
von: Deng, Yicheng, et al.
Veröffentlicht: (2024)
von: Deng, Yicheng, et al.
Veröffentlicht: (2024)
Stable Diffusion Exposed: Gender Bias from Prompt to Image
von: Wu, Yankun, et al.
Veröffentlicht: (2023)
von: Wu, Yankun, et al.
Veröffentlicht: (2023)
CALICO: Confident Active Learning with Integrated Calibration
von: Querol, Lorenzo S., et al.
Veröffentlicht: (2024)
von: Querol, Lorenzo S., et al.
Veröffentlicht: (2024)
NeISF++: Neural Incident Stokes Field for Polarized Inverse Rendering of Conductors and Dielectrics
von: Li, Chenhao, et al.
Veröffentlicht: (2024)
von: Li, Chenhao, et al.
Veröffentlicht: (2024)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
Single-Image Depth from Defocus with Coded Aperture and Diffusion Posterior Sampling
von: Kawachi, Hodaka, et al.
Veröffentlicht: (2025)
von: Kawachi, Hodaka, et al.
Veröffentlicht: (2025)
Explainable Part-Based Vehicle Classifier with Spatial Awareness
von: Caduff, Andreas, et al.
Veröffentlicht: (2026)
von: Caduff, Andreas, et al.
Veröffentlicht: (2026)
Exploring Visual Prompting: Robustness Inheritance and Beyond
von: Li, Qi, et al.
Veröffentlicht: (2025)
von: Li, Qi, et al.
Veröffentlicht: (2025)
Measure Twice, Cut Once: A Semantic-Oriented Approach to Video Temporal Localization with Video LLMs
von: Pang, Zongshang, et al.
Veröffentlicht: (2025)
von: Pang, Zongshang, et al.
Veröffentlicht: (2025)
EMMA: Concept Erasure Benchmark with Comprehensive Semantic Metrics and Diverse Categories
von: Wei, Lu, et al.
Veröffentlicht: (2025)
von: Wei, Lu, et al.
Veröffentlicht: (2025)
Privacy in Image Datasets: A Case Study on Pregnancy Ultrasounds
von: Lohanimit, Rawisara, et al.
Veröffentlicht: (2026)
von: Lohanimit, Rawisara, et al.
Veröffentlicht: (2026)
UniFormer: Unifying Convolution and Self-attention for Visual Recognition
von: Li, Kunchang, et al.
Veröffentlicht: (2022)
von: Li, Kunchang, et al.
Veröffentlicht: (2022)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
From Descriptive Richness to Bias: Unveiling the Dark Side of Generative Image Caption Enrichment
von: Hirota, Yusuke, et al.
Veröffentlicht: (2024)
von: Hirota, Yusuke, et al.
Veröffentlicht: (2024)
Acquiring a Dynamic Light Field through a Single-Shot Coded Image
von: Mizuno, Ryoya, et al.
Veröffentlicht: (2022)
von: Mizuno, Ryoya, et al.
Veröffentlicht: (2022)
EMWaveNet: Physically Explainable Neural Network Based on Electromagnetic Propagation for SAR Target Recognition
von: Li, Zhuoxuan, et al.
Veröffentlicht: (2024)
von: Li, Zhuoxuan, et al.
Veröffentlicht: (2024)
OpenSlot: Mixed Open-Set Recognition with Object-Centric Learning
von: Yin, Xu, et al.
Veröffentlicht: (2024)
von: Yin, Xu, et al.
Veröffentlicht: (2024)
Predicting Video Slot Attention Queries from Random Slot-Feature Pairs
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
Target Refocusing via Attention Redistribution for Open-Vocabulary Semantic Segmentation: An Explainability Perspective
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
von: Li, Jiahao, et al.
Veröffentlicht: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
Time-Efficient Light-Field Acquisition Using Coded Aperture and Events
von: Habuchi, Shuji, et al.
Veröffentlicht: (2024)
von: Habuchi, Shuji, et al.
Veröffentlicht: (2024)
Continuous Sign Language Recognition Based on Motor attention mechanism and frame-level Self-distillation
von: Zhu, Qidan, et al.
Veröffentlicht: (2024)
von: Zhu, Qidan, et al.
Veröffentlicht: (2024)
Slot-VLM: SlowFast Slots for Video-Language Modeling
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
When Slots Compete: Slot Merging in Object-Centric Learning
von: Chatzisavvas, Christos, et al.
Veröffentlicht: (2026)
von: Chatzisavvas, Christos, et al.
Veröffentlicht: (2026)
From Global to Local: Social Bias Transfer in CLIP
von: Ramos, Ryan, et al.
Veröffentlicht: (2025)
von: Ramos, Ryan, et al.
Veröffentlicht: (2025)
Towards Counterfactual and Contrastive Explainability and Transparency of DCNN Image Classifiers
von: Tariq, Syed Ali, et al.
Veröffentlicht: (2025)
von: Tariq, Syed Ali, et al.
Veröffentlicht: (2025)
CNS-Bench: Benchmarking Image Classifier Robustness Under Continuous Nuisance Shifts
von: Dünkel, Olaf, et al.
Veröffentlicht: (2025)
von: Dünkel, Olaf, et al.
Veröffentlicht: (2025)
Enhancing Cross-Dataset Performance of Distracted Driving Detection With Score Softmax Classifier And Dynamic Gaussian Smoothing Supervision
von: Duan, Cong, et al.
Veröffentlicht: (2023)
von: Duan, Cong, et al.
Veröffentlicht: (2023)
Learning Global Object-Centric Representations via Disentangled Slot Attention
von: Chen, Tonglin, et al.
Veröffentlicht: (2024)
von: Chen, Tonglin, et al.
Veröffentlicht: (2024)
Unsupervised Part Discovery via Descriptor-Based Masked Image Restoration with Optimized Constraints
von: Xia, Jiahao, et al.
Veröffentlicht: (2025)
von: Xia, Jiahao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025) -
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025) -
Enhancing Ambiguous Dynamic Facial Expression Recognition with Soft Label-based Data Augmentation
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025) -
MIDAS: Mixing Ambiguous Data with Soft Labels for Dynamic Facial Expression Recognition
von: Kawamura, Ryosuke, et al.
Veröffentlicht: (2025) -
Point-Supervised Facial Expression Spotting with Gaussian-Based Instance-Adaptive Intensity Modeling
von: Deng, Yicheng, et al.
Veröffentlicht: (2025)