CNNtention: Can CNNs do better with Attention?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kapila, Nikhil, Glattki, Julian, Rathi, Tejas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LayerAct: Advanced Activation Mechanism for Robust Inference of CNNs
von: Yoon, Kihyuk, et al.
Veröffentlicht: (2023)
von: Yoon, Kihyuk, et al.
Veröffentlicht: (2023)
Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
Poisson Flow Consistency Training
von: Zhang, Anthony, et al.
Veröffentlicht: (2025)
von: Zhang, Anthony, et al.
Veröffentlicht: (2025)
MambaNetLK: Enhancing Colonoscopy Point Cloud Registration with Mamba
von: Jiang, Linzhe, et al.
Veröffentlicht: (2025)
von: Jiang, Linzhe, et al.
Veröffentlicht: (2025)
LRVS-Fashion: Extending Visual Search with Referring Instructions
von: Lepage, Simon, et al.
Veröffentlicht: (2023)
von: Lepage, Simon, et al.
Veröffentlicht: (2023)
Style Transfer with Diffusion Models for Synthetic-to-Real Domain Adaptation
von: Chigot, Estelle, et al.
Veröffentlicht: (2025)
von: Chigot, Estelle, et al.
Veröffentlicht: (2025)
Attention Maps in 3D Shape Classification for Dental Stage Estimation with Class Node Graph Attention Networks
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
Empowering Image Recovery_ A Multi-Attention Approach
von: Wen, Juan, et al.
Veröffentlicht: (2024)
von: Wen, Juan, et al.
Veröffentlicht: (2024)
Search Multilayer Perceptron-Based Fusion for Efficient and Accurate Siamese Tracking
von: Shen, Tianqi, et al.
Veröffentlicht: (2026)
von: Shen, Tianqi, et al.
Veröffentlicht: (2026)
VLM@school -- Evaluation of AI image understanding on German middle school knowledge
von: Peinl, René, et al.
Veröffentlicht: (2025)
von: Peinl, René, et al.
Veröffentlicht: (2025)
Archival Faces: Detection of Faces in Digitized Historical Documents
von: Vaško, Marek, et al.
Veröffentlicht: (2025)
von: Vaško, Marek, et al.
Veröffentlicht: (2025)
How many samples to label for an application given a foundation model? Chest X-ray classification study
von: Nechaev, Nikolay, et al.
Veröffentlicht: (2025)
von: Nechaev, Nikolay, et al.
Veröffentlicht: (2025)
JVLGS: Joint Vision-Language Gas Leak Segmentation
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
Attention Gathers, MLPs Compose: A Causal Analysis of an Action-Outcome Circuit in VideoViT
von: Chereddy, Sai V R
Veröffentlicht: (2026)
von: Chereddy, Sai V R
Veröffentlicht: (2026)
Unsupervised Anomaly Detection Using Diffusion Trend Analysis for Display Inspection
von: Kim, Eunwoo, et al.
Veröffentlicht: (2024)
von: Kim, Eunwoo, et al.
Veröffentlicht: (2024)
Nearest Neighbor Projection Removal Adversarial Training
von: Singh, Himanshu, et al.
Veröffentlicht: (2025)
von: Singh, Himanshu, et al.
Veröffentlicht: (2025)
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
von: Françani, André O., et al.
Veröffentlicht: (2024)
von: Françani, André O., et al.
Veröffentlicht: (2024)
Deep Learning-Based Image Recovery and Pose Estimation for Resident Space Objects
von: Aberdeen, Louis, et al.
Veröffentlicht: (2025)
von: Aberdeen, Louis, et al.
Veröffentlicht: (2025)
RailSafeNet: Visual Scene Understanding for Tram Safety
von: Valach, Ondřej, et al.
Veröffentlicht: (2025)
von: Valach, Ondřej, et al.
Veröffentlicht: (2025)
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
Does CLIP perceive art the same way we do?
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
EatGAN: An Edge-Attention Guided Generative Adversarial Network for Single Image Super-Resolution
von: Rao, Penghao, et al.
Veröffentlicht: (2025)
von: Rao, Penghao, et al.
Veröffentlicht: (2025)
Application of deep learning approaches for medieval historical documents transcription
von: Voloshchuk, Maksym, et al.
Veröffentlicht: (2025)
von: Voloshchuk, Maksym, et al.
Veröffentlicht: (2025)
BAAF: Universal Transformation of One-Class Classifiers for Unsupervised Image Anomaly Detection
von: McIntosh, Declan, et al.
Veröffentlicht: (2026)
von: McIntosh, Declan, et al.
Veröffentlicht: (2026)
Interactive Image Selection and Training for Brain Tumor Segmentation Network
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
Performance Decay in Deepfake Detection: The Limitations of Training on Outdated Data
von: Richings, Jack, et al.
Veröffentlicht: (2025)
von: Richings, Jack, et al.
Veröffentlicht: (2025)
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
von: Muna, Ummay Maria, et al.
Veröffentlicht: (2025)
COCO-Urdu: A Large-Scale Urdu Image-Caption Dataset with Multimodal Quality Estimation
von: Hassan, Umair
Veröffentlicht: (2025)
von: Hassan, Umair
Veröffentlicht: (2025)
Temporal Attention Evolutional Graph Convolutional Network for Multivariate Time Series Forecasting
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
Fine-grained spatial-temporal perception for gas leak segmentation
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
Comparative Analysis of Lightweight Deep Learning Models for Memory-Constrained Devices
von: Shahriar, Tasnim
Veröffentlicht: (2025)
von: Shahriar, Tasnim
Veröffentlicht: (2025)
Handling Out-of-Distribution Data: A Survey
von: Tamang, Lakpa, et al.
Veröffentlicht: (2025)
von: Tamang, Lakpa, et al.
Veröffentlicht: (2025)
HelloMeme: Integrating Spatial Knitting Attentions to Embed High-Level and Fidelity-Rich Conditions in Diffusion Models
von: Zhang, Shengkai, et al.
Veröffentlicht: (2024)
von: Zhang, Shengkai, et al.
Veröffentlicht: (2024)
On Memory: A comparison of memory mechanisms in world models
von: Laird, Eli J., et al.
Veröffentlicht: (2025)
von: Laird, Eli J., et al.
Veröffentlicht: (2025)
Benchmarking Vision Language Models on German Factual Data
von: Peinl, René, et al.
Veröffentlicht: (2025)
von: Peinl, René, et al.
Veröffentlicht: (2025)
Depth Priors in Removal Neural Radiance Fields
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
Data Augmentation and Resolution Enhancement using GANs and Diffusion Models for Tree Segmentation
von: Ferreira, Alessandro dos Santos, et al.
Veröffentlicht: (2025)
von: Ferreira, Alessandro dos Santos, et al.
Veröffentlicht: (2025)
Language Guided Adversarial Purification
von: Singh, Himanshu, et al.
Veröffentlicht: (2023)
von: Singh, Himanshu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LayerAct: Advanced Activation Mechanism for Robust Inference of CNNs
von: Yoon, Kihyuk, et al.
Veröffentlicht: (2023) -
Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024) -
Poisson Flow Consistency Training
von: Zhang, Anthony, et al.
Veröffentlicht: (2025) -
MambaNetLK: Enhancing Colonoscopy Point Cloud Registration with Mamba
von: Jiang, Linzhe, et al.
Veröffentlicht: (2025) -
LRVS-Fashion: Extending Visual Search with Referring Instructions
von: Lepage, Simon, et al.
Veröffentlicht: (2023)