Trustworthy Self-Attention: Enabling the Network to Focus Only on the Most Relevant References
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jing, Yu, Yujuan, Tan, Ao, Ren, Duo, Liu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
YOIO: You Only Iterate Once by mining and fusing multiple necessary global information in the optical flow estimation
von: Jing, Yu, et al.
Veröffentlicht: (2024)
von: Jing, Yu, et al.
Veröffentlicht: (2024)
YOLO-PRO: Enhancing Instance-Specific Object Detection with Full-Channel Global Self-Attention
von: Huang, Lin, et al.
Veröffentlicht: (2025)
von: Huang, Lin, et al.
Veröffentlicht: (2025)
Lightweight Backbone Networks Only Require Adaptive Lightweight Self-Attention Mechanisms
von: Li, Fengyun, et al.
Veröffentlicht: (2025)
von: Li, Fengyun, et al.
Veröffentlicht: (2025)
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer
von: Guo, Yufei, et al.
Veröffentlicht: (2025)
von: Guo, Yufei, et al.
Veröffentlicht: (2025)
High-Performance Fine Defect Detection in Artificial Leather Using Dual Feature Pool Object Detection
von: Huang, Lin, et al.
Veröffentlicht: (2023)
von: Huang, Lin, et al.
Veröffentlicht: (2023)
YOLOCS: Object Detection based on Dense Channel Compression for Feature Spatial Solidification
von: Huang, Lin, et al.
Veröffentlicht: (2023)
von: Huang, Lin, et al.
Veröffentlicht: (2023)
RADE-Net: Robust Attention Network for Radar-Only Object Detection in Adverse Weather
von: Leitgeb, Christof, et al.
Veröffentlicht: (2026)
von: Leitgeb, Christof, et al.
Veröffentlicht: (2026)
You Only Need Less Attention at Each Stage in Vision Transformers
von: Zhang, Shuoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Shuoxi, et al.
Veröffentlicht: (2024)
Learning Object Focused Attention
von: Trivedy, Vivek, et al.
Veröffentlicht: (2025)
von: Trivedy, Vivek, et al.
Veröffentlicht: (2025)
YOLO-DS: Fine-Grained Feature Decoupling via Dual-Statistic Synergy Operator for Object Detection
von: Huang, Lin, et al.
Veröffentlicht: (2026)
von: Huang, Lin, et al.
Veröffentlicht: (2026)
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search
von: Tan, Lei, et al.
Veröffentlicht: (2024)
von: Tan, Lei, et al.
Veröffentlicht: (2024)
ToDRE: Effective Visual Token Pruning via Token Diversity and Task Relevance
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
Fully Attentional Networks with Self-emerging Token Labeling
von: Zhao, Bingyin, et al.
Veröffentlicht: (2024)
von: Zhao, Bingyin, et al.
Veröffentlicht: (2024)
You Only Submit One Image to Find the Most Suitable Generative Model
von: Zhou, Zhi, et al.
Veröffentlicht: (2024)
von: Zhou, Zhi, et al.
Veröffentlicht: (2024)
Three-Stream Temporal-Shift Attention Network Based on Self-Knowledge Distillation for Micro-Expression Recognition
von: Zhu, Guanghao, et al.
Veröffentlicht: (2024)
von: Zhu, Guanghao, et al.
Veröffentlicht: (2024)
Hierarchical Graph Attention Network for No-Reference Omnidirectional Image Quality Assessment
von: Yang, Hao, et al.
Veröffentlicht: (2025)
von: Yang, Hao, et al.
Veröffentlicht: (2025)
LSNet: See Large, Focus Small
von: Wang, Ao, et al.
Veröffentlicht: (2025)
von: Wang, Ao, et al.
Veröffentlicht: (2025)
Progressively Normalized Self-Attention Network for Video Polyp Segmentation
von: Ji, Ge-Peng, et al.
Veröffentlicht: (2021)
von: Ji, Ge-Peng, et al.
Veröffentlicht: (2021)
ConFoThinking: Consolidated Focused Attention Driven Thinking for Visual Question Answering
von: Wu, Zhaodong, et al.
Veröffentlicht: (2026)
von: Wu, Zhaodong, et al.
Veröffentlicht: (2026)
HARIS: Human-Like Attention for Reference Image Segmentation
von: Zhang, Mengxi, et al.
Veröffentlicht: (2024)
von: Zhang, Mengxi, et al.
Veröffentlicht: (2024)
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
von: Liu, Wenjie, et al.
Veröffentlicht: (2026)
von: Liu, Wenjie, et al.
Veröffentlicht: (2026)
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers
von: Cao, Ke, et al.
Veröffentlicht: (2025)
von: Cao, Ke, et al.
Veröffentlicht: (2025)
You Only Sample Once: Taming One-Step Text-to-Image Synthesis by Self-Cooperative Diffusion GANs
von: Luo, Yihong, et al.
Veröffentlicht: (2024)
von: Luo, Yihong, et al.
Veröffentlicht: (2024)
MixSA: Training-free Reference-based Sketch Extraction via Mixture-of-Self-Attention
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
Diagnose Like A REAL Pathologist: An Uncertainty-Focused Approach for Trustworthy Multi-Resolution Multiple Instance Learning
von: Hong, Sungrae, et al.
Veröffentlicht: (2025)
von: Hong, Sungrae, et al.
Veröffentlicht: (2025)
Encoder-Only Image Registration
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
von: Chen, Xiang, et al.
Veröffentlicht: (2025)
Optimizing Vision-Language Interactions Through Decoder-Only Models
von: Tanaka, Kaito, et al.
Veröffentlicht: (2024)
von: Tanaka, Kaito, et al.
Veröffentlicht: (2024)
Hierarchical Cross-Attention Network for Virtual Try-On
von: Tang, Hao, et al.
Veröffentlicht: (2024)
von: Tang, Hao, et al.
Veröffentlicht: (2024)
Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning
von: Zhang, Shaojie, et al.
Veröffentlicht: (2025)
von: Zhang, Shaojie, et al.
Veröffentlicht: (2025)
CAMixerSR: Only Details Need More "Attention"
von: Wang, Yan, et al.
Veröffentlicht: (2024)
von: Wang, Yan, et al.
Veröffentlicht: (2024)
TrackRef3D: Multi-View Consistent Track-then-Label for Open-World Referring Segmentation in 3D Gaussian Splatting
von: Tan, Yuyang, et al.
Veröffentlicht: (2026)
von: Tan, Yuyang, et al.
Veröffentlicht: (2026)
CM-MaskSD: Cross-Modality Masked Self-Distillation for Referring Image Segmentation
von: Wang, Wenxuan, et al.
Veröffentlicht: (2023)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2023)
RISAM: Referring Image Segmentation via Mutual-Aware Attention Features
von: Zhang, Mengxi, et al.
Veröffentlicht: (2023)
von: Zhang, Mengxi, et al.
Veröffentlicht: (2023)
Unlocking Generalization in Polyp Segmentation with DINO Self-Attention "keys"
von: Monteiro, Carla, et al.
Veröffentlicht: (2025)
von: Monteiro, Carla, et al.
Veröffentlicht: (2025)
MS-Twins: Multi-Scale Deep Self-Attention Networks for Medical Image Segmentation
von: Xu, Jing
Veröffentlicht: (2023)
von: Xu, Jing
Veröffentlicht: (2023)
Self-Supervised Implicit Attention Priors for Point Cloud Reconstruction
von: Fogarty, Kyle, et al.
Veröffentlicht: (2025)
von: Fogarty, Kyle, et al.
Veröffentlicht: (2025)
Learning Robust Convolutional Neural Networks with Relevant Feature Focusing via Explanations
von: Adachi, Kazuki, et al.
Veröffentlicht: (2022)
von: Adachi, Kazuki, et al.
Veröffentlicht: (2022)
WeakMCN: Multi-task Collaborative Network for Weakly Supervised Referring Expression Comprehension and Segmentation
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Federated Distillation for Medical Image Classification: Towards Trustworthy Computer-Aided Diagnosis
von: Ren, Sufen, et al.
Veröffentlicht: (2024)
von: Ren, Sufen, et al.
Veröffentlicht: (2024)
Low-Resolution Self-Attention for Semantic Segmentation
von: Wu, Yu-Huan, et al.
Veröffentlicht: (2023)
von: Wu, Yu-Huan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
YOIO: You Only Iterate Once by mining and fusing multiple necessary global information in the optical flow estimation
von: Jing, Yu, et al.
Veröffentlicht: (2024) -
YOLO-PRO: Enhancing Instance-Specific Object Detection with Full-Channel Global Self-Attention
von: Huang, Lin, et al.
Veröffentlicht: (2025) -
Lightweight Backbone Networks Only Require Adaptive Lightweight Self-Attention Mechanisms
von: Li, Fengyun, et al.
Veröffentlicht: (2025) -
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer
von: Guo, Yufei, et al.
Veröffentlicht: (2025) -
High-Performance Fine Defect Detection in Artificial Leather Using Dual Feature Pool Object Detection
von: Huang, Lin, et al.
Veröffentlicht: (2023)