Interpreting and Improving Attention From the Perspective of Large Kernel Convolution
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Chenghao, Zhang, Chaoning, Zeng, Boheng, Lu, Yi, Shi, Pengbo, Chen, Qingzi, Liu, Jirui, Zhu, Lingyun, Yang, Yang, Shen, Heng Tao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring Kernel Transformations for Implicit Neural Representations
di: Zheng, Sheng, et al.
Pubblicazione: (2025)
di: Zheng, Sheng, et al.
Pubblicazione: (2025)
SAM Meets UAP: Attacking Segment Anything Model With Universal Adversarial Perturbation
di: Han, Dongshen, et al.
Pubblicazione: (2023)
di: Han, Dongshen, et al.
Pubblicazione: (2023)
LKCell: Efficient Cell Nuclei Instance Segmentation with Large Convolution Kernels
di: Cui, Ziwei, et al.
Pubblicazione: (2024)
di: Cui, Ziwei, et al.
Pubblicazione: (2024)
KPConvX: Modernizing Kernel Point Convolution with Kernel Attention
di: Thomas, Hugues, et al.
Pubblicazione: (2024)
di: Thomas, Hugues, et al.
Pubblicazione: (2024)
RFAConv: Receptive-Field Attention Convolution for Improving Convolutional Neural Networks
di: Zhang, Xin, et al.
Pubblicazione: (2023)
di: Zhang, Xin, et al.
Pubblicazione: (2023)
Syzygy of Thoughts: Improving LLM CoT with the Minimal Free Resolution
di: Li, Chenghao, et al.
Pubblicazione: (2025)
di: Li, Chenghao, et al.
Pubblicazione: (2025)
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention
di: Zhang, Jiangling, et al.
Pubblicazione: (2025)
di: Zhang, Jiangling, et al.
Pubblicazione: (2025)
$ShiftwiseConv:$ Small Convolutional Kernel with Large Kernel Effect
di: Li, Dachong, et al.
Pubblicazione: (2024)
di: Li, Dachong, et al.
Pubblicazione: (2024)
Joint Lossless Compression and Steganography for Medical Images via Large Language Models
di: Zheng, Pengcheng, et al.
Pubblicazione: (2025)
di: Zheng, Pengcheng, et al.
Pubblicazione: (2025)
GateAttentionPose: Enhancing Pose Estimation with Agent Attention and Improved Gated Convolutions
di: Feng, Liang, et al.
Pubblicazione: (2024)
di: Feng, Liang, et al.
Pubblicazione: (2024)
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens
di: Jiang, Zhangqi, et al.
Pubblicazione: (2024)
di: Jiang, Zhangqi, et al.
Pubblicazione: (2024)
Fast SAM2 with Text-Driven Token Pruning
di: Mandal, Avilasha, et al.
Pubblicazione: (2025)
di: Mandal, Avilasha, et al.
Pubblicazione: (2025)
High-Precision Fabric Defect Detection via Adaptive Shape Convolutions and Large Kernel Spatial Modeling
di: Wang, Shuai, et al.
Pubblicazione: (2025)
di: Wang, Shuai, et al.
Pubblicazione: (2025)
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
di: Zeng, Haoxi, et al.
Pubblicazione: (2025)
di: Zeng, Haoxi, et al.
Pubblicazione: (2025)
Partial Convolution Meets Visual Attention
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
di: Bin, Yi, et al.
Pubblicazione: (2024)
di: Bin, Yi, et al.
Pubblicazione: (2024)
Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation
di: Chen, Beitao, et al.
Pubblicazione: (2025)
di: Chen, Beitao, et al.
Pubblicazione: (2025)
Improving Dental Diagnostics: Enhanced Convolution with Spatial Attention Mechanism
di: Rezaie, Shahriar, et al.
Pubblicazione: (2024)
di: Rezaie, Shahriar, et al.
Pubblicazione: (2024)
FreeLong: Training-Free Long Video Generation with SpectralBlend Temporal Attention
di: Lu, Yu, et al.
Pubblicazione: (2024)
di: Lu, Yu, et al.
Pubblicazione: (2024)
Kernel Normalized Convolutional Networks
di: Nasirigerdeh, Reza, et al.
Pubblicazione: (2022)
di: Nasirigerdeh, Reza, et al.
Pubblicazione: (2022)
PeLK: Parameter-efficient Large Kernel ConvNets with Peripheral Convolution
di: Chen, Honghao, et al.
Pubblicazione: (2024)
di: Chen, Honghao, et al.
Pubblicazione: (2024)
Unified Generation and Self-Verification for Vision-Language Models via Advantage Decoupled Preference Optimization
di: Qiu, Xinyu, et al.
Pubblicazione: (2026)
di: Qiu, Xinyu, et al.
Pubblicazione: (2026)
Flattening Singular Values of Factorized Convolution for Medical Images
di: Feng, Zexin, et al.
Pubblicazione: (2024)
di: Feng, Zexin, et al.
Pubblicazione: (2024)
Generative AI meets 3D: A Survey on Text-to-3D in AIGC Era
di: Li, Chenghao, et al.
Pubblicazione: (2023)
di: Li, Chenghao, et al.
Pubblicazione: (2023)
ELA: Efficient Local Attention for Deep Convolutional Neural Networks
di: Xu, Wei, et al.
Pubblicazione: (2024)
di: Xu, Wei, et al.
Pubblicazione: (2024)
Rethinking Normalization Strategies and Convolutional Kernels for Multimodal Image Fusion
di: He, Dan, et al.
Pubblicazione: (2024)
di: He, Dan, et al.
Pubblicazione: (2024)
GRASP: Guided Region-Aware Sparse Prompting for Adapting MLLMs to Remote Sensing
di: Sun, Qigan, et al.
Pubblicazione: (2026)
di: Sun, Qigan, et al.
Pubblicazione: (2026)
Self-supervised Monocular Depth Estimation with Large Kernel Attention
di: Xiang, Xuezhi, et al.
Pubblicazione: (2024)
di: Xiang, Xuezhi, et al.
Pubblicazione: (2024)
Few-Shot Medical Image Segmentation with Large Kernel Attention
di: Wu, Xiaoxiao, et al.
Pubblicazione: (2024)
di: Wu, Xiaoxiao, et al.
Pubblicazione: (2024)
From Attenuation to Attention: Variational Information Flow Manipulation for Fine-Grained Visual Perception
di: Zhu, Jilong, et al.
Pubblicazione: (2026)
di: Zhu, Jilong, et al.
Pubblicazione: (2026)
Revisiting Data Auditing in Large Vision-Language Models
di: Zhu, Hongyu, et al.
Pubblicazione: (2025)
di: Zhu, Hongyu, et al.
Pubblicazione: (2025)
LSK3DNet: Towards Effective and Efficient 3D Perception with Large Sparse Kernels
di: Feng, Tuo, et al.
Pubblicazione: (2024)
di: Feng, Tuo, et al.
Pubblicazione: (2024)
Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
di: Zhang, Chengsheng, et al.
Pubblicazione: (2026)
CACE-Net: Co-guidance Attention and Contrastive Enhancement for Effective Audio-Visual Event Localization
di: He, Xiang, et al.
Pubblicazione: (2024)
di: He, Xiang, et al.
Pubblicazione: (2024)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
di: Yang, Xiangpeng, et al.
Pubblicazione: (2025)
di: Yang, Xiangpeng, et al.
Pubblicazione: (2025)
Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks
di: Wang, Lehan, et al.
Pubblicazione: (2024)
di: Wang, Lehan, et al.
Pubblicazione: (2024)
LKA-ReID:Vehicle Re-Identification with Large Kernel Attention
di: Xiang, Xuezhi, et al.
Pubblicazione: (2024)
di: Xiang, Xuezhi, et al.
Pubblicazione: (2024)
InceptionMamba: An Efficient Hybrid Network with Large Band Convolution and Bottleneck Mamba
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
di: Wang, Yuhang, et al.
Pubblicazione: (2025)
DiffLoRA: Generating Personalized Low-Rank Adaptation Weights with Diffusion
di: Wu, Yujia, et al.
Pubblicazione: (2024)
di: Wu, Yujia, et al.
Pubblicazione: (2024)
Enhanced Convolutional Neural Networks for Improved Image Classification
di: Yang, Xiaoran, et al.
Pubblicazione: (2025)
di: Yang, Xiaoran, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Exploring Kernel Transformations for Implicit Neural Representations
di: Zheng, Sheng, et al.
Pubblicazione: (2025) -
SAM Meets UAP: Attacking Segment Anything Model With Universal Adversarial Perturbation
di: Han, Dongshen, et al.
Pubblicazione: (2023) -
LKCell: Efficient Cell Nuclei Instance Segmentation with Large Convolution Kernels
di: Cui, Ziwei, et al.
Pubblicazione: (2024) -
KPConvX: Modernizing Kernel Point Convolution with Kernel Attention
di: Thomas, Hugues, et al.
Pubblicazione: (2024) -
RFAConv: Receptive-Field Attention Convolution for Improving Convolutional Neural Networks
di: Zhang, Xin, et al.
Pubblicazione: (2023)