Attention Retention for Continual Learning with Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Yue, Zhou, Xiangyu, Zhang, Shizhou, Xing, Yinghui, Liang, Guoqiang, Zhang, Wencong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Prompt Tuning in Null Space for Continual Learning
von: Lu, Yue, et al.
Veröffentlicht: (2024)
von: Lu, Yue, et al.
Veröffentlicht: (2024)
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
von: Zhang, Shizhou, et al.
Veröffentlicht: (2021)
von: Zhang, Shizhou, et al.
Veröffentlicht: (2021)
Improved Belief-Attention in Vision Task
von: Zhang, Guoqiang
Veröffentlicht: (2026)
von: Zhang, Guoqiang
Veröffentlicht: (2026)
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
von: Xing, Yinghui, et al.
Veröffentlicht: (2023)
von: Xing, Yinghui, et al.
Veröffentlicht: (2023)
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Frequency-Guided Spatial Adaptation for Camouflaged Object Detection
von: Zhang, Shizhou, et al.
Veröffentlicht: (2024)
von: Zhang, Shizhou, et al.
Veröffentlicht: (2024)
WeatherSeg: Weather-Robust Image Segmentation using Teacher-Student Dual Learning and Classifier-Updating Attention
von: Zhang, Zhang, et al.
Veröffentlicht: (2026)
von: Zhang, Zhang, et al.
Veröffentlicht: (2026)
DMAT: A Dynamic Mask-Aware Transformer for Human De-occlusion
von: Liang, Guoqiang, et al.
Veröffentlicht: (2024)
von: Liang, Guoqiang, et al.
Veröffentlicht: (2024)
Probing Routing-Conditional Calibration in Attention-Residual Transformers
von: Liang, Wenhao, et al.
Veröffentlicht: (2026)
von: Liang, Wenhao, et al.
Veröffentlicht: (2026)
PolaFormer: Polarity-aware Linear Attention for Vision Transformers
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
von: Meng, Weikang, et al.
Veröffentlicht: (2025)
EDIT: Enhancing Vision Transformers by Mitigating Attention Sink through an Encoder-Decoder Architecture
von: Feng, Wenfeng, et al.
Veröffentlicht: (2025)
von: Feng, Wenfeng, et al.
Veröffentlicht: (2025)
Dynamic Accumulated Attention Map for Interpreting Evolution of Decision-Making in Vision Transformer
von: Liao, Yi, et al.
Veröffentlicht: (2025)
von: Liao, Yi, et al.
Veröffentlicht: (2025)
The Linear Attention Resurrection in Vision Transformer
von: Zheng, Chuanyang
Veröffentlicht: (2025)
von: Zheng, Chuanyang
Veröffentlicht: (2025)
JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search
von: Zou, Dongyun, et al.
Veröffentlicht: (2026)
von: Zou, Dongyun, et al.
Veröffentlicht: (2026)
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation
von: Cai, Minghong, et al.
Veröffentlicht: (2024)
von: Cai, Minghong, et al.
Veröffentlicht: (2024)
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion
von: Hua, Wei, et al.
Veröffentlicht: (2025)
von: Hua, Wei, et al.
Veröffentlicht: (2025)
LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs
von: Luo, Kun, et al.
Veröffentlicht: (2026)
von: Luo, Kun, et al.
Veröffentlicht: (2026)
Efficient Adaptation of Pre-trained Vision Transformer via Householder Transformation
von: Dong, Wei, et al.
Veröffentlicht: (2024)
von: Dong, Wei, et al.
Veröffentlicht: (2024)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
von: Leem, Saebom, et al.
Veröffentlicht: (2024)
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
von: Yu, Lu, et al.
Veröffentlicht: (2024)
von: Yu, Lu, et al.
Veröffentlicht: (2024)
SPARO: Selective Attention for Robust and Compositional Transformer Encodings for Vision
von: Vani, Ankit, et al.
Veröffentlicht: (2024)
von: Vani, Ankit, et al.
Veröffentlicht: (2024)
CEAT: Continual Expansion and Absorption Transformer for Non-Exemplar Class-Incremental Learning
von: Gao, Xinyuan, et al.
Veröffentlicht: (2024)
von: Gao, Xinyuan, et al.
Veröffentlicht: (2024)
Semi-Supervised Semantic Segmentation Based on Pseudo-Labels: A Survey
von: Ran, Lingyan, et al.
Veröffentlicht: (2024)
von: Ran, Lingyan, et al.
Veröffentlicht: (2024)
Continual Adaptation of Vision Transformers for Federated Learning
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
Hierarchical Modeling for Medical Visual Question Answering with Cross-Attention Fusion
von: Zhang, Junkai, et al.
Veröffentlicht: (2025)
von: Zhang, Junkai, et al.
Veröffentlicht: (2025)
EFTViT: Efficient Federated Training of Vision Transformers with Masked Images on Resource-Constrained Clients
von: Wu, Meihan, et al.
Veröffentlicht: (2024)
von: Wu, Meihan, et al.
Veröffentlicht: (2024)
Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReID
von: Cheng, De, et al.
Veröffentlicht: (2023)
von: Cheng, De, et al.
Veröffentlicht: (2023)
UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives
von: Jiang, Kai, et al.
Veröffentlicht: (2025)
von: Jiang, Kai, et al.
Veröffentlicht: (2025)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
Revisiting the Integration of Convolution and Attention for Vision Backbone
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
Vision Search Assistant: Empower Vision-Language Models as Multimodal Search Engines
von: Zhang, Zhixin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhixin, et al.
Veröffentlicht: (2024)
Continual Vision-and-Language Navigation
von: Jeong, Seongjun, et al.
Veröffentlicht: (2024)
von: Jeong, Seongjun, et al.
Veröffentlicht: (2024)
SOMA: Feature Gradient Enhanced Affine-Flow Matching for SAR-Optical Registration
von: Wang, Haodong, et al.
Veröffentlicht: (2025)
von: Wang, Haodong, et al.
Veröffentlicht: (2025)
Cognitive Alignment At No Cost: Inducing Human Attention Biases For Interpretable Vision Transformers
von: Knights, Ethan
Veröffentlicht: (2026)
von: Knights, Ethan
Veröffentlicht: (2026)
ECViT: Efficient Convolutional Vision Transformer with Local-Attention and Multi-scale Stages
von: Qian, Zhoujie
Veröffentlicht: (2025)
von: Qian, Zhoujie
Veröffentlicht: (2025)
Symbolic Rule Extraction from Attention-Guided Sparse Representations in Vision Transformers
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
GACO-CAD: Geometry-Augmented and Conciseness-Optimized CAD Model Generation from Single Image
von: Wang, Yinghui, et al.
Veröffentlicht: (2025)
von: Wang, Yinghui, et al.
Veröffentlicht: (2025)
A-VL: Adaptive Attention for Large Vision-Language Models
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Visual Prompt Tuning in Null Space for Continual Learning
von: Lu, Yue, et al.
Veröffentlicht: (2024) -
On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective
von: Yang, Shuo, et al.
Veröffentlicht: (2025) -
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
von: Zhang, Shizhou, et al.
Veröffentlicht: (2021) -
Improved Belief-Attention in Vision Task
von: Zhang, Guoqiang
Veröffentlicht: (2026) -
MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balanced Optimization
von: Xing, Yinghui, et al.
Veröffentlicht: (2023)