Calibration Attention: Learning Reliability-Aware Representations for Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Wenhao, Zhang, Wei Emma, Yue, Lin, Xu, Miao, Guo, Mingyu, Maennel, Olaf, Chen, Weitong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probing Routing-Conditional Calibration in Attention-Residual Transformers
by: Liang, Wenhao, et al.
Published: (2026)
by: Liang, Wenhao, et al.
Published: (2026)
PostHoc FREE Calibrating on Kolmogorov Arnold Networks
by: Liang, Wenhao, et al.
Published: (2025)
by: Liang, Wenhao, et al.
Published: (2025)
We Care Each Pixel: Calibrating on Medical Segmentation Model
by: Liang, Wenhao, et al.
Published: (2025)
by: Liang, Wenhao, et al.
Published: (2025)
Test-Time Attention Purification for Backdoored Large Vision Language Models
by: Zhang, Zhifang, et al.
Published: (2026)
by: Zhang, Zhifang, et al.
Published: (2026)
Calibrating Deep Neural Network using Euclidean Distance
by: Liang, Wenhao, et al.
Published: (2024)
by: Liang, Wenhao, et al.
Published: (2024)
Trajectory-Consistent Calibration for Cache-Accelerated Diffusion Models
by: Liang, Mingyu, et al.
Published: (2026)
by: Liang, Mingyu, et al.
Published: (2026)
Attention Retention for Continual Learning with Vision Transformers
by: Lu, Yue, et al.
Published: (2026)
by: Lu, Yue, et al.
Published: (2026)
VAT: Vision Action Transformer by Unlocking Full Representation of ViT
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Decision-Aware Attention Propagation for Vision Transformer Explainability
by: Jo, Sehyeong, et al.
Published: (2026)
by: Jo, Sehyeong, et al.
Published: (2026)
Representative Attention For Vision Transformers
by: Li, Yuntong, et al.
Published: (2026)
by: Li, Yuntong, et al.
Published: (2026)
SPFormer: Enhancing Vision Transformer with Superpixel Representation
by: Mei, Jieru, et al.
Published: (2024)
by: Mei, Jieru, et al.
Published: (2024)
MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models
by: Guo, Yuncheng, et al.
Published: (2025)
by: Guo, Yuncheng, et al.
Published: (2025)
A Survey of Deep Learning-based Radiology Report Generation Using Multimodal Data
by: Wang, Xinyi, et al.
Published: (2024)
by: Wang, Xinyi, et al.
Published: (2024)
Calibration-Aware Prompt Learning for Medical Vision-Language Models
by: Basu, Abhishek, et al.
Published: (2025)
by: Basu, Abhishek, et al.
Published: (2025)
S2AFormer: Strip Self-Attention for Efficient Vision Transformer
by: Xu, Guoan, et al.
Published: (2025)
by: Xu, Guoan, et al.
Published: (2025)
Tackling the Abstraction and Reasoning Corpus with Vision Transformers: the Importance of 2D Representation, Positions, and Objects
by: Li, Wenhao, et al.
Published: (2024)
by: Li, Wenhao, et al.
Published: (2024)
3D Human Pose Estimation via Spatial Graph Order Attention and Temporal Body Aware Transformer
by: Aouaidjia, Kamel, et al.
Published: (2025)
by: Aouaidjia, Kamel, et al.
Published: (2025)
DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer
by: Tang, Xiaoya, et al.
Published: (2025)
by: Tang, Xiaoya, et al.
Published: (2025)
Polyline Path Masked Attention for Vision Transformer
by: Zhao, Zhongchen, et al.
Published: (2025)
by: Zhao, Zhongchen, et al.
Published: (2025)
Understanding Self-Supervised Pretraining with Part-Aware Representation Learning
by: Zhu, Jie, et al.
Published: (2023)
by: Zhu, Jie, et al.
Published: (2023)
Learning Visual Prompts for Guiding the Attention of Vision Transformers
by: Rezaei, Razieh, et al.
Published: (2024)
by: Rezaei, Razieh, et al.
Published: (2024)
ROI-Aware Multiscale Cross-Attention Vision Transformer for Pest Image Identification
by: Kim, Ga-Eun, et al.
Published: (2023)
by: Kim, Ga-Eun, et al.
Published: (2023)
Visual Imitation Learning with Calibrated Contrastive Representation
by: Wang, Yunke, et al.
Published: (2024)
by: Wang, Yunke, et al.
Published: (2024)
VA-AR: Learning Velocity-Aware Action Representations with Mixture of Window Attention
by: Wei, Jiangning, et al.
Published: (2025)
by: Wei, Jiangning, et al.
Published: (2025)
GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations
by: Guo, Wenxuan, et al.
Published: (2026)
by: Guo, Wenxuan, et al.
Published: (2026)
Forecast then Calibrate: Feature Caching as ODE for Efficient Diffusion Transformers
by: Zheng, Shikang, et al.
Published: (2025)
by: Zheng, Shikang, et al.
Published: (2025)
Calibrated and Resource-Aware Super-Resolution for Reliable Driver Behavior Analysis
by: Shihab, Ibne Farabi, et al.
Published: (2025)
by: Shihab, Ibne Farabi, et al.
Published: (2025)
Vision Transformers with Hierarchical Attention
by: Liu, Yun, et al.
Published: (2021)
by: Liu, Yun, et al.
Published: (2021)
Image Recognition with Online Lightweight Vision Transformer: A Survey
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
Distilling Vision Transformers for Distortion-Robust Representation Learning
by: Alexis, Konstantinos, et al.
Published: (2026)
by: Alexis, Konstantinos, et al.
Published: (2026)
FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation
by: Zhang, Zherui, et al.
Published: (2025)
by: Zhang, Zherui, et al.
Published: (2025)
Pay Attention to the Atlas: Atlas-Guided Test-Time Adaptation Method for Robust 3D Medical Image Segmentation
by: Guo, Jingjie, et al.
Published: (2023)
by: Guo, Jingjie, et al.
Published: (2023)
Structured Initialization for Attention in Vision Transformers
by: Zheng, Jianqiao, et al.
Published: (2024)
by: Zheng, Jianqiao, et al.
Published: (2024)
Vision Transformers are Circulant Attention Learners
by: Han, Dongchen, et al.
Published: (2025)
by: Han, Dongchen, et al.
Published: (2025)
Multi-manifold Attention for Vision Transformers
by: Konstantinidis, Dimitrios, et al.
Published: (2022)
by: Konstantinidis, Dimitrios, et al.
Published: (2022)
HAViT: Historical Attention Vision Transformer
by: Banik, Swarnendu, et al.
Published: (2026)
by: Banik, Swarnendu, et al.
Published: (2026)
You Only Need Less Attention at Each Stage in Vision Transformers
by: Zhang, Shuoxi, et al.
Published: (2024)
by: Zhang, Shuoxi, et al.
Published: (2024)
Optimization of Prompt Learning via Multi-Knowledge Representation for Vision-Language Models
by: Zhang, Enming, et al.
Published: (2024)
by: Zhang, Enming, et al.
Published: (2024)
Learning Where to Edit Vision Transformers
by: Yang, Yunqiao, et al.
Published: (2024)
by: Yang, Yunqiao, et al.
Published: (2024)
Local-Global Context Aware Transformer for Language-Guided Video Segmentation
by: Liang, Chen, et al.
Published: (2022)
by: Liang, Chen, et al.
Published: (2022)
Similar Items
-
Probing Routing-Conditional Calibration in Attention-Residual Transformers
by: Liang, Wenhao, et al.
Published: (2026) -
PostHoc FREE Calibrating on Kolmogorov Arnold Networks
by: Liang, Wenhao, et al.
Published: (2025) -
We Care Each Pixel: Calibrating on Medical Segmentation Model
by: Liang, Wenhao, et al.
Published: (2025) -
Test-Time Attention Purification for Backdoored Large Vision Language Models
by: Zhang, Zhifang, et al.
Published: (2026) -
Calibrating Deep Neural Network using Euclidean Distance
by: Liang, Wenhao, et al.
Published: (2024)