Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Yunxiang, Ma, Mengmeng, Yao, Ziyu, Peng, Xi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
von: Li, Tang, et al.
Veröffentlicht: (2024)
von: Li, Tang, et al.
Veröffentlicht: (2024)
Beyond Accuracy: Ensuring Correct Predictions With Correct Rationales
von: Li, Tang, et al.
Veröffentlicht: (2024)
von: Li, Tang, et al.
Veröffentlicht: (2024)
Attention Transfer Is Not Universally Effective for Vision Transformers
von: Qin, Huaiyuan, et al.
Veröffentlicht: (2026)
von: Qin, Huaiyuan, et al.
Veröffentlicht: (2026)
Out-Of-Distribution Detection with Diversification (Provably)
von: Yao, Haiyun, et al.
Veröffentlicht: (2024)
von: Yao, Haiyun, et al.
Veröffentlicht: (2024)
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
von: Peng, Bo, et al.
Veröffentlicht: (2026)
von: Peng, Bo, et al.
Veröffentlicht: (2026)
Slicing Vision Transformer for Flexible Inference
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
Recovering Global Data Distribution Locally in Federated Learning
von: Yao, Ziyu
Veröffentlicht: (2024)
von: Yao, Ziyu
Veröffentlicht: (2024)
Efficient Adaptation of Large Vision Transformer via Adapter Re-Composing
von: Dong, Wei, et al.
Veröffentlicht: (2023)
von: Dong, Wei, et al.
Veröffentlicht: (2023)
Point Cloud Synthesis Using Inner Product Transforms
von: Röell, Ernst, et al.
Veröffentlicht: (2024)
von: Röell, Ernst, et al.
Veröffentlicht: (2024)
When Accuracy Is Not Enough: Uncertainty Collapse between Noisy Label Learning and Out-of-Distribution Detection
von: Peng, Ningkang, et al.
Veröffentlicht: (2026)
von: Peng, Ningkang, et al.
Veröffentlicht: (2026)
SeafloorAI: A Large-scale Vision-Language Dataset for Seafloor Geological Survey
von: Nguyen, Kien X., et al.
Veröffentlicht: (2024)
von: Nguyen, Kien X., et al.
Veröffentlicht: (2024)
Privacy-Preserving in Connected and Autonomous Vehicles Through Vision to Text Transformation
von: Rezaei, Abdolazim, et al.
Veröffentlicht: (2025)
von: Rezaei, Abdolazim, et al.
Veröffentlicht: (2025)
SOLO: A Single Transformer for Scalable Vision-Language Modeling
von: Chen, Yangyi, et al.
Veröffentlicht: (2024)
von: Chen, Yangyi, et al.
Veröffentlicht: (2024)
DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval
von: Yang, Ruohong, et al.
Veröffentlicht: (2025)
von: Yang, Ruohong, et al.
Veröffentlicht: (2025)
Benchmarking Bias Mitigation Toward Fairness Without Harm from Vision to LVLMs
von: Tan, Xuwei, et al.
Veröffentlicht: (2026)
von: Tan, Xuwei, et al.
Veröffentlicht: (2026)
Relating CNN-Transformer Fusion Network for Change Detection
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
WriteViT: Handwritten Text Generation with Vision Transformer
von: Nam, Dang Hoai, et al.
Veröffentlicht: (2025)
von: Nam, Dang Hoai, et al.
Veröffentlicht: (2025)
Optimal Transport-Induced Samples against Out-of-Distribution Overconfidence
von: Tang, Keke, et al.
Veröffentlicht: (2026)
von: Tang, Keke, et al.
Veröffentlicht: (2026)
T-QPM: Enabling Temporal Out-Of-Distribution Detection and Domain Generalization for Vision-Language Models in Open-World
von: Naiknaware, Aditi, et al.
Veröffentlicht: (2026)
von: Naiknaware, Aditi, et al.
Veröffentlicht: (2026)
Matryoshka Query Transformer for Large Vision-Language Models
von: Hu, Wenbo, et al.
Veröffentlicht: (2024)
von: Hu, Wenbo, et al.
Veröffentlicht: (2024)
Image Restoration Through Generalized Ornstein-Uhlenbeck Bridge
von: Yue, Conghan, et al.
Veröffentlicht: (2023)
von: Yue, Conghan, et al.
Veröffentlicht: (2023)
Cross-modal Active Complementary Learning with Self-refining Correspondence
von: Qin, Yang, et al.
Veröffentlicht: (2023)
von: Qin, Yang, et al.
Veröffentlicht: (2023)
DiFiC: Your Diffusion Model Holds the Secret to Fine-Grained Clustering
von: Yang, Ruohong, et al.
Veröffentlicht: (2024)
von: Yang, Ruohong, et al.
Veröffentlicht: (2024)
Rethinking the Use of Vision Transformers for AI-Generated Image Detection
von: Park, NaHyeon, et al.
Veröffentlicht: (2025)
von: Park, NaHyeon, et al.
Veröffentlicht: (2025)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
von: Yu, Geng, et al.
Veröffentlicht: (2024)
von: Yu, Geng, et al.
Veröffentlicht: (2024)
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
von: Lou, Meng, et al.
Veröffentlicht: (2026)
von: Lou, Meng, et al.
Veröffentlicht: (2026)
LAION-C: An Out-of-Distribution Benchmark for Web-Scale Vision Models
von: Li, Fanfei, et al.
Veröffentlicht: (2025)
von: Li, Fanfei, et al.
Veröffentlicht: (2025)
Unveil Benign Overfitting for Transformer in Vision: Training Dynamics, Convergence, and Generalization
von: Jiang, Jiarui, et al.
Veröffentlicht: (2024)
von: Jiang, Jiarui, et al.
Veröffentlicht: (2024)
On the Detection of Anomalous or Out-Of-Distribution Data in Vision Models Using Statistical Techniques
von: O'Mahony, Laura, et al.
Veröffentlicht: (2024)
von: O'Mahony, Laura, et al.
Veröffentlicht: (2024)
Detecting Out-of-Distribution Objects through Class-Conditioned Inpainting
von: Nguyen, Quang-Huy, et al.
Veröffentlicht: (2024)
von: Nguyen, Quang-Huy, et al.
Veröffentlicht: (2024)
Understanding the Failure Modes of Out-of-Distribution Generalization
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2020)
von: Nagarajan, Vaishnavh, et al.
Veröffentlicht: (2020)
Native Segmentation Vision Transformers
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
Learning with Mixture of Prototypes for Out-of-Distribution Detection
von: Lu, Haodong, et al.
Veröffentlicht: (2024)
von: Lu, Haodong, et al.
Veröffentlicht: (2024)
Oscillation-Reduced MXFP4 Training for Vision Transformers
von: Chen, Yuxiang, et al.
Veröffentlicht: (2025)
von: Chen, Yuxiang, et al.
Veröffentlicht: (2025)
ELSA: Exact Linear-Scan Attention for Fast and Memory-Light Vision Transformers
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2026)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2026)
SelaFD:Seamless Adaptation of Vision Transformer Fine-tuning for Radar-based Human Activity Recognition
von: Wang, Yijun, et al.
Veröffentlicht: (2025)
von: Wang, Yijun, et al.
Veröffentlicht: (2025)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
von: Wang, Siqi, et al.
Veröffentlicht: (2025)
BOOD: Boundary-based Out-Of-Distribution Data Generation
von: Liao, Qilin, et al.
Veröffentlicht: (2025)
von: Liao, Qilin, et al.
Veröffentlicht: (2025)
Hypercone Assisted Contour Generation for Out-of-Distribution Detection
von: Vapsi, Annita, et al.
Veröffentlicht: (2025)
von: Vapsi, Annita, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
von: Li, Tang, et al.
Veröffentlicht: (2024) -
Beyond Accuracy: Ensuring Correct Predictions With Correct Rationales
von: Li, Tang, et al.
Veröffentlicht: (2024) -
Attention Transfer Is Not Universally Effective for Vision Transformers
von: Qin, Huaiyuan, et al.
Veröffentlicht: (2026) -
Out-Of-Distribution Detection with Diversification (Provably)
von: Yao, Haiyun, et al.
Veröffentlicht: (2024) -
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
von: Peng, Bo, et al.
Veröffentlicht: (2026)