Understanding vision transformer robustness through the lens of out-of-distribution detection
Fuente:
arXiv
Saved in:
| Main Authors: | Kuang, Joey, Wong, Alexander |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial robustness of VAEs through the lens of local geometry
by: Khan, Asif, et al.
Published: (2022)
by: Khan, Asif, et al.
Published: (2022)
A noisy elephant in the room: Is your out-of-distribution detector robust to label noise?
by: Humblot-Renaux, Galadrielle, et al.
Published: (2024)
by: Humblot-Renaux, Galadrielle, et al.
Published: (2024)
Steering CLIP's vision transformer with sparse autoencoders
by: Joseph, Sonia, et al.
Published: (2025)
by: Joseph, Sonia, et al.
Published: (2025)
NECO: NEural Collapse Based Out-of-distribution detection
by: Ammar, Mouïn Ben, et al.
Published: (2023)
by: Ammar, Mouïn Ben, et al.
Published: (2023)
Understanding normalization in contrastive representation learning and out-of-distribution detection
by: Le-Gia, Tai, et al.
Published: (2023)
by: Le-Gia, Tai, et al.
Published: (2023)
Your CLIP has 164 dimensions of noise: Exploring the embeddings covariance eigenspectrum of contrastively pretrained vision-language transformers
by: Grzywaczewski, Jakub, et al.
Published: (2026)
by: Grzywaczewski, Jakub, et al.
Published: (2026)
DOSE3 : Diffusion-based Out-of-distribution detection on SE(3) trajectories
by: Cheng, Hongzhe, et al.
Published: (2025)
by: Cheng, Hongzhe, et al.
Published: (2025)
Confidence Trigger Detection: Accelerating Real-time Tracking-by-detection Systems
by: Ding, Zhicheng, et al.
Published: (2019)
by: Ding, Zhicheng, et al.
Published: (2019)
An explainable vision transformer with transfer learning based efficient drought stress identification
by: Patra, Aswini Kumar, et al.
Published: (2024)
by: Patra, Aswini Kumar, et al.
Published: (2024)
Group-robust Sample Reweighting for Subpopulation Shifts via Influence Functions
by: Qiao, Rui, et al.
Published: (2025)
by: Qiao, Rui, et al.
Published: (2025)
QUEST: A robust attention formulation using query-modulated spherical attention
by: Govindarajan, Hariprasath, et al.
Published: (2026)
by: Govindarajan, Hariprasath, et al.
Published: (2026)
Saliency strikes back: How filtering out high frequencies improves white-box explanations
by: Muzellec, Sabine, et al.
Published: (2023)
by: Muzellec, Sabine, et al.
Published: (2023)
Self-supervised video pretraining yields robust and more human-aligned visual representations
by: Parthasarathy, Nikhil, et al.
Published: (2022)
by: Parthasarathy, Nikhil, et al.
Published: (2022)
Two-Stream temporal transformer for video action classification
by: Kurpukdee, Nattapong, et al.
Published: (2026)
by: Kurpukdee, Nattapong, et al.
Published: (2026)
SELECTOR: Heterogeneous graph network with convolutional masked autoencoder for multimodal robust prediction of cancer survival
by: Pan, Liangrui, et al.
Published: (2024)
by: Pan, Liangrui, et al.
Published: (2024)
Random forest-based out-of-distribution detection for robust lung cancer segmentation
by: Rangnekar, Aneesh, et al.
Published: (2025)
by: Rangnekar, Aneesh, et al.
Published: (2025)
Direct Distillation between Different Domains
by: Tang, Jialiang, et al.
Published: (2024)
by: Tang, Jialiang, et al.
Published: (2024)
BRAVE: Broadening the visual encoding of vision-language models
by: Kar, Oğuzhan Fatih, et al.
Published: (2024)
by: Kar, Oğuzhan Fatih, et al.
Published: (2024)
CTARR: A fast and robust method for identifying anatomical regions on CT images via atlas registration
by: Buddenkotte, Thomas, et al.
Published: (2024)
by: Buddenkotte, Thomas, et al.
Published: (2024)
Evidential Neural Radiance Fields
by: Duan, Ruxiao, et al.
Published: (2026)
by: Duan, Ruxiao, et al.
Published: (2026)
STLDM: Spatio-Temporal Latent Diffusion Model for Precipitation Nowcasting
by: Foo, Shi Quan, et al.
Published: (2025)
by: Foo, Shi Quan, et al.
Published: (2025)
Harnessing small projectors and multiple views for efficient vision pretraining
by: Agrawal, Kumar Krishna, et al.
Published: (2023)
by: Agrawal, Kumar Krishna, et al.
Published: (2023)
Transferring Textual Preferences to Vision-Language Understanding through Model Merging
by: Li, Chen-An, et al.
Published: (2025)
by: Li, Chen-An, et al.
Published: (2025)
Decoding Diffusion: A Scalable Framework for Unsupervised Analysis of Latent Space Biases and Representations Using Natural Language Prompts
by: Zeng, E. Zhixuan, et al.
Published: (2024)
by: Zeng, E. Zhixuan, et al.
Published: (2024)
PIF: Anomaly detection via preference embedding
by: Leveni, Filippo, et al.
Published: (2025)
by: Leveni, Filippo, et al.
Published: (2025)
Defect detection using weakly supervised learning
by: Sevetlidis, Vasileios, et al.
Published: (2023)
by: Sevetlidis, Vasileios, et al.
Published: (2023)
A survey on GANs for computer vision: Recent research, analysis and taxonomy
by: Iglesias, Guillermo, et al.
Published: (2022)
by: Iglesias, Guillermo, et al.
Published: (2022)
CLIP Can Understand Depth
by: Kim, Sohee, et al.
Published: (2024)
by: Kim, Sohee, et al.
Published: (2024)
Understanding Multi-View Transformers
by: Stary, Michal, et al.
Published: (2025)
by: Stary, Michal, et al.
Published: (2025)
See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
Out-of-distribution forgetting: vulnerability of continual learning to intra-class distribution shift
by: Guo, Liangxuan, et al.
Published: (2023)
by: Guo, Liangxuan, et al.
Published: (2023)
Improving deep learning with prior knowledge and cognitive models: A survey on enhancing explainability, adversarial robustness and zero-shot learning
by: Mumuni, Fuseinin, et al.
Published: (2024)
by: Mumuni, Fuseinin, et al.
Published: (2024)
Influence of color correction on pathology detection in Capsule Endoscopy
by: Agossou, Bidossessi Emmanuel, et al.
Published: (2025)
by: Agossou, Bidossessi Emmanuel, et al.
Published: (2025)
De-biasing facial detection system using VAE
by: Kandge, Vedant V., et al.
Published: (2022)
by: Kandge, Vedant V., et al.
Published: (2022)
Leaf diseases detection using deep learning methods
by: Fatimi, El Houcine El
Published: (2024)
by: Fatimi, El Houcine El
Published: (2024)
Advanced computer vision for extracting georeferenced vehicle trajectories from drone imagery
by: Fonod, Robert, et al.
Published: (2024)
by: Fonod, Robert, et al.
Published: (2024)
Reproducibility study on how to find Spurious Correlations, Shortcut Learning, Clever Hans or Group-Distributional non-robustness and how to fix them
by: Delzer, Ole, et al.
Published: (2026)
by: Delzer, Ole, et al.
Published: (2026)
Do Language Models Understand Time?
by: Ding, Xi, et al.
Published: (2024)
by: Ding, Xi, et al.
Published: (2024)
Understanding Visual Concepts Across Models
by: Trabucco, Brandon, et al.
Published: (2024)
by: Trabucco, Brandon, et al.
Published: (2024)
VISion On Request: Enhanced VLLM efficiency with sparse, dynamically selected, vision-language interactions
by: Bulat, Adrian, et al.
Published: (2026)
by: Bulat, Adrian, et al.
Published: (2026)
Similar Items
-
Adversarial robustness of VAEs through the lens of local geometry
by: Khan, Asif, et al.
Published: (2022) -
A noisy elephant in the room: Is your out-of-distribution detector robust to label noise?
by: Humblot-Renaux, Galadrielle, et al.
Published: (2024) -
Steering CLIP's vision transformer with sparse autoencoders
by: Joseph, Sonia, et al.
Published: (2025) -
NECO: NEural Collapse Based Out-of-distribution detection
by: Ammar, Mouïn Ben, et al.
Published: (2023) -
Understanding normalization in contrastive representation learning and out-of-distribution detection
by: Le-Gia, Tai, et al.
Published: (2023)