LipShiFT: A Certifiably Robust Shift-based Vision Transformer
Fuente:
arXiv
Salvato in:
| Autori principali: | Menon, Rohan, Franco, Nicola, Günnemann, Stephan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Assessing Robustness via Score-Based Adversarial Image Generation
di: Kollovieh, Marcel, et al.
Pubblicazione: (2023)
di: Kollovieh, Marcel, et al.
Pubblicazione: (2023)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
di: Ki, Taekyung, et al.
Pubblicazione: (2023)
di: Ki, Taekyung, et al.
Pubblicazione: (2023)
Hierarchical Randomized Smoothing
di: Scholten, Yan, et al.
Pubblicazione: (2023)
di: Scholten, Yan, et al.
Pubblicazione: (2023)
Accelerating Vision Transformers with Adaptive Patch Sizes
di: Choudhury, Rohan, et al.
Pubblicazione: (2025)
di: Choudhury, Rohan, et al.
Pubblicazione: (2025)
Configuring Data Augmentations to Reduce Variance Shift in Positional Embedding of Vision Transformers
di: Kim, Bum Jun, et al.
Pubblicazione: (2024)
di: Kim, Bum Jun, et al.
Pubblicazione: (2024)
Unsupervised Dynamic Feature Selection for Robust Latent Spaces in Vision Tasks
di: Corcuera, Bruno, et al.
Pubblicazione: (2025)
di: Corcuera, Bruno, et al.
Pubblicazione: (2025)
Rethinking Evaluation Paradigms in IBP-based Certified Training
di: Kaulen, Konstantin, et al.
Pubblicazione: (2026)
di: Kaulen, Konstantin, et al.
Pubblicazione: (2026)
LayerShuffle: Enhancing Robustness in Vision Transformers by Randomizing Layer Execution Order
di: Freiberger, Matthias, et al.
Pubblicazione: (2024)
di: Freiberger, Matthias, et al.
Pubblicazione: (2024)
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
di: Sui, Elaine, et al.
Pubblicazione: (2024)
di: Sui, Elaine, et al.
Pubblicazione: (2024)
Generalized Category Discovery under Domain Shifts: From Vision to Vision-Language Models
di: Wang, Hongjun, et al.
Pubblicazione: (2026)
di: Wang, Hongjun, et al.
Pubblicazione: (2026)
Diverse Prototypical Ensembles Improve Robustness to Subpopulation Shift
di: To, Minh Nguyen Nhat, et al.
Pubblicazione: (2025)
di: To, Minh Nguyen Nhat, et al.
Pubblicazione: (2025)
Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts
di: Crabbé, Jonathan, et al.
Pubblicazione: (2023)
di: Crabbé, Jonathan, et al.
Pubblicazione: (2023)
Robustness Tokens: Towards Adversarial Robustness of Transformers
di: Pulfer, Brian, et al.
Pubblicazione: (2025)
di: Pulfer, Brian, et al.
Pubblicazione: (2025)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
di: Schlarmann, Christian, et al.
Pubblicazione: (2024)
di: Schlarmann, Christian, et al.
Pubblicazione: (2024)
NeuroLip: An Event-driven Spatiotemporal Learning Framework for Cross-Scene Lip-Motion-based Visual Speaker Recognition
di: Yao, Junguang, et al.
Pubblicazione: (2026)
di: Yao, Junguang, et al.
Pubblicazione: (2026)
From Ground to Air: Noise Robustness in Vision Transformers and CNNs for Event-Based Vehicle Classification with Potential UAV Applications
di: Almesafri, Nouf, et al.
Pubblicazione: (2025)
di: Almesafri, Nouf, et al.
Pubblicazione: (2025)
CDF Transform-and-Shift: An effective way to deal with datasets of inhomogeneous cluster densities
di: Zhu, Ye, et al.
Pubblicazione: (2018)
di: Zhu, Ye, et al.
Pubblicazione: (2018)
Improving Interpretation Faithfulness for Vision Transformers
di: Hu, Lijie, et al.
Pubblicazione: (2023)
di: Hu, Lijie, et al.
Pubblicazione: (2023)
Block-Recurrent Dynamics in Vision Transformers
di: Jacobs, Mozes, et al.
Pubblicazione: (2025)
di: Jacobs, Mozes, et al.
Pubblicazione: (2025)
Pixel-level Certified Explanations via Randomized Smoothing
di: Anani, Alaa, et al.
Pubblicazione: (2025)
di: Anani, Alaa, et al.
Pubblicazione: (2025)
VariViT: A Vision Transformer for Variable Image Sizes
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
di: Varma, Aswathi, et al.
Pubblicazione: (2026)
A Survey of the Self Supervised Learning Mechanisms for Vision Transformers
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
di: Khan, Asifullah, et al.
Pubblicazione: (2024)
VaPR -- Vision-language Preference alignment for Reasoning
di: Wadhawan, Rohan, et al.
Pubblicazione: (2025)
di: Wadhawan, Rohan, et al.
Pubblicazione: (2025)
Continual Adaptation of Vision Transformers for Federated Learning
di: Halbe, Shaunak, et al.
Pubblicazione: (2023)
di: Halbe, Shaunak, et al.
Pubblicazione: (2023)
Mechanisms of Non-Monotonic Scaling in Vision Transformers
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
di: Kumar, Anantha Padmanaban Krishna
Pubblicazione: (2025)
Discovering Influential Neuron Path in Vision Transformers
di: Wang, Yifan, et al.
Pubblicazione: (2025)
di: Wang, Yifan, et al.
Pubblicazione: (2025)
DiffiT: Diffusion Vision Transformers for Image Generation
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
ADAPT to Robustify Prompt Tuning Vision Transformers
di: Eskandar, Masih, et al.
Pubblicazione: (2024)
di: Eskandar, Masih, et al.
Pubblicazione: (2024)
Class-Discriminative Attention Maps for Vision Transformers
di: Brocki, Lennart, et al.
Pubblicazione: (2023)
di: Brocki, Lennart, et al.
Pubblicazione: (2023)
Exploring Curriculum Learning for Vision-Language Tasks: A Study on Small-Scale Multimodal Training
di: Saha, Rohan, et al.
Pubblicazione: (2024)
di: Saha, Rohan, et al.
Pubblicazione: (2024)
Hierarchically Robust Zero-shot Vision-language Models
di: Dong, Junhao, et al.
Pubblicazione: (2026)
di: Dong, Junhao, et al.
Pubblicazione: (2026)
Margin and Consistency Supervision for Calibrated and Robust Vision Models
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
Intriguing Equivalence Structures of the Embedding Space of Vision Transformers
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
Oscillation-Reduced MXFP4 Training for Vision Transformers
di: Chen, Yuxiang, et al.
Pubblicazione: (2025)
di: Chen, Yuxiang, et al.
Pubblicazione: (2025)
Enhancing Vision Transformer Explainability Using Artificial Astrocytes
di: Echevarrieta-Catalan, Nicolas, et al.
Pubblicazione: (2025)
di: Echevarrieta-Catalan, Nicolas, et al.
Pubblicazione: (2025)
FasterViT: Fast Vision Transformers with Hierarchical Attention
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
di: Hatamizadeh, Ali, et al.
Pubblicazione: (2023)
Confidence-aware Denoised Fine-tuning of Off-the-shelf Models for Certified Robustness
di: Jang, Suhyeok, et al.
Pubblicazione: (2024)
di: Jang, Suhyeok, et al.
Pubblicazione: (2024)
ViT-2SPN: Vision Transformer-based Dual-Stream Self-Supervised Pretraining Networks for Retinal OCT Classification
di: Saraei, Mohammadreza, et al.
Pubblicazione: (2025)
di: Saraei, Mohammadreza, et al.
Pubblicazione: (2025)
Localized Randomized Smoothing for Collective Robustness Certification
di: Schuchardt, Jan, et al.
Pubblicazione: (2022)
di: Schuchardt, Jan, et al.
Pubblicazione: (2022)
VisTabNet: Adapting Vision Transformers for Tabular Data
di: Wydmański, Witold, et al.
Pubblicazione: (2024)
di: Wydmański, Witold, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Assessing Robustness via Score-Based Adversarial Image Generation
di: Kollovieh, Marcel, et al.
Pubblicazione: (2023) -
StyleLipSync: Style-based Personalized Lip-sync Video Generation
di: Ki, Taekyung, et al.
Pubblicazione: (2023) -
Hierarchical Randomized Smoothing
di: Scholten, Yan, et al.
Pubblicazione: (2023) -
Accelerating Vision Transformers with Adaptive Patch Sizes
di: Choudhury, Rohan, et al.
Pubblicazione: (2025) -
Configuring Data Augmentations to Reduce Variance Shift in Positional Embedding of Vision Transformers
di: Kim, Bum Jun, et al.
Pubblicazione: (2024)