SpecFormer: Guarding Vision Transformer Robustness via Maximum Singular Value Penalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Hu, Xixu, Zheng, Runkai, Wang, Jindong, Leung, Cheuk Hang, Wu, Qi, Xie, Xing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Exploring Scale Shift in Crowd Localization under the Context of Domain Generalization
di: Wang, Juncheng, et al.
Pubblicazione: (2025)
di: Wang, Juncheng, et al.
Pubblicazione: (2025)
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
SpecGuard: Spectral Projection-based Advanced Invisible Watermarking
di: Alam, Inzamamul, et al.
Pubblicazione: (2025)
di: Alam, Inzamamul, et al.
Pubblicazione: (2025)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
di: Li, Wenxi, et al.
Pubblicazione: (2025)
di: Li, Wenxi, et al.
Pubblicazione: (2025)
PolaFormer: Polarity-aware Linear Attention for Vision Transformers
di: Meng, Weikang, et al.
Pubblicazione: (2025)
di: Meng, Weikang, et al.
Pubblicazione: (2025)
Distributionally Robust Policy Evaluation and Learning for Continuous Treatment with Observational Data
di: Leung, Cheuk Hang, et al.
Pubblicazione: (2025)
di: Leung, Cheuk Hang, et al.
Pubblicazione: (2025)
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction
di: Chen, Fangda, et al.
Pubblicazione: (2026)
di: Chen, Fangda, et al.
Pubblicazione: (2026)
Risk-Neutral Generative Networks
di: Xian, Zhonghao, et al.
Pubblicazione: (2024)
di: Xian, Zhonghao, et al.
Pubblicazione: (2024)
PlankFormer: Robust Plankton Instance Segmentation via MAE-Pretrained Vision Transformers and Pseudo Community Image Generation
di: Miyazaki, Masaharu, et al.
Pubblicazione: (2026)
di: Miyazaki, Masaharu, et al.
Pubblicazione: (2026)
SLAM-Former: Putting SLAM into One Transformer
di: Yuan, Yijun, et al.
Pubblicazione: (2025)
di: Yuan, Yijun, et al.
Pubblicazione: (2025)
PatchGuard: Adversarially Robust Anomaly Detection and Localization through Vision Transformers and Pseudo Anomalies
di: Nafez, Mojtaba, et al.
Pubblicazione: (2025)
di: Nafez, Mojtaba, et al.
Pubblicazione: (2025)
Probabilistic Learning of Multivariate Time Series with Temporal Irregularity
di: Li, Yijun, et al.
Pubblicazione: (2023)
di: Li, Yijun, et al.
Pubblicazione: (2023)
Distribution-valued Causal Machine Learning: Implications of Credit on Spending Patterns
di: Leung, Cheuk Hang, et al.
Pubblicazione: (2025)
di: Leung, Cheuk Hang, et al.
Pubblicazione: (2025)
SALT: Parameter-Efficient Fine-Tuning via Singular Value Adaptation with Low-Rank Transformation
di: Elsayed, Abdelrahman, et al.
Pubblicazione: (2025)
di: Elsayed, Abdelrahman, et al.
Pubblicazione: (2025)
Finetune Like You Pretrain: Boosting Zero-shot Adversarial Robustness in Vision-language Models
di: Xing, Songlong, et al.
Pubblicazione: (2026)
di: Xing, Songlong, et al.
Pubblicazione: (2026)
HexFormer: Hyperbolic Vision Transformer with Exponential Map Aggregation
di: Alyoussef, Haya, et al.
Pubblicazione: (2026)
di: Alyoussef, Haya, et al.
Pubblicazione: (2026)
TreeFormers -- An Exploration of Vision Transformers for Deforestation Driver Classification
di: Ochuba, Uche
Pubblicazione: (2024)
di: Ochuba, Uche
Pubblicazione: (2024)
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
di: Koleilat, Taha, et al.
Pubblicazione: (2025)
di: Koleilat, Taha, et al.
Pubblicazione: (2025)
CountFormer: Multi-View Crowd Counting Transformer
di: Mo, Hong, et al.
Pubblicazione: (2024)
di: Mo, Hong, et al.
Pubblicazione: (2024)
Boosting 3D Neuron Segmentation with 2D Vision Transformer Pre-trained on Natural Images
di: Cheng, Yik San, et al.
Pubblicazione: (2024)
di: Cheng, Yik San, et al.
Pubblicazione: (2024)
AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer
di: Shan, Jiquan, et al.
Pubblicazione: (2025)
di: Shan, Jiquan, et al.
Pubblicazione: (2025)
SpecSAR-Former: A Lightweight Transformer-based Network for Global LULC Mapping Using Integrated Sentinel-1 and Sentinel-2
di: Yu, Hao, et al.
Pubblicazione: (2024)
di: Yu, Hao, et al.
Pubblicazione: (2024)
PartFormer: Awakening Latent Diverse Representation from Vision Transformer for Object Re-Identification
di: Tan, Lei, et al.
Pubblicazione: (2024)
di: Tan, Lei, et al.
Pubblicazione: (2024)
ImplantFormer: Vision Transformer based Implant Position Regression Using Dental CBCT Data
di: Yang, Xinquan, et al.
Pubblicazione: (2022)
di: Yang, Xinquan, et al.
Pubblicazione: (2022)
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
di: Huang, Yiyan, et al.
Pubblicazione: (2024)
di: Huang, Yiyan, et al.
Pubblicazione: (2024)
SplatFormer: Point Transformer for Robust 3D Gaussian Splatting
di: Chen, Yutong, et al.
Pubblicazione: (2024)
di: Chen, Yutong, et al.
Pubblicazione: (2024)
LoFormer: Local Frequency Transformer for Image Deblurring
di: Mao, Xintian, et al.
Pubblicazione: (2024)
di: Mao, Xintian, et al.
Pubblicazione: (2024)
TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection
di: Chen, Tianxiang, et al.
Pubblicazione: (2024)
di: Chen, Tianxiang, et al.
Pubblicazione: (2024)
AuthGuard: Generalizable Deepfake Detection via Language Guidance
di: Shen, Guangyu, et al.
Pubblicazione: (2025)
di: Shen, Guangyu, et al.
Pubblicazione: (2025)
PDiscoFormer: Relaxing Part Discovery Constraints with Vision Transformers
di: Aniraj, Ananthu, et al.
Pubblicazione: (2024)
di: Aniraj, Ananthu, et al.
Pubblicazione: (2024)
Bootstrapping SparseFormers from Vision Foundation Models
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
VistaFormer: Scalable Vision Transformers for Satellite Image Time Series Segmentation
di: MacDonald, Ezra, et al.
Pubblicazione: (2024)
di: MacDonald, Ezra, et al.
Pubblicazione: (2024)
DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer
di: Tang, Xiaoya, et al.
Pubblicazione: (2025)
di: Tang, Xiaoya, et al.
Pubblicazione: (2025)
TGBFormer: Transformer-GraphFormer Blender Network for Video Object Detection
di: Qi, Qiang, et al.
Pubblicazione: (2025)
di: Qi, Qiang, et al.
Pubblicazione: (2025)
360U-Former: HDR Illumination Estimation with Panoramic Adapted Vision Transformers
di: Hilliard, Jack, et al.
Pubblicazione: (2024)
di: Hilliard, Jack, et al.
Pubblicazione: (2024)
iFormer: Integrating ConvNet and Transformer for Mobile Application
di: Zheng, Chuanyang
Pubblicazione: (2025)
di: Zheng, Chuanyang
Pubblicazione: (2025)
SigFormer: Sparse Signal-Guided Transformer for Multi-Modal Human Action Segmentation
di: Liu, Qi, et al.
Pubblicazione: (2023)
di: Liu, Qi, et al.
Pubblicazione: (2023)
LoLA-SpecViT: Local Attention SwiGLU Vision Transformer with LoRA for Hyperspectral Imaging
di: Zidi, Fadi Abdeladhim, et al.
Pubblicazione: (2025)
di: Zidi, Fadi Abdeladhim, et al.
Pubblicazione: (2025)
DeBiFormer: Vision Transformer with Deformable Agent Bi-level Routing Attention
di: Long, Nguyen Huu Bao, et al.
Pubblicazione: (2024)
di: Long, Nguyen Huu Bao, et al.
Pubblicazione: (2024)
Learning Visual Prompts for Guiding the Attention of Vision Transformers
di: Rezaei, Razieh, et al.
Pubblicazione: (2024)
di: Rezaei, Razieh, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Exploring Scale Shift in Crowd Localization under the Context of Domain Generalization
di: Wang, Juncheng, et al.
Pubblicazione: (2025) -
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
di: Hu, Shengkai, et al.
Pubblicazione: (2025) -
SpecGuard: Spectral Projection-based Advanced Invisible Watermarking
di: Alam, Inzamamul, et al.
Pubblicazione: (2025) -
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
di: Li, Wenxi, et al.
Pubblicazione: (2025) -
PolaFormer: Polarity-aware Linear Attention for Vision Transformers
di: Meng, Weikang, et al.
Pubblicazione: (2025)