Zero-Ablation Overstates Register Content Dependence in DINO Vision Transformers
Fuente:
arXiv
Salvato in:
| Autori principali: | Parodi, Felipe, Matelsky, Jordan, Segado, Melanie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Grounding Intelligence in Movement
di: Segado, Melanie, et al.
Pubblicazione: (2025)
di: Segado, Melanie, et al.
Pubblicazione: (2025)
Leveraging Registers in Vision Transformers for Robust Adaptation
di: Yellapragada, Srikar, et al.
Pubblicazione: (2025)
di: Yellapragada, Srikar, et al.
Pubblicazione: (2025)
DINO Pre-training for Vision-based End-to-end Autonomous Driving
di: Juneja, Shubham, et al.
Pubblicazione: (2024)
di: Juneja, Shubham, et al.
Pubblicazione: (2024)
Activation Quantization of Vision Encoders Needs Prefixing Registers
di: Kim, Seunghyeon, et al.
Pubblicazione: (2025)
di: Kim, Seunghyeon, et al.
Pubblicazione: (2025)
CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2024)
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2024)
PASTA: Vision Transformer Patch Aggregation for Weakly Supervised Target and Anomaly Segmentation
di: Neubauer, Melanie, et al.
Pubblicazione: (2026)
di: Neubauer, Melanie, et al.
Pubblicazione: (2026)
Progressive Semantic-Guided Vision Transformer for Zero-Shot Learning
di: Chen, Shiming, et al.
Pubblicazione: (2024)
di: Chen, Shiming, et al.
Pubblicazione: (2024)
Vision-language models for decoding provider attention during neonatal resuscitation
di: Parodi, Felipe, et al.
Pubblicazione: (2024)
di: Parodi, Felipe, et al.
Pubblicazione: (2024)
SatDINO: A Deep Dive into Self-Supervised Pretraining for Remote Sensing
di: Straka, Jakub, et al.
Pubblicazione: (2025)
di: Straka, Jakub, et al.
Pubblicazione: (2025)
Anomaly Detection by Clustering DINO Embeddings using a Dirichlet Process Mixture
di: Schulthess, Nico, et al.
Pubblicazione: (2025)
di: Schulthess, Nico, et al.
Pubblicazione: (2025)
VET-DINO: Learning Anatomical Understanding Through Multi-View Distillation in Veterinary Imaging
di: Dourson, Andre, et al.
Pubblicazione: (2025)
di: Dourson, Andre, et al.
Pubblicazione: (2025)
Transform-Dependent Adversarial Attacks
di: Tan, Yaoteng, et al.
Pubblicazione: (2024)
di: Tan, Yaoteng, et al.
Pubblicazione: (2024)
I-Segmenter: Integer-Only Vision Transformer for Efficient Semantic Segmentation
di: Sassoon, Jordan, et al.
Pubblicazione: (2025)
di: Sassoon, Jordan, et al.
Pubblicazione: (2025)
VDInstruct: Zero-Shot Key Information Extraction via Content-Aware Vision Tokenization
di: Nguyen, Son, et al.
Pubblicazione: (2025)
di: Nguyen, Son, et al.
Pubblicazione: (2025)
DINO-Detect: A Simple yet Effective Framework for Blur-Robust AI-Generated Image Detection
di: Shen, Jialiang, et al.
Pubblicazione: (2025)
di: Shen, Jialiang, et al.
Pubblicazione: (2025)
Oh-A-DINO: Understanding and Enhancing Attribute-Level Information in Self-Supervised Object-Centric Representations
di: Wagner, Stefan Sylvius, et al.
Pubblicazione: (2025)
di: Wagner, Stefan Sylvius, et al.
Pubblicazione: (2025)
ZACH-ViT: A Zero-Token Vision Transformer with ShuffleStrides Data Augmentation for Robust Lung Ultrasound Classification
di: Angelakis, Athanasios, et al.
Pubblicazione: (2025)
di: Angelakis, Athanasios, et al.
Pubblicazione: (2025)
DINO as a von Mises-Fisher mixture model
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2024)
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2024)
MM-Zero: Self-Evolving Multi-Model Vision Language Models From Zero Data
di: Li, Zongxia, et al.
Pubblicazione: (2026)
di: Li, Zongxia, et al.
Pubblicazione: (2026)
Cryo-CARE: Content-Aware Image Restoration for Cryo-Transmission Electron Microscopy Data
di: Buchholz, Tim-Oliver, et al.
Pubblicazione: (2018)
di: Buchholz, Tim-Oliver, et al.
Pubblicazione: (2018)
On Partial Prototype Collapse in the DINO Family of Self-Supervised Methods
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2024)
di: Govindarajan, Hariprasath, et al.
Pubblicazione: (2024)
Native Segmentation Vision Transformers
di: Brasó, Guillem, et al.
Pubblicazione: (2025)
di: Brasó, Guillem, et al.
Pubblicazione: (2025)
Label Propagation for Zero-shot Classification with Vision-Language Models
di: Stojnić, Vladan, et al.
Pubblicazione: (2024)
di: Stojnić, Vladan, et al.
Pubblicazione: (2024)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
di: Huo, Simin, et al.
Pubblicazione: (2025)
di: Huo, Simin, et al.
Pubblicazione: (2025)
RAViT: Resolution-Adaptive Vision Transformer
di: Guidez, Martial, et al.
Pubblicazione: (2026)
di: Guidez, Martial, et al.
Pubblicazione: (2026)
Slicing Vision Transformer for Flexible Inference
di: Zhang, Yitian, et al.
Pubblicazione: (2024)
di: Zhang, Yitian, et al.
Pubblicazione: (2024)
Rotary Position Embedding for Vision Transformer
di: Heo, Byeongho, et al.
Pubblicazione: (2024)
di: Heo, Byeongho, et al.
Pubblicazione: (2024)
Decorrelation Speeds Up Vision Transformers
di: Carrigg, Kieran, et al.
Pubblicazione: (2025)
di: Carrigg, Kieran, et al.
Pubblicazione: (2025)
Vision Transformers Need Registers
di: Darcet, Timothée, et al.
Pubblicazione: (2023)
di: Darcet, Timothée, et al.
Pubblicazione: (2023)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
di: Nag, Shashank, et al.
Pubblicazione: (2025)
di: Nag, Shashank, et al.
Pubblicazione: (2025)
Semi-Supervised Fine-Tuning of Vision Foundation Models with Content-Style Decomposition
di: Drozdova, Mariia, et al.
Pubblicazione: (2024)
di: Drozdova, Mariia, et al.
Pubblicazione: (2024)
Intriguing Differences Between Zero-Shot and Systematic Evaluations of Vision-Language Transformer Models
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
di: Salman, Shaeke, et al.
Pubblicazione: (2024)
Elastic Attention Cores for Scalable Vision Transformers
di: Song, Alan Z., et al.
Pubblicazione: (2026)
di: Song, Alan Z., et al.
Pubblicazione: (2026)
Attention Transfer Is Not Universally Effective for Vision Transformers
di: Qin, Huaiyuan, et al.
Pubblicazione: (2026)
di: Qin, Huaiyuan, et al.
Pubblicazione: (2026)
SPoT: Subpixel Placement of Tokens in Vision Transformers
di: Hjelkrem-Tan, Martine, et al.
Pubblicazione: (2025)
di: Hjelkrem-Tan, Martine, et al.
Pubblicazione: (2025)
Instance-Aware Group Quantization for Vision Transformers
di: Moon, Jaehyeon, et al.
Pubblicazione: (2024)
di: Moon, Jaehyeon, et al.
Pubblicazione: (2024)
Vision Transformer-based Adversarial Domain Adaptation
di: Li, Yahan, et al.
Pubblicazione: (2024)
di: Li, Yahan, et al.
Pubblicazione: (2024)
Compact Vision Transformer by Reduction of Kernel Complexity
di: Wang, Yancheng, et al.
Pubblicazione: (2025)
di: Wang, Yancheng, et al.
Pubblicazione: (2025)
Split Adaptation for Pre-trained Vision Transformers
di: Wang, Lixu, et al.
Pubblicazione: (2025)
di: Wang, Lixu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Grounding Intelligence in Movement
di: Segado, Melanie, et al.
Pubblicazione: (2025) -
Leveraging Registers in Vision Transformers for Robust Adaptation
di: Yellapragada, Srikar, et al.
Pubblicazione: (2025) -
DINO Pre-training for Vision-based End-to-end Autonomous Driving
di: Juneja, Shubham, et al.
Pubblicazione: (2024) -
Activation Quantization of Vision Encoders Needs Prefixing Registers
di: Kim, Seunghyeon, et al.
Pubblicazione: (2025) -
CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2024)