Token Cropr: Faster ViTs for Quite a Few Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Bergner, Benjamin, Lippert, Christoph, Mahendran, Aravindh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
di: Chattopadhyay, Nandish, et al.
Pubblicazione: (2026)
di: Chattopadhyay, Nandish, et al.
Pubblicazione: (2026)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
di: Han, Donghoon, et al.
Pubblicazione: (2023)
di: Han, Donghoon, et al.
Pubblicazione: (2023)
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
di: Chen, Lu, et al.
Pubblicazione: (2025)
di: Chen, Lu, et al.
Pubblicazione: (2025)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
di: Wang, Zhibo, et al.
Pubblicazione: (2026)
di: Wang, Zhibo, et al.
Pubblicazione: (2026)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)
Training-Free Acceleration of ViTs with Delayed Spatial Merging
di: Heo, Jung Hwan, et al.
Pubblicazione: (2023)
di: Heo, Jung Hwan, et al.
Pubblicazione: (2023)
Octic Vision Transformers: Quicker ViTs Through Equivariance
di: Nordström, David, et al.
Pubblicazione: (2025)
di: Nordström, David, et al.
Pubblicazione: (2025)
Exploring the Synergies of Hybrid CNNs and ViTs Architectures for Computer Vision: A survey
di: Yunusa, Haruna, et al.
Pubblicazione: (2024)
di: Yunusa, Haruna, et al.
Pubblicazione: (2024)
Causality $\neq$ Decodability, and Vice Versa: Lessons from Interpreting Counting ViTs
di: Huang, Lianghuan, et al.
Pubblicazione: (2025)
di: Huang, Lianghuan, et al.
Pubblicazione: (2025)
Register and [CLS] tokens yield a decoupling of local and global features in large ViTs
di: Lappe, Alexander, et al.
Pubblicazione: (2025)
di: Lappe, Alexander, et al.
Pubblicazione: (2025)
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
Communication Efficient Split Learning of ViTs with Attention-based Double Compression
di: Alvetreti, Federico, et al.
Pubblicazione: (2025)
di: Alvetreti, Federico, et al.
Pubblicazione: (2025)
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
di: Balasubramanian, Sriram, et al.
Pubblicazione: (2024)
di: Balasubramanian, Sriram, et al.
Pubblicazione: (2024)
CubistMerge: Spatial-Preserving Token Merging For Diverse ViT Backbones
di: Gong, Wenyi, et al.
Pubblicazione: (2025)
di: Gong, Wenyi, et al.
Pubblicazione: (2025)
Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness
di: Elisha, Yehonatan, et al.
Pubblicazione: (2026)
di: Elisha, Yehonatan, et al.
Pubblicazione: (2026)
ViTCAE: ViT-based Class-conditioned Autoencoder
di: Jebraeeli, Vahid, et al.
Pubblicazione: (2025)
di: Jebraeeli, Vahid, et al.
Pubblicazione: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
di: Shah, Arya, et al.
Pubblicazione: (2025)
di: Shah, Arya, et al.
Pubblicazione: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
di: Zhong, Yunshan, et al.
Pubblicazione: (2023)
di: Zhong, Yunshan, et al.
Pubblicazione: (2023)
How to train your ViT for OOD Detection
di: Mueller, Maximilian, et al.
Pubblicazione: (2024)
di: Mueller, Maximilian, et al.
Pubblicazione: (2024)
HydraViT: Stacking Heads for a Scalable ViT
di: Haberer, Janek, et al.
Pubblicazione: (2024)
di: Haberer, Janek, et al.
Pubblicazione: (2024)
ViT-ProtoNet for Few-Shot Image Classification: A Multi-Benchmark Evaluation
di: Mutlu, Abdulvahap, et al.
Pubblicazione: (2025)
di: Mutlu, Abdulvahap, et al.
Pubblicazione: (2025)
U-REPA: Aligning Diffusion U-Nets to ViTs
di: Tian, Yuchuan, et al.
Pubblicazione: (2025)
di: Tian, Yuchuan, et al.
Pubblicazione: (2025)
Elastic ViTs from Pretrained Models without Retraining
di: Simoncini, Walter, et al.
Pubblicazione: (2025)
di: Simoncini, Walter, et al.
Pubblicazione: (2025)
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
di: Liu, Jiani, et al.
Pubblicazione: (2025)
di: Liu, Jiani, et al.
Pubblicazione: (2025)
Pretrained ViTs Yield Versatile Representations For Medical Images
di: Matsoukas, Christos, et al.
Pubblicazione: (2023)
di: Matsoukas, Christos, et al.
Pubblicazione: (2023)
ZACH-ViT: A Zero-Token Vision Transformer with ShuffleStrides Data Augmentation for Robust Lung Ultrasound Classification
di: Angelakis, Athanasios, et al.
Pubblicazione: (2025)
di: Angelakis, Athanasios, et al.
Pubblicazione: (2025)
Which Direction to Choose? An Analysis on the Representation Power of Self-Supervised ViTs in Downstream Tasks
di: Kaltampanidis, Yannis, et al.
Pubblicazione: (2025)
di: Kaltampanidis, Yannis, et al.
Pubblicazione: (2025)
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
di: Liu, Longfei, et al.
Pubblicazione: (2026)
di: Liu, Longfei, et al.
Pubblicazione: (2026)
Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring
di: Huang, Youhan, et al.
Pubblicazione: (2026)
di: Huang, Youhan, et al.
Pubblicazione: (2026)
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer
di: Zhu, Wentao
Pubblicazione: (2024)
di: Zhu, Wentao
Pubblicazione: (2024)
Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay
di: Tong, Jin, et al.
Pubblicazione: (2026)
di: Tong, Jin, et al.
Pubblicazione: (2026)
ProtoS-ViT: Visual foundation models for sparse self-explainable classifications
di: Turbé, Hugues, et al.
Pubblicazione: (2024)
di: Turbé, Hugues, et al.
Pubblicazione: (2024)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
di: Kang, Ben, et al.
Pubblicazione: (2025)
di: Kang, Ben, et al.
Pubblicazione: (2025)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
di: Nag, Shashank, et al.
Pubblicazione: (2025)
di: Nag, Shashank, et al.
Pubblicazione: (2025)
Layer by layer, module by module: Choose both for optimal OOD probing of ViT
di: Odonnat, Ambroise, et al.
Pubblicazione: (2026)
di: Odonnat, Ambroise, et al.
Pubblicazione: (2026)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
di: Siméoni, Oriane, et al.
Pubblicazione: (2023)
CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
di: Ramachandran, Akshat, et al.
Pubblicazione: (2024)
di: Ramachandran, Akshat, et al.
Pubblicazione: (2024)
ViT-MUL: A Baseline Study on Recent Machine Unlearning Methods Applied to Vision Transformers
di: Cho, Ikhyun, et al.
Pubblicazione: (2024)
di: Cho, Ikhyun, et al.
Pubblicazione: (2024)
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
di: Bourriez, Nicolas, et al.
Pubblicazione: (2023)
di: Bourriez, Nicolas, et al.
Pubblicazione: (2023)
GrowTAS: Progressive Expansion from Small to Large Subnets for Efficient ViT Architecture Search
di: Lee, Hyunju, et al.
Pubblicazione: (2025)
di: Lee, Hyunju, et al.
Pubblicazione: (2025)
Documenti analoghi
-
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
di: Chattopadhyay, Nandish, et al.
Pubblicazione: (2026) -
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
di: Han, Donghoon, et al.
Pubblicazione: (2023) -
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
di: Chen, Lu, et al.
Pubblicazione: (2025) -
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
di: Wang, Zhibo, et al.
Pubblicazione: (2026) -
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
di: Hwang, Dongyoon, et al.
Pubblicazione: (2024)