Causality $\neq$ Decodability, and Vice Versa: Lessons from Interpreting Counting ViTs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Lianghuan, Chang, Yingshan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
von: Chen, Lu, et al.
Veröffentlicht: (2025)
von: Chen, Lu, et al.
Veröffentlicht: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
Token Cropr: Faster ViTs for Quite a Few Tasks
von: Bergner, Benjamin, et al.
Veröffentlicht: (2024)
von: Bergner, Benjamin, et al.
Veröffentlicht: (2024)
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
von: Hwang, Dongyoon, et al.
Veröffentlicht: (2024)
von: Hwang, Dongyoon, et al.
Veröffentlicht: (2024)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
Training-Free Acceleration of ViTs with Delayed Spatial Merging
von: Heo, Jung Hwan, et al.
Veröffentlicht: (2023)
von: Heo, Jung Hwan, et al.
Veröffentlicht: (2023)
Exploring the Synergies of Hybrid CNNs and ViTs Architectures for Computer Vision: A survey
von: Yunusa, Haruna, et al.
Veröffentlicht: (2024)
von: Yunusa, Haruna, et al.
Veröffentlicht: (2024)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
Register and [CLS] tokens yield a decoupling of local and global features in large ViTs
von: Lappe, Alexander, et al.
Veröffentlicht: (2025)
von: Lappe, Alexander, et al.
Veröffentlicht: (2025)
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs
von: Kim, Donghyun, et al.
Veröffentlicht: (2024)
von: Kim, Donghyun, et al.
Veröffentlicht: (2024)
Communication Efficient Split Learning of ViTs with Attention-based Double Compression
von: Alvetreti, Federico, et al.
Veröffentlicht: (2025)
von: Alvetreti, Federico, et al.
Veröffentlicht: (2025)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness
von: Elisha, Yehonatan, et al.
Veröffentlicht: (2026)
von: Elisha, Yehonatan, et al.
Veröffentlicht: (2026)
ViTCAE: ViT-based Class-conditioned Autoencoder
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
von: Liu, Jiani, et al.
Veröffentlicht: (2025)
von: Liu, Jiani, et al.
Veröffentlicht: (2025)
How to train your ViT for OOD Detection
von: Mueller, Maximilian, et al.
Veröffentlicht: (2024)
von: Mueller, Maximilian, et al.
Veröffentlicht: (2024)
Elastic ViTs from Pretrained Models without Retraining
von: Simoncini, Walter, et al.
Veröffentlicht: (2025)
von: Simoncini, Walter, et al.
Veröffentlicht: (2025)
U-REPA: Aligning Diffusion U-Nets to ViTs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
Pretrained ViTs Yield Versatile Representations For Medical Images
von: Matsoukas, Christos, et al.
Veröffentlicht: (2023)
von: Matsoukas, Christos, et al.
Veröffentlicht: (2023)
HydraViT: Stacking Heads for a Scalable ViT
von: Haberer, Janek, et al.
Veröffentlicht: (2024)
von: Haberer, Janek, et al.
Veröffentlicht: (2024)
Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay
von: Tong, Jin, et al.
Veröffentlicht: (2026)
von: Tong, Jin, et al.
Veröffentlicht: (2026)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
von: Kang, Ben, et al.
Veröffentlicht: (2025)
von: Kang, Ben, et al.
Veröffentlicht: (2025)
CubistMerge: Spatial-Preserving Token Merging For Diverse ViT Backbones
von: Gong, Wenyi, et al.
Veröffentlicht: (2025)
von: Gong, Wenyi, et al.
Veröffentlicht: (2025)
ProtoS-ViT: Visual foundation models for sparse self-explainable classifications
von: Turbé, Hugues, et al.
Veröffentlicht: (2024)
von: Turbé, Hugues, et al.
Veröffentlicht: (2024)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
Layer by layer, module by module: Choose both for optimal OOD probing of ViT
von: Odonnat, Ambroise, et al.
Veröffentlicht: (2026)
von: Odonnat, Ambroise, et al.
Veröffentlicht: (2026)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
GrowTAS: Progressive Expansion from Small to Large Subnets for Efficient ViT Architecture Search
von: Lee, Hyunju, et al.
Veröffentlicht: (2025)
von: Lee, Hyunju, et al.
Veröffentlicht: (2025)
Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring
von: Huang, Youhan, et al.
Veröffentlicht: (2026)
von: Huang, Youhan, et al.
Veröffentlicht: (2026)
Triamese-ViT: A 3D-Aware Method for Robust Brain Age Estimation from MRIs
von: Zhang, Zhaonian, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaonian, et al.
Veröffentlicht: (2024)
CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
ViT-MUL: A Baseline Study on Recent Machine Unlearning Methods Applied to Vision Transformers
von: Cho, Ikhyun, et al.
Veröffentlicht: (2024)
von: Cho, Ikhyun, et al.
Veröffentlicht: (2024)
Parameter Efficient Fine-tuning of Self-supervised ViTs without Catastrophic Forgetting
von: Bafghi, Reza Akbarian, et al.
Veröffentlicht: (2024)
von: Bafghi, Reza Akbarian, et al.
Veröffentlicht: (2024)
Class Is Invariant to Context and Vice Versa: On Learning Invariance for Out-Of-Distribution Generalization
von: Qi, Jiaxin, et al.
Veröffentlicht: (2022)
von: Qi, Jiaxin, et al.
Veröffentlicht: (2022)
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
Tex-ViT: A Generalizable, Robust, Texture-based dual-branch cross-attention deepfake detector
von: Dagar, Deepak, et al.
Veröffentlicht: (2024)
von: Dagar, Deepak, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
von: Chen, Lu, et al.
Veröffentlicht: (2025) -
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026) -
Token Cropr: Faster ViTs for Quite a Few Tasks
von: Bergner, Benjamin, et al.
Veröffentlicht: (2024) -
Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2024) -
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
von: Hwang, Dongyoon, et al.
Veröffentlicht: (2024)