A Manifold Representation of the Key in Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meng, Li, Goodwin, Morten, Yazidi, Anis, Engelstad, Paal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Maximum Manifold Capacity Representations in State Representation Learning
von: Meng, Li, et al.
Veröffentlicht: (2024)
von: Meng, Li, et al.
Veröffentlicht: (2024)
State Representation Learning Using an Unbalanced Atlas
von: Meng, Li, et al.
Veröffentlicht: (2023)
von: Meng, Li, et al.
Veröffentlicht: (2023)
Deep Reinforcement Learning with Swin Transformers
von: Meng, Li, et al.
Veröffentlicht: (2022)
von: Meng, Li, et al.
Veröffentlicht: (2022)
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
von: Meng, Li, et al.
Veröffentlicht: (2022)
von: Meng, Li, et al.
Veröffentlicht: (2022)
Expert Q-learning: Deep Reinforcement Learning with Coarse State Values from Offline Expert Examples
von: Meng, Li, et al.
Veröffentlicht: (2021)
von: Meng, Li, et al.
Veröffentlicht: (2021)
From Video to EEG: Adapting Joint Embedding Predictive Architecture to Uncover Saptiotemporal Dynamics in Brain Signal Analysis
von: Hojjati, Amirabbas, et al.
Veröffentlicht: (2025)
von: Hojjati, Amirabbas, et al.
Veröffentlicht: (2025)
Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders
von: Kumar, Amandeep, et al.
Veröffentlicht: (2026)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2026)
Mechanistic Understandings of Representation Vulnerabilities and Engineering Robust Vision Transformers
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
GeoViSTA: Geospatial Vision-Tabular Transformer for Multimodal Environment Representation
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
SpaceMesh: A Continuous Representation for Learning Manifold Surface Meshes
von: Shen, Tianchang, et al.
Veröffentlicht: (2024)
von: Shen, Tianchang, et al.
Veröffentlicht: (2024)
Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
von: Huo, Simin, et al.
Veröffentlicht: (2025)
von: Huo, Simin, et al.
Veröffentlicht: (2025)
Vision Transformer-based Adversarial Domain Adaptation
von: Li, Yahan, et al.
Veröffentlicht: (2024)
von: Li, Yahan, et al.
Veröffentlicht: (2024)
Diffusion Transformers with Representation Autoencoders
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
Probing the Representational Power of Sparse Autoencoders in Vision Models
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
von: Olson, Matthew Lyle, et al.
Veröffentlicht: (2025)
Split Adaptation for Pre-trained Vision Transformers
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
von: Wang, Lixu, et al.
Veröffentlicht: (2025)
HEAL-SWIN: A Vision Transformer On The Sphere
von: Carlsson, Oscar, et al.
Veröffentlicht: (2023)
von: Carlsson, Oscar, et al.
Veröffentlicht: (2023)
Approximate Nullspace Augmented Finetuning for Robust Vision Transformers
von: Liu, Haoyang, et al.
Veröffentlicht: (2024)
von: Liu, Haoyang, et al.
Veröffentlicht: (2024)
HQViT: Hybrid Quantum Vision Transformer for Image Classification
von: Zhang, Hui, et al.
Veröffentlicht: (2025)
von: Zhang, Hui, et al.
Veröffentlicht: (2025)
Native Segmentation Vision Transformers
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
Unlocking Noise-Resistant Vision: Key Architectural Secrets for Robust Models
von: Kim, Bum Jun, et al.
Veröffentlicht: (2025)
von: Kim, Bum Jun, et al.
Veröffentlicht: (2025)
Multilingual Diversity Improves Vision-Language Representations
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
von: Nguyen, Thao, et al.
Veröffentlicht: (2024)
Slicing Vision Transformer for Flexible Inference
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
Rotary Position Embedding for Vision Transformer
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
RAViT: Resolution-Adaptive Vision Transformer
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
Decorrelation Speeds Up Vision Transformers
von: Carrigg, Kieran, et al.
Veröffentlicht: (2025)
von: Carrigg, Kieran, et al.
Veröffentlicht: (2025)
HiAP: A Multi-Granular Stochastic Auto-Pruning Framework for Vision Transformers
von: Li, Andy, et al.
Veröffentlicht: (2026)
von: Li, Andy, et al.
Veröffentlicht: (2026)
Using Interleaved Ensemble Unlearning to Keep Backdoors at Bay for Finetuning Vision Transformers
von: Li, Zeyu Michael
Veröffentlicht: (2024)
von: Li, Zeyu Michael
Veröffentlicht: (2024)
PerFormer: A Permutation Based Vision Transformer for Remaining Useful Life Prediction
von: Fan, Zhengyang, et al.
Veröffentlicht: (2025)
von: Fan, Zhengyang, et al.
Veröffentlicht: (2025)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations
von: Jiang, Nick, et al.
Veröffentlicht: (2024)
von: Jiang, Nick, et al.
Veröffentlicht: (2024)
When Does Perceptual Alignment Benefit Vision Representations?
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
von: Sundaram, Shobhita, et al.
Veröffentlicht: (2024)
Dynamic Scene Understanding from Vision-Language Representations
von: Pruss, Shahaf, et al.
Veröffentlicht: (2025)
von: Pruss, Shahaf, et al.
Veröffentlicht: (2025)
Contrastive Forward-Forward: A Training Algorithm of Vision Transformer
von: Aghagolzadeh, Hossein, et al.
Veröffentlicht: (2025)
von: Aghagolzadeh, Hossein, et al.
Veröffentlicht: (2025)
SpectralKD: A Unified Framework for Interpreting and Distilling Vision Transformers via Spectral Analysis
von: Tian, Huiyuan, et al.
Veröffentlicht: (2024)
von: Tian, Huiyuan, et al.
Veröffentlicht: (2024)
A Study on Inference Latency for Vision Transformers on Mobile Devices
von: Li, Zhuojin, et al.
Veröffentlicht: (2025)
von: Li, Zhuojin, et al.
Veröffentlicht: (2025)
Instance-Aware Group Quantization for Vision Transformers
von: Moon, Jaehyeon, et al.
Veröffentlicht: (2024)
von: Moon, Jaehyeon, et al.
Veröffentlicht: (2024)
Self-Supervised Vision Transformers for Writer Retrieval
von: Raven, Tim, et al.
Veröffentlicht: (2024)
von: Raven, Tim, et al.
Veröffentlicht: (2024)
SPoT: Subpixel Placement of Tokens in Vision Transformers
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
Elastic Attention Cores for Scalable Vision Transformers
von: Song, Alan Z., et al.
Veröffentlicht: (2026)
von: Song, Alan Z., et al.
Veröffentlicht: (2026)
Compact Vision Transformer by Reduction of Kernel Complexity
von: Wang, Yancheng, et al.
Veröffentlicht: (2025)
von: Wang, Yancheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Maximum Manifold Capacity Representations in State Representation Learning
von: Meng, Li, et al.
Veröffentlicht: (2024) -
State Representation Learning Using an Unbalanced Atlas
von: Meng, Li, et al.
Veröffentlicht: (2023) -
Deep Reinforcement Learning with Swin Transformers
von: Meng, Li, et al.
Veröffentlicht: (2022) -
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
von: Meng, Li, et al.
Veröffentlicht: (2022) -
Expert Q-learning: Deep Reinforcement Learning with Coarse State Values from Offline Expert Examples
von: Meng, Li, et al.
Veröffentlicht: (2021)