Understanding Transformer-based Vision Models through Inversion
Fuente:
arXiv
Saved in:
| Main Authors: | Rathjens, Jan, Reyhanian, Shirin, Kappel, David, Wiskott, Laurenz |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Hierarchical Quantization Essential for Optimal Reconstruction?
by: Reyhanian, Shirin, et al.
Published: (2026)
by: Reyhanian, Shirin, et al.
Published: (2026)
STREAM: A Universal State-Space Model for Sparse Geometric Data
by: Schöne, Mark, et al.
Published: (2024)
by: Schöne, Mark, et al.
Published: (2024)
SMART: Scalable Mesh-free Aerodynamic Simulations from Raw Geometries using a Transformer-based Surrogate Model
by: Hagnberger, Jan, et al.
Published: (2026)
by: Hagnberger, Jan, et al.
Published: (2026)
On the Surprising Effectiveness of Attention Transfer for Vision Transformers
by: Li, Alexander C., et al.
Published: (2024)
by: Li, Alexander C., et al.
Published: (2024)
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer
by: Zhu, Wentao
Published: (2024)
by: Zhu, Wentao
Published: (2024)
Kolmogorov-Arnold Transformer
by: Yang, Xingyi, et al.
Published: (2024)
by: Yang, Xingyi, et al.
Published: (2024)
Neural Sentinel: Unified Vision Language Model (VLM) for License Plate Recognition with Human-in-the-Loop Continual Learning
by: Sivakoti, Karthik
Published: (2026)
by: Sivakoti, Karthik
Published: (2026)
STAS: Spatio-Temporal Adaptive Computation Time for Spiking Transformers
by: Kang, Donghwa, et al.
Published: (2025)
by: Kang, Donghwa, et al.
Published: (2025)
It's Not a Lottery, It's a Race: Understanding How Gradient Descent Adapts the Network's Capacity to the Task
by: Pinson, Hannah
Published: (2026)
by: Pinson, Hannah
Published: (2026)
DS2TA: Denoising Spiking Transformer with Attenuated Spatiotemporal Attention
by: Xu, Boxun, et al.
Published: (2024)
by: Xu, Boxun, et al.
Published: (2024)
Domain Generalization through Meta-Learning: A Survey
by: Khoee, Arsham Gholamzadeh, et al.
Published: (2024)
by: Khoee, Arsham Gholamzadeh, et al.
Published: (2024)
Making Images from Images: Interleaving Denoising and Transformation
by: Baluja, Shumeet, et al.
Published: (2024)
by: Baluja, Shumeet, et al.
Published: (2024)
Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models
by: McAllister, David, et al.
Published: (2026)
by: McAllister, David, et al.
Published: (2026)
Encodings for Prediction-based Neural Architecture Search
by: Akhauri, Yash, et al.
Published: (2024)
by: Akhauri, Yash, et al.
Published: (2024)
The Cooperative Network Architecture: Learning Structured Networks as Representation of Sensory Patterns
by: Sager, Pascal J., et al.
Published: (2024)
by: Sager, Pascal J., et al.
Published: (2024)
A Neural Architecture Search Method using Auxiliary Evaluation Metric based on ResNet Architecture
by: Wang, Shang, et al.
Published: (2025)
by: Wang, Shang, et al.
Published: (2025)
Analyzing and Improving the Training Dynamics of Diffusion Models
by: Karras, Tero, et al.
Published: (2023)
by: Karras, Tero, et al.
Published: (2023)
Connective Viewpoints of Signal-to-Noise Diffusion Models
by: Doan, Khanh, et al.
Published: (2024)
by: Doan, Khanh, et al.
Published: (2024)
Multistep Distillation of Diffusion Models via Moment Matching
by: Salimans, Tim, et al.
Published: (2024)
by: Salimans, Tim, et al.
Published: (2024)
Foundry: Distilling 3D Foundation Models for the Edge
by: Letellier, Guillaume, et al.
Published: (2025)
by: Letellier, Guillaume, et al.
Published: (2025)
Guiding a Diffusion Model with a Bad Version of Itself
by: Karras, Tero, et al.
Published: (2024)
by: Karras, Tero, et al.
Published: (2024)
Elastic Spiking Transformers for Efficient Gesture Understanding
by: Ancilotto, Alberto, et al.
Published: (2026)
by: Ancilotto, Alberto, et al.
Published: (2026)
Fooling Contrastive Language-Image Pre-trained Models with CLIPMasterPrints
by: Freiberger, Matthias, et al.
Published: (2023)
by: Freiberger, Matthias, et al.
Published: (2023)
Breaking Free: How to Hack Safety Guardrails in Black-Box Diffusion Models!
by: Kotyan, Shashank, et al.
Published: (2024)
by: Kotyan, Shashank, et al.
Published: (2024)
Applying Guidance in a Limited Interval Improves Sample and Distribution Quality in Diffusion Models
by: Kynkäänniemi, Tuomas, et al.
Published: (2024)
by: Kynkäänniemi, Tuomas, et al.
Published: (2024)
STanHop: Sparse Tandem Hopfield Model for Memory-Enhanced Time Series Prediction
by: Wu, Dennis, et al.
Published: (2023)
by: Wu, Dennis, et al.
Published: (2023)
CALM-PDE: Continuous and Adaptive Convolutions for Latent Space Modeling of Time-dependent PDEs
by: Hagnberger, Jan, et al.
Published: (2025)
by: Hagnberger, Jan, et al.
Published: (2025)
State-of-the-Art Transformer Models for Image Super-Resolution: Techniques, Challenges, and Applications
by: Dutta, Debasish, et al.
Published: (2025)
by: Dutta, Debasish, et al.
Published: (2025)
NeuRN: Neuro-inspired Domain Generalization for Image Classification
by: Jalil, Hamd, et al.
Published: (2025)
by: Jalil, Hamd, et al.
Published: (2025)
Mice to Machines: Neural Representations from Visual Cortex for Domain Generalization
by: Qazi, Ahmed, et al.
Published: (2025)
by: Qazi, Ahmed, et al.
Published: (2025)
Concept Probing: Where to Find Human-Defined Concepts (Extended Version)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
by: Ribeiro, Manuel de Sousa, et al.
Published: (2025)
APTx Neuron: A Unified Trainable Neuron Architecture Integrating Activation and Computation
by: Kumar, Ravin
Published: (2025)
by: Kumar, Ravin
Published: (2025)
The Platonic Representation Hypothesis
by: Huh, Minyoung, et al.
Published: (2024)
by: Huh, Minyoung, et al.
Published: (2024)
Neuromimetic metaplasticity for adaptive continual learning
by: Cho, Suhee, et al.
Published: (2024)
by: Cho, Suhee, et al.
Published: (2024)
CoNO: Complex Neural Operator for Continous Dynamical Physical Systems
by: Tiwari, Karn, et al.
Published: (2024)
by: Tiwari, Karn, et al.
Published: (2024)
Benchmarking Spiking Neural Network Learning Methods with Varying Locality
by: Lin, Jiaqi, et al.
Published: (2024)
by: Lin, Jiaqi, et al.
Published: (2024)
Winning the Lottery by Preserving Network Training Dynamics with Concrete Ticket Search
by: Arora, Tanay, et al.
Published: (2025)
by: Arora, Tanay, et al.
Published: (2025)
Learning Evolution via Optimization Knowledge Adaptation
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
When Person Re-Identification Meets Event Camera: A Benchmark Dataset and An Attribute-guided Re-Identification Framework
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Temporal Flexibility in Spiking Neural Networks: Towards Generalization Across Time Steps and Deployment Friendliness
by: Du, Kangrui, et al.
Published: (2025)
by: Du, Kangrui, et al.
Published: (2025)
Similar Items
-
Is Hierarchical Quantization Essential for Optimal Reconstruction?
by: Reyhanian, Shirin, et al.
Published: (2026) -
STREAM: A Universal State-Space Model for Sparse Geometric Data
by: Schöne, Mark, et al.
Published: (2024) -
SMART: Scalable Mesh-free Aerodynamic Simulations from Raw Geometries using a Transformer-based Surrogate Model
by: Hagnberger, Jan, et al.
Published: (2026) -
On the Surprising Effectiveness of Attention Transfer for Vision Transformers
by: Li, Alexander C., et al.
Published: (2024) -
TPC-ViT: Token Propagation Controller for Efficient Vision Transformer
by: Zhu, Wentao
Published: (2024)