MuViT: Multi-Resolution Vision Transformers for Learning Across Scales in Microscopy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mantes, Albert Dominguez, La Manno, Gioele, Weigert, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
Federated EndoViT: Pretraining Vision Transformers via Federated Learning on Endoscopic Image Collections
von: Kirchner, Max, et al.
Veröffentlicht: (2025)
von: Kirchner, Max, et al.
Veröffentlicht: (2025)
WriteViT: Handwritten Text Generation with Vision Transformer
von: Nam, Dang Hoai, et al.
Veröffentlicht: (2025)
von: Nam, Dang Hoai, et al.
Veröffentlicht: (2025)
LetheViT: Selective Machine Unlearning for Vision Transformers via Attention-Guided Contrastive Learning
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
RAViT: Resolution-Adaptive Vision Transformer
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
GeoViSTA: Geospatial Vision-Tabular Transformer for Multimodal Environment Representation
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
von: Liu, Yuhao, et al.
Veröffentlicht: (2026)
LL-ViT: Edge Deployable Vision Transformers with Look Up Table Neurons
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
von: Nag, Shashank, et al.
Veröffentlicht: (2025)
ViTNT-FIQA: Training-Free Face Image Quality Assessment with Vision Transformers
von: Ozgur, Guray, et al.
Veröffentlicht: (2026)
von: Ozgur, Guray, et al.
Veröffentlicht: (2026)
GNN-ViTCap: GNN-Enhanced Multiple Instance Learning with Vision Transformers for Whole Slide Image Classification and Captioning
von: Raju, S M Taslim Uddin, et al.
Veröffentlicht: (2025)
von: Raju, S M Taslim Uddin, et al.
Veröffentlicht: (2025)
ViViDex: Learning Vision-based Dexterous Manipulation from Human Videos
von: Chen, Zerui, et al.
Veröffentlicht: (2024)
von: Chen, Zerui, et al.
Veröffentlicht: (2024)
FasterViT: Fast Vision Transformers with Hierarchical Attention
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
ViT-MUL: A Baseline Study on Recent Machine Unlearning Methods Applied to Vision Transformers
von: Cho, Ikhyun, et al.
Veröffentlicht: (2024)
von: Cho, Ikhyun, et al.
Veröffentlicht: (2024)
BornoViT: A Novel Efficient Vision Transformer for Bengali Handwritten Basic Characters Classification
von: Chowdhury, Rafi Hassan, et al.
Veröffentlicht: (2026)
von: Chowdhury, Rafi Hassan, et al.
Veröffentlicht: (2026)
ViTGAN: Training GANs with Vision Transformers
von: Lee, Kwonjoon, et al.
Veröffentlicht: (2021)
von: Lee, Kwonjoon, et al.
Veröffentlicht: (2021)
VariViT: A Vision Transformer for Variable Image Sizes
von: Varma, Aswathi, et al.
Veröffentlicht: (2026)
von: Varma, Aswathi, et al.
Veröffentlicht: (2026)
ScriptViT: Vision Transformer-Based Personalized Handwriting Generation
von: Acharya, Sajjan, et al.
Veröffentlicht: (2025)
von: Acharya, Sajjan, et al.
Veröffentlicht: (2025)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
MuM: Multi-View Masked Image Modeling for 3D Vision
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
TRecViT: A Recurrent Video Transformer
von: Pătrăucean, Viorica, et al.
Veröffentlicht: (2024)
von: Pătrăucean, Viorica, et al.
Veröffentlicht: (2024)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
GCI-ViTAL: Gradual Confidence Improvement with Vision Transformers for Active Learning on Label Noise
von: Mots'oehli, Moseli, et al.
Veröffentlicht: (2024)
von: Mots'oehli, Moseli, et al.
Veröffentlicht: (2024)
ZACH-ViT: A Zero-Token Vision Transformer with ShuffleStrides Data Augmentation for Robust Lung Ultrasound Classification
von: Angelakis, Athanasios, et al.
Veröffentlicht: (2025)
von: Angelakis, Athanasios, et al.
Veröffentlicht: (2025)
S-E Pipeline: A Vision Transformer (ViT) based Resilient Classification Pipeline for Medical Imaging Against Adversarial Attacks
von: S, Neha A, et al.
Veröffentlicht: (2024)
von: S, Neha A, et al.
Veröffentlicht: (2024)
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024)
von: Yao, Ting, et al.
Veröffentlicht: (2024)
Adaptive Resolution Residual Networks -- Generalizing Across Resolutions Easily and Efficiently
von: Demeule, Léa, et al.
Veröffentlicht: (2024)
von: Demeule, Léa, et al.
Veröffentlicht: (2024)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
ViTok-v2: Scaling Native Resolution Auto-Encoders to 5 Billion Parameters
von: Hansen-Estruch, Philippe, et al.
Veröffentlicht: (2026)
von: Hansen-Estruch, Philippe, et al.
Veröffentlicht: (2026)
Multi-Scale High-Resolution Logarithmic Grapher Module for Efficient Vision GNNs
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
Sequence Length Scaling in Vision Transformers for Scientific Images on Frontier
von: Tsaris, Aristeidis, et al.
Veröffentlicht: (2024)
von: Tsaris, Aristeidis, et al.
Veröffentlicht: (2024)
GCond: Gradient Conflict Resolution via Accumulation-based Stabilization for Large-Scale Multi-Task Learning
von: Limarenko, Evgeny Alves, et al.
Veröffentlicht: (2025)
von: Limarenko, Evgeny Alves, et al.
Veröffentlicht: (2025)
ViTAR: Vision Transformer with Any Resolution
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
SkipViT: Speeding Up Vision Transformers with a Token-Level Skip Connection
von: Ataiefard, Foozhan, et al.
Veröffentlicht: (2024)
von: Ataiefard, Foozhan, et al.
Veröffentlicht: (2024)
HistoViT: Vision Transformer for Accurate and Scalable Histopathological Cancer Diagnosis
von: Ahmed, Faisal
Veröffentlicht: (2025)
von: Ahmed, Faisal
Veröffentlicht: (2025)
ViTally Consistent: Scaling Biological Representation Learning for Cell Microscopy
von: Kenyon-Dean, Kian, et al.
Veröffentlicht: (2024)
von: Kenyon-Dean, Kian, et al.
Veröffentlicht: (2024)
ViTCAE: ViT-based Class-conditioned Autoencoder
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
Multi-Scale Deep Learning for Colon Histopathology: A Hybrid Graph-Transformer Approach
von: Saremi, Sadra, et al.
Veröffentlicht: (2025)
von: Saremi, Sadra, et al.
Veröffentlicht: (2025)
ViTCoD: Vision Transformer Acceleration via Dedicated Algorithm and Accelerator Co-Design
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
Exploring the Synergies of Hybrid CNNs and ViTs Architectures for Computer Vision: A survey
von: Yunusa, Haruna, et al.
Veröffentlicht: (2024)
von: Yunusa, Haruna, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025) -
Federated EndoViT: Pretraining Vision Transformers via Federated Learning on Endoscopic Image Collections
von: Kirchner, Max, et al.
Veröffentlicht: (2025) -
WriteViT: Handwritten Text Generation with Vision Transformer
von: Nam, Dang Hoai, et al.
Veröffentlicht: (2025) -
LetheViT: Selective Machine Unlearning for Vision Transformers via Attention-Guided Contrastive Learning
von: Tong, Yujia, et al.
Veröffentlicht: (2025) -
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)