Identifying Critical Tokens for Accurate Predictions in Transformer-based Medical Imaging Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kang, Solha, Vankerschaver, Joris, Ozbulak, Utku |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Affordable Tumor Segmentation and Visualization for 3D Breast MRI Using SAM2
by: Kang, Solha, et al.
Published: (2025)
by: Kang, Solha, et al.
Published: (2025)
Detecting Regional Spurious Correlations in Vision Transformers via Token Discarding
by: Kang, Solha, et al.
Published: (2025)
by: Kang, Solha, et al.
Published: (2025)
Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets?
by: Ozbulak, Utku, et al.
Published: (2025)
by: Ozbulak, Utku, et al.
Published: (2025)
Exploring Patient Data Requirements in Training Effective AI Models for MRI-based Breast Cancer Classification
by: Kang, Solha, et al.
Published: (2025)
by: Kang, Solha, et al.
Published: (2025)
Improved Sub-Visible Particle Classification in Flow Imaging Microscopy via Generative AI-Based Image Synthesis
by: Ozbulak, Utku, et al.
Published: (2025)
by: Ozbulak, Utku, et al.
Published: (2025)
SpurBreast: A Curated Dataset for Investigating Spurious Correlations in Real-world Breast MRI Classification
by: Won, Jong Bum, et al.
Published: (2025)
by: Won, Jong Bum, et al.
Published: (2025)
Evaluating Visual Explanations of Attention Maps for Transformer-based Medical Imaging
by: Chung, Minjae, et al.
Published: (2025)
by: Chung, Minjae, et al.
Published: (2025)
One Patient's Annotation is Another One's Initialization: Towards Zero-Shot Surgical Video Segmentation with Cross-Patient Initialization
by: Mousavi, Seyed Amir, et al.
Published: (2025)
by: Mousavi, Seyed Amir, et al.
Published: (2025)
Revisiting the Evaluation Bias Introduced by Frame Sampling Strategies in Surgical Video Segmentation Using SAM2
by: Ozbulak, Utku, et al.
Published: (2025)
by: Ozbulak, Utku, et al.
Published: (2025)
When Tracking Fails: Analyzing Failure Modes of SAM2 for Point-Based Tracking in Surgical Videos
by: Jang, Woowon, et al.
Published: (2025)
by: Jang, Woowon, et al.
Published: (2025)
Color Flow Imaging Microscopy Improves Identification of Stress Sources of Protein Aggregates in Biopharmaceuticals
by: Cohrs, Michaela, et al.
Published: (2025)
by: Cohrs, Michaela, et al.
Published: (2025)
Advancing Medical Image Segmentation: Morphology-Driven Learning with Diffusion Transformer
by: Kang, Sungmin, et al.
Published: (2024)
by: Kang, Sungmin, et al.
Published: (2024)
AT-SNN: Adaptive Tokens for Vision Transformer on Spiking Neural Network
by: Kang, Donghwa, et al.
Published: (2024)
by: Kang, Donghwa, et al.
Published: (2024)
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
by: Zhou, Chunting, et al.
Published: (2024)
by: Zhou, Chunting, et al.
Published: (2024)
TransGUNet: Transformer Meets Graph-based Skip Connection for Medical Image Segmentation
by: Nam, Ju-Hyeon, et al.
Published: (2025)
by: Nam, Ju-Hyeon, et al.
Published: (2025)
Identifiable Token Correspondence for World Models
by: Kim, Youngin, et al.
Published: (2026)
by: Kim, Youngin, et al.
Published: (2026)
Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
by: Das, Dabbrata, et al.
Published: (2025)
by: Das, Dabbrata, et al.
Published: (2025)
Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing
by: Wang, Weixing, et al.
Published: (2025)
by: Wang, Weixing, et al.
Published: (2025)
MedPruner: Training-Free Hierarchical Token Pruning for Efficient 3D Medical Image Understanding in Vision-Language Models
by: Liu, Shengyuan, et al.
Published: (2026)
by: Liu, Shengyuan, et al.
Published: (2026)
Inline Critic Steers Image Editing
by: Kang, Weitai, et al.
Published: (2026)
by: Kang, Weitai, et al.
Published: (2026)
Benchmarking Foundation Models and Parameter-Efficient Fine-Tuning for Prognosis Prediction in Medical Imaging
by: Ruffini, Filippo, et al.
Published: (2025)
by: Ruffini, Filippo, et al.
Published: (2025)
GTP-ViT: Efficient Vision Transformers via Graph-based Token Propagation
by: Xu, Xuwei, et al.
Published: (2023)
by: Xu, Xuwei, et al.
Published: (2023)
Discriminative Class Tokens for Text-to-Image Diffusion Models
by: Schwartz, Idan, et al.
Published: (2023)
by: Schwartz, Idan, et al.
Published: (2023)
DCMM-Transformer: Degree-Corrected Mixed-Membership Attention for Medical Imaging
by: Cheng, Huimin, et al.
Published: (2025)
by: Cheng, Huimin, et al.
Published: (2025)
HU-based Foreground Masking for 3D Medical Masked Image Modeling
by: Lee, Jin, et al.
Published: (2025)
by: Lee, Jin, et al.
Published: (2025)
TinyDrop: Tiny Model Guided Token Dropping for Vision Transformers
by: Wang, Guoxin, et al.
Published: (2025)
by: Wang, Guoxin, et al.
Published: (2025)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
by: Zhao, Min, et al.
Published: (2024)
by: Zhao, Min, et al.
Published: (2024)
ENAT: Rethinking Spatial-temporal Interactions in Token-based Image Synthesis
by: Ni, Zanlin, et al.
Published: (2024)
by: Ni, Zanlin, et al.
Published: (2024)
CGI: Identifying Conditional Generative Models with Example Images
by: Zhou, Zhi, et al.
Published: (2025)
by: Zhou, Zhi, et al.
Published: (2025)
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
by: Qu, Liao, et al.
Published: (2024)
by: Qu, Liao, et al.
Published: (2024)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
by: Ghasemi, Narges, et al.
Published: (2025)
by: Ghasemi, Narges, et al.
Published: (2025)
Multimodal Foundation Models Exploit Text to Make Medical Image Predictions
by: Buckley, Thomas, et al.
Published: (2023)
by: Buckley, Thomas, et al.
Published: (2023)
Accurate Explanation Model for Image Classifiers using Class Association Embedding
by: Xie, Ruitao, et al.
Published: (2024)
by: Xie, Ruitao, et al.
Published: (2024)
Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging
by: Shams, Montasir, et al.
Published: (2025)
by: Shams, Montasir, et al.
Published: (2025)
Compute-Efficient Medical Image Classification with Softmax-Free Transformers and Sequence Normalization
by: Khader, Firas, et al.
Published: (2024)
by: Khader, Firas, et al.
Published: (2024)
GaussianToken: An Effective Image Tokenizer with 2D Gaussian Splatting
by: Dong, Jiajun, et al.
Published: (2025)
by: Dong, Jiajun, et al.
Published: (2025)
Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice
by: Huang, Xiusheng, et al.
Published: (2026)
by: Huang, Xiusheng, et al.
Published: (2026)
Promoting Segment Anything Model towards Highly Accurate Dichotomous Image Segmentation
by: Liu, Xianjie, et al.
Published: (2023)
by: Liu, Xianjie, et al.
Published: (2023)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
by: Zheng, Anlin, et al.
Published: (2025)
by: Zheng, Anlin, et al.
Published: (2025)
QTSeg: A Query Token-Based Dual-Mix Attention Framework with Multi-Level Feature Distribution for Medical Image Segmentation
by: Tran, Phuong-Nam, et al.
Published: (2024)
by: Tran, Phuong-Nam, et al.
Published: (2024)
Similar Items
-
Towards Affordable Tumor Segmentation and Visualization for 3D Breast MRI Using SAM2
by: Kang, Solha, et al.
Published: (2025) -
Detecting Regional Spurious Correlations in Vision Transformers via Token Discarding
by: Kang, Solha, et al.
Published: (2025) -
Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets?
by: Ozbulak, Utku, et al.
Published: (2025) -
Exploring Patient Data Requirements in Training Effective AI Models for MRI-based Breast Cancer Classification
by: Kang, Solha, et al.
Published: (2025) -
Improved Sub-Visible Particle Classification in Flow Imaging Microscopy via Generative AI-Based Image Synthesis
by: Ozbulak, Utku, et al.
Published: (2025)