Random Token Fusion for Multi-View Medical Diagnosis
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Jingyu, Matsoukas, Christos, Strand, Fredrik, Smith, Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning What Helps: Task-Aligned Context Selection for Vision Tasks
by: Guo, Jingyu, et al.
Published: (2025)
by: Guo, Jingyu, et al.
Published: (2025)
MuM: Multi-View Masked Image Modeling for 3D Vision
by: Nordström, David, et al.
Published: (2025)
by: Nordström, David, et al.
Published: (2025)
3D-Consistent Multi-View Editing by Correspondence Guidance
by: Bengtson, Josef, et al.
Published: (2025)
by: Bengtson, Josef, et al.
Published: (2025)
VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
by: Scabini, Leonardo, et al.
Published: (2025)
by: Scabini, Leonardo, et al.
Published: (2025)
ARTA: Adaptive Mixed-Resolution Token Allocation for Efficient Dense Feature Extraction
by: Hagerman, David, et al.
Published: (2026)
by: Hagerman, David, et al.
Published: (2026)
k-NN as a Simple and Effective Estimator of Transferability
by: Sorkhei, Moein, et al.
Published: (2025)
by: Sorkhei, Moein, et al.
Published: (2025)
Efficient Self-Supervised Adaptation for Medical Image Analysis
by: Sorkhei, Moein, et al.
Published: (2025)
by: Sorkhei, Moein, et al.
Published: (2025)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
by: Xu, Jie, et al.
Published: (2025)
by: Xu, Jie, et al.
Published: (2025)
Multi-scale Quaternion CNN and BiGRU with Cross Self-attention Feature Fusion for Fault Diagnosis of Bearing
by: Liu, Huanbai, et al.
Published: (2024)
by: Liu, Huanbai, et al.
Published: (2024)
Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation
by: Zhang, Yanhua, et al.
Published: (2024)
by: Zhang, Yanhua, et al.
Published: (2024)
MFAF: An EVA02-Based Multi-scale Frequency Attention Fusion Method for Cross-View Geo-Localization
by: Liu, YiTong, et al.
Published: (2025)
by: Liu, YiTong, et al.
Published: (2025)
Robust and Explainable Bicuspid Aortic Valve Diagnosis Using Stacked Ensembles on Echocardiography
by: Nikolaidis, Christos Chrysanthos, et al.
Published: (2026)
by: Nikolaidis, Christos Chrysanthos, et al.
Published: (2026)
MemFusionMap: Working Memory Fusion for Online Vectorized HD Map Construction
by: Song, Jingyu, et al.
Published: (2024)
by: Song, Jingyu, et al.
Published: (2024)
Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection
by: Senadeera, Damith Chamalke, et al.
Published: (2025)
by: Senadeera, Damith Chamalke, et al.
Published: (2025)
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation
by: Rand, Amit, et al.
Published: (2025)
by: Rand, Amit, et al.
Published: (2025)
Understanding Multi-View Transformers
by: Stary, Michal, et al.
Published: (2025)
by: Stary, Michal, et al.
Published: (2025)
ToDo: Token Downsampling for Efficient Generation of High-Resolution Images
by: Smith, Ethan, et al.
Published: (2024)
by: Smith, Ethan, et al.
Published: (2024)
Multi-Object Tracking Consistently Improves Wildlife Inference
by: Muthivhi, Mufhumudzi, et al.
Published: (2026)
by: Muthivhi, Mufhumudzi, et al.
Published: (2026)
Interpretable Alzheimer's Diagnosis via Multimodal Fusion of Regional Brain Experts
by: Zhuang, Farica, et al.
Published: (2025)
by: Zhuang, Farica, et al.
Published: (2025)
Iterative Refinement Strategy for Automated Data Labeling: Facial Landmark Diagnosis in Medical Imaging
by: Chen, Yu-Hsi
Published: (2024)
by: Chen, Yu-Hsi
Published: (2024)
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
by: Ogezi, Michael, et al.
Published: (2026)
by: Ogezi, Michael, et al.
Published: (2026)
Resolving Token-Space Gradient Conflicts: Token Space Manipulation for Transformer-Based Multi-Task Learning
by: Jeong, Wooseong, et al.
Published: (2025)
by: Jeong, Wooseong, et al.
Published: (2025)
DreamComposer: Controllable 3D Object Generation via Multi-View Conditions
by: Yang, Yunhan, et al.
Published: (2023)
by: Yang, Yunhan, et al.
Published: (2023)
Inference Optimal VLMs Need Fewer Visual Tokens and More Parameters
by: Li, Kevin Y., et al.
Published: (2024)
by: Li, Kevin Y., et al.
Published: (2024)
Toward Efficient Convolutional Neural Networks With Structured Ternary Patterns
by: Kyrkou, Christos
Published: (2024)
by: Kyrkou, Christos
Published: (2024)
Multi-View Hypercomplex Learning for Breast Cancer Screening
by: Lopez, Eleonora, et al.
Published: (2022)
by: Lopez, Eleonora, et al.
Published: (2022)
Multi-View and Multi-Scale Alignment for Contrastive Language-Image Pre-training in Mammography
by: Du, Yuexi, et al.
Published: (2024)
by: Du, Yuexi, et al.
Published: (2024)
UniFusion: Vision-Language Model as Unified Encoder in Image Generation
by: Li, Kevin, et al.
Published: (2025)
by: Li, Kevin, et al.
Published: (2025)
Multi-View 3D Reconstruction using Knowledge Distillation
by: Dutt, Aditya, et al.
Published: (2024)
by: Dutt, Aditya, et al.
Published: (2024)
Towards Robust Uncertainty-Aware Incomplete Multi-View Classification
by: Chen, Mulin, et al.
Published: (2024)
by: Chen, Mulin, et al.
Published: (2024)
FusionCell: Cross-Attentive Fusion of Layout Geometry and Netlist Topology for Standard-Cell Performance Prediction
by: Zhang, Haoyi, et al.
Published: (2026)
by: Zhang, Haoyi, et al.
Published: (2026)
Learnable Expansion of Graph Operators for Multi-Modal Feature Fusion
by: Ding, Dexuan, et al.
Published: (2024)
by: Ding, Dexuan, et al.
Published: (2024)
Contrastive Continual Multi-view Clustering with Filtered Structural Fusion
by: Wan, Xinhang, et al.
Published: (2023)
by: Wan, Xinhang, et al.
Published: (2023)
Multi-Token Prediction Needs Registers
by: Gerontopoulos, Anastasios, et al.
Published: (2025)
by: Gerontopoulos, Anastasios, et al.
Published: (2025)
VidTok: A Versatile and Open-Source Video Tokenizer
by: Tang, Anni, et al.
Published: (2024)
by: Tang, Anni, et al.
Published: (2024)
Hypergraph-based Multi-View Action Recognition using Event Cameras
by: Gao, Yue, et al.
Published: (2024)
by: Gao, Yue, et al.
Published: (2024)
Conformal Trajectory Prediction with Multi-View Data Integration in Cooperative Driving
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Uncertainty Quantification via Hölder Divergence for Multi-View Representation Learning
by: Zhang, Yan, et al.
Published: (2024)
by: Zhang, Yan, et al.
Published: (2024)
Ranking vs. Assignment: The Metric Mismatch in Multi-View Object Association
by: Shelukhan, Matvei, et al.
Published: (2026)
by: Shelukhan, Matvei, et al.
Published: (2026)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
by: Miyato, Takeru, et al.
Published: (2023)
by: Miyato, Takeru, et al.
Published: (2023)
Similar Items
-
Learning What Helps: Task-Aligned Context Selection for Vision Tasks
by: Guo, Jingyu, et al.
Published: (2025) -
MuM: Multi-View Masked Image Modeling for 3D Vision
by: Nordström, David, et al.
Published: (2025) -
3D-Consistent Multi-View Editing by Correspondence Guidance
by: Bengtson, Josef, et al.
Published: (2025) -
VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
by: Scabini, Leonardo, et al.
Published: (2025) -
ARTA: Adaptive Mixed-Resolution Token Allocation for Efficient Dense Feature Extraction
by: Hagerman, David, et al.
Published: (2026)