Multimodal Fusion Learning with Dual Attention for Medical Imaging
Fuente:
arXiv
Saved in:
| Main Authors: | Dhar, Joy, Zaidi, Nayyar, Haghighat, Maryam, Goyal, Puneet, Roy, Sudipta, Alavi, Azadeh, Kumar, Vikas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effective and Robust Multimodal Medical Image Analysis
by: Dhar, Joy, et al.
Published: (2026)
by: Dhar, Joy, et al.
Published: (2026)
Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity
by: Dhar, Joy, et al.
Published: (2026)
by: Dhar, Joy, et al.
Published: (2026)
HyPCA-Net: Advancing Multimodal Fusion in Medical Image Analysis
by: Dhar, J., et al.
Published: (2026)
by: Dhar, J., et al.
Published: (2026)
LightMedSeg: Lightweight 3D Medical Image Segmentation with Learned Spatial Anchors
by: Tyagi, Kavyansh, et al.
Published: (2026)
by: Tyagi, Kavyansh, et al.
Published: (2026)
Beyond Images: Adaptive Fusion of Visual and Textual Data for Food Classification
by: Mittal, Prateek, et al.
Published: (2023)
by: Mittal, Prateek, et al.
Published: (2023)
LINGUAL: Language-INtegrated GUidance in Active Learning for Medical Image Segmentation
by: Islam, Md Shazid, et al.
Published: (2025)
by: Islam, Md Shazid, et al.
Published: (2025)
Fruit Classification System with Deep Learning and Neural Architecture Search
by: Dewi, Christine, et al.
Published: (2024)
by: Dewi, Christine, et al.
Published: (2024)
Dual-Attention Frequency Fusion at Multi-Scale for Joint Segmentation and Deformable Medical Image Registration
by: Zhou, Hongchao, et al.
Published: (2024)
by: Zhou, Hongchao, et al.
Published: (2024)
RefineFormer3D: Efficient 3D Medical Image Segmentation via Adaptive Multi-Scale Transformer with Cross Attention Fusion
by: Tyagi, Kavyansh, et al.
Published: (2026)
by: Tyagi, Kavyansh, et al.
Published: (2026)
Dual Interaction Network with Cross-Image Attention for Medical Image Segmentation
by: Noh, Jeonghyun, et al.
Published: (2025)
by: Noh, Jeonghyun, et al.
Published: (2025)
Intra-Class Probabilistic Embeddings for Uncertainty Estimation in Vision-Language Models
by: Lin, Zhenxiang, et al.
Published: (2025)
by: Lin, Zhenxiang, et al.
Published: (2025)
Towards Efficient Information Fusion: Concentric Dual Fusion Attention Based Multiple Instance Learning for Whole Slide Images
by: Liu, Yujian, et al.
Published: (2024)
by: Liu, Yujian, et al.
Published: (2024)
Multi-dimension Transformer with Attention-based Filtering for Medical Image Segmentation
by: Wang, Wentao, et al.
Published: (2024)
by: Wang, Wentao, et al.
Published: (2024)
SMFusion: Semantic-Preserving Fusion of Multimodal Medical Images for Enhanced Clinical Diagnosis
by: Xiang, Haozhe, et al.
Published: (2025)
by: Xiang, Haozhe, et al.
Published: (2025)
FIAS: Feature Imbalance-Aware Medical Image Segmentation with Dynamic Fusion and Mixing Attention
by: Liu, Xiwei, et al.
Published: (2024)
by: Liu, Xiwei, et al.
Published: (2024)
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism
by: Agarwal, Lakshita, et al.
Published: (2025)
by: Agarwal, Lakshita, et al.
Published: (2025)
MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
CTRL-F: Pairing Convolution with Transformer for Image Classification via Multi-Level Feature Cross-Attention and Representation Learning Fusion
by: EL-Assiouti, Hosam S., et al.
Published: (2024)
by: EL-Assiouti, Hosam S., et al.
Published: (2024)
HyperPointFormer: Multimodal Fusion in 3D Space with Dual-Branch Cross-Attention Transformers
by: Rizaldy, Aldino, et al.
Published: (2025)
by: Rizaldy, Aldino, et al.
Published: (2025)
VISTANet: VIsual Spoken Textual Additive Net for Interpretable Multimodal Emotion Recognition
by: Kumar, Puneet, et al.
Published: (2022)
by: Kumar, Puneet, et al.
Published: (2022)
Multimodal Diffusion Bridge with Attention-Based SAR Fusion for Satellite Image Cloud Removal
by: Hu, Yuyang, et al.
Published: (2025)
by: Hu, Yuyang, et al.
Published: (2025)
Edge-Enhanced Dilated Residual Attention Network for Multimodal Medical Image Fusion
by: Zhou, Meng, et al.
Published: (2024)
by: Zhou, Meng, et al.
Published: (2024)
Pre-training with Random Orthogonal Projection Image Modeling
by: Haghighat, Maryam, et al.
Published: (2023)
by: Haghighat, Maryam, et al.
Published: (2023)
DCAU-Net: Differential Cross Attention and Channel-Spatial Feature Fusion for Medical Image Segmentation
by: Li, Yanxin, et al.
Published: (2026)
by: Li, Yanxin, et al.
Published: (2026)
Interpretable Image Emotion Recognition: A Domain Adaptation Approach Using Facial Expressions
by: Kumar, Puneet, et al.
Published: (2020)
by: Kumar, Puneet, et al.
Published: (2020)
PraNet-V2: Dual-Supervised Reverse Attention for Medical Image Segmentation
by: Hu, Bo-Cheng, et al.
Published: (2025)
by: Hu, Bo-Cheng, et al.
Published: (2025)
VisioPhysioENet: Visual Physiological Engagement Detection Network
by: Singh, Alakhsimar, et al.
Published: (2024)
by: Singh, Alakhsimar, et al.
Published: (2024)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
Primus: Enforcing Attention Usage for 3D Medical Image Segmentation
by: Wald, Tassilo, et al.
Published: (2025)
by: Wald, Tassilo, et al.
Published: (2025)
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
by: Kumar, Deepak, et al.
Published: (2026)
by: Kumar, Deepak, et al.
Published: (2026)
Timing Is Everything: Finding the Optimal Fusion Points in Multimodal Medical Imaging
by: Guarrasi, Valerio, et al.
Published: (2025)
by: Guarrasi, Valerio, et al.
Published: (2025)
TTTFusion: A Test-Time Training-Based Strategy for Multimodal Medical Image Fusion in Surgical Robots
by: Xie, Qinhua, et al.
Published: (2025)
by: Xie, Qinhua, et al.
Published: (2025)
MoCaE: Mixture of Calibrated Experts Significantly Improves Object Detection
by: Oksuz, Kemal, et al.
Published: (2023)
by: Oksuz, Kemal, et al.
Published: (2023)
MambaCAFU: Hybrid Multi-Scale and Multi-Attention Model with Mamba-Based Fusion for Medical Image Segmentation
by: Bui, T-Mai, et al.
Published: (2025)
by: Bui, T-Mai, et al.
Published: (2025)
Disentangled and Interpretable Multimodal Attention Fusion for Cancer Survival Prediction
by: Eijpe, Aniek, et al.
Published: (2025)
by: Eijpe, Aniek, et al.
Published: (2025)
VQA-MHUG: A Gaze Dataset to Study Multimodal Neural Attention in Visual Question Answering
by: Sood, Ekta, et al.
Published: (2021)
by: Sood, Ekta, et al.
Published: (2021)
Multimodal Outer Arithmetic Block Dual Fusion of Whole Slide Images and Omics Data for Precision Oncology
by: Alwazzan, Omnia, et al.
Published: (2024)
by: Alwazzan, Omnia, et al.
Published: (2024)
Task-Generalized Adaptive Cross-Domain Learning for Multimodal Image Fusion
by: Wang, Mengyu, et al.
Published: (2025)
by: Wang, Mengyu, et al.
Published: (2025)
Multimodal Fusion SLAM with Fourier Attention
by: Zhou, Youjie, et al.
Published: (2025)
by: Zhou, Youjie, et al.
Published: (2025)
Detecting Near-Duplicate Face Images
by: Banerjee, Sudipta, et al.
Published: (2024)
by: Banerjee, Sudipta, et al.
Published: (2024)
Similar Items
-
Effective and Robust Multimodal Medical Image Analysis
by: Dhar, Joy, et al.
Published: (2026) -
Certified vs. Empirical Adversarial Robust-ness via Hybrid Convolutions with Attention Stochasticity
by: Dhar, Joy, et al.
Published: (2026) -
HyPCA-Net: Advancing Multimodal Fusion in Medical Image Analysis
by: Dhar, J., et al.
Published: (2026) -
LightMedSeg: Lightweight 3D Medical Image Segmentation with Learned Spatial Anchors
by: Tyagi, Kavyansh, et al.
Published: (2026) -
Beyond Images: Adaptive Fusion of Visual and Textual Data for Food Classification
by: Mittal, Prateek, et al.
Published: (2023)