MIRA: A Novel Framework for Fusing Modalities in Medical RAG
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jinhong, Ashraf, Tajamul, Han, Zongyan, Laaksonen, Jorma, Anwer, Rao Mohammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
A Fusion-Guided Inception Network for Hyperspectral Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
All in One: Visual-Description-Guided Unified Point Cloud Segmentation
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
von: Thawakar, Omkar, et al.
Veröffentlicht: (2023)
von: Thawakar, Omkar, et al.
Veröffentlicht: (2023)
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark
von: Ghaboura, Sara, et al.
Veröffentlicht: (2025)
von: Ghaboura, Sara, et al.
Veröffentlicht: (2025)
DACN: Dual-Attention Convolutional Network for Hyperspectral Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
HF-Fed: Hierarchical based customized Federated Learning Framework for X-Ray Imaging
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
Hybrid Deep Learning for Hyperspectral Single Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
Generalizable Federated Learning using Client Adaptive Focal Modulation
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
TITAN: Query-Token based Domain Adaptive Adversarial Learning
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
FATE: Focal-modulated Attention Encoder for Multivariate Time-series Forecasting
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
Towards Lightweight Hyperspectral Image Super-Resolution with Depthwise Separable Dilated Convolutional Network
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
MATRIX: Multimodal Agent Tuning for Robust Tool-Use Reasoning
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
A Dual-Domain Convolutional Network for Hyperspectral Single-Image Super-Resolution
von: Karayaka, Murat, et al.
Veröffentlicht: (2025)
von: Karayaka, Murat, et al.
Veröffentlicht: (2025)
MedSPOT: A Workflow-Aware Sequential Grounding Benchmark for Clinical GUI
von: Shakeel, Rozain, et al.
Veröffentlicht: (2026)
von: Shakeel, Rozain, et al.
Veröffentlicht: (2026)
LLM Post-Training: A Deep Dive into Reasoning Large Language Models
von: Kumar, Komal, et al.
Veröffentlicht: (2025)
von: Kumar, Komal, et al.
Veröffentlicht: (2025)
D-MASTER: Mask Annealed Transformer for Unsupervised Domain Adaptation in Breast Cancer Detection from Mammograms
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2024)
CarePilot: A Multi-Agent Framework for Long-Horizon Computer Task Automation in Healthcare
von: Ghosh, Akash, et al.
Veröffentlicht: (2026)
von: Ghosh, Akash, et al.
Veröffentlicht: (2026)
BiMediX2: Bio-Medical EXpert LMM for Diverse Medical Modalities
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2024)
CAMEL-Bench: A Comprehensive Arabic LMM Benchmark
von: Ghaboura, Sara, et al.
Veröffentlicht: (2024)
von: Ghaboura, Sara, et al.
Veröffentlicht: (2024)
Imagine How To Change: Explicit Procedure Modeling for Change Captioning
von: Sun, Jiayang, et al.
Veröffentlicht: (2026)
von: Sun, Jiayang, et al.
Veröffentlicht: (2026)
Context Aware Grounded Teacher for Source Free Object Detection
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
Multi-modal Generation via Cross-Modal In-Context Learning
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
MedROV: Towards Real-Time Open-Vocabulary Detection Across Diverse Medical Imaging Modalities
von: Sheikh, Tooba Tehreem, et al.
Veröffentlicht: (2025)
von: Sheikh, Tooba Tehreem, et al.
Veröffentlicht: (2025)
AgriChain Visually Grounded Expert Verified Reasoning for Interpretable Agricultural Vision Language Models
von: Mahmood, Hazza, et al.
Veröffentlicht: (2026)
von: Mahmood, Hazza, et al.
Veröffentlicht: (2026)
Bilateral Reference for High-Resolution Dichotomous Image Segmentation
von: Zheng, Peng, et al.
Veröffentlicht: (2024)
von: Zheng, Peng, et al.
Veröffentlicht: (2024)
UniFuse: A Unified All-in-One Framework for Multi-Modal Medical Image Fusion Under Diverse Degradations and Misalignments
von: Su, Dayong, et al.
Veröffentlicht: (2025)
von: Su, Dayong, et al.
Veröffentlicht: (2025)
A Benchmark and Agentic Framework for Omni-Modal Reasoning and Tool Use in Long Videos
von: Kurpath, Mohammed Irfan, et al.
Veröffentlicht: (2025)
von: Kurpath, Mohammed Irfan, et al.
Veröffentlicht: (2025)
GroundedSurg: A Multi-Procedure Benchmark for Language-Conditioned Surgical Tool Segmentation
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
DFR: A Decompose-Fuse-Reconstruct Framework for Multi-Modal Few-Shot Segmentation
von: Chen, Shuai, et al.
Veröffentlicht: (2025)
von: Chen, Shuai, et al.
Veröffentlicht: (2025)
QTrack: Query-Driven Reasoning for Multi-modal MOT
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2026)
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
MIRA: Multimodal Iterative Reasoning Agent for Image Editing
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
Fuse4Seg: Image Fusion for Multi-Modal Medical Segmentation via Bi-level Optimization
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
von: Guo, Yuchen, et al.
Veröffentlicht: (2024)
Learning to Fuse and Reconstruct Multi-View Graphs for Diabetic Retinopathy Grading
von: Li, Haoran, et al.
Veröffentlicht: (2026)
von: Li, Haoran, et al.
Veröffentlicht: (2026)
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
MediX-R1: Open Ended Medical Reinforcement Learning
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2026)
von: Mullappilly, Sahal Shaji, et al.
Veröffentlicht: (2026)
Conti-Fuse: A Novel Continuous Decomposition-based Fusion Framework for Infrared and Visible Images
von: Li, Hui, et al.
Veröffentlicht: (2024)
von: Li, Hui, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning
von: Han, Zongyan, et al.
Veröffentlicht: (2025) -
A Fusion-Guided Inception Network for Hyperspectral Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025) -
All in One: Visual-Description-Guided Unified Point Cloud Segmentation
von: Han, Zongyan, et al.
Veröffentlicht: (2025) -
XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
von: Thawakar, Omkar, et al.
Veröffentlicht: (2023) -
Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)