Incentivizing Cardiologist-Like Reasoning in MLLMs for Interpretable Echocardiographic Diagnosis
Fuente:
arXiv
Salvato in:
| Autori principali: | Qin, Yi, Wang, Lehan, Zhao, Chenxu, Lee, Alex P. W., Li, Xiaomeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Proactive Reasoning-with-Retrieval Framework for Medical Multimodal Large Language Models
di: Wang, Lehan, et al.
Pubblicazione: (2025)
di: Wang, Lehan, et al.
Pubblicazione: (2025)
Neurons: Emulating the Human Visual Cortex Improves Fidelity and Interpretability in fMRI-to-Video Reconstruction
di: Wang, Haonan, et al.
Pubblicazione: (2025)
di: Wang, Haonan, et al.
Pubblicazione: (2025)
Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks
di: Wang, Lehan, et al.
Pubblicazione: (2024)
di: Wang, Lehan, et al.
Pubblicazione: (2024)
VideoRFT: Incentivizing Video Reasoning Capability in MLLMs via Reinforced Fine-Tuning
di: Wang, Qi, et al.
Pubblicazione: (2025)
di: Wang, Qi, et al.
Pubblicazione: (2025)
VAMPIRE: Uncovering Vessel Directional and Morphological Information from OCTA Images for Cardiovascular Disease Risk Factor Prediction
di: Wang, Lehan, et al.
Pubblicazione: (2025)
di: Wang, Lehan, et al.
Pubblicazione: (2025)
HumanSense: From Multimodal Perception to Empathetic Context-Aware Responses through Reasoning MLLMs
di: Qin, Zheng, et al.
Pubblicazione: (2025)
di: Qin, Zheng, et al.
Pubblicazione: (2025)
Evidence-Based Actor-Verifier Reasoning for Echocardiographic Agents
di: Huang, Peng, et al.
Pubblicazione: (2026)
di: Huang, Peng, et al.
Pubblicazione: (2026)
Multi-Modal Explainable Medical AI Assistant for Trustworthy Human-AI Collaboration
di: Yang, Honglong, et al.
Pubblicazione: (2025)
di: Yang, Honglong, et al.
Pubblicazione: (2025)
MultiEYE: Dataset and Benchmark for OCT-Enhanced Retinal Disease Recognition from Fundus Images
di: Wang, Lehan, et al.
Pubblicazione: (2024)
di: Wang, Lehan, et al.
Pubblicazione: (2024)
GeoZero: Incentivizing Reasoning from Scratch on Geospatial Scenes
di: Wang, Di, et al.
Pubblicazione: (2025)
di: Wang, Di, et al.
Pubblicazione: (2025)
PathMR: Multimodal Visual Reasoning for Interpretable Pathology Diagnosis
di: Zhang, Ye, et al.
Pubblicazione: (2025)
di: Zhang, Ye, et al.
Pubblicazione: (2025)
DAFTED: Decoupled Asymmetric Fusion of Tabular and Echocardiographic Data for Cardiac Hypertension Diagnosis
di: Stym-Popper, Jérémie, et al.
Pubblicazione: (2025)
di: Stym-Popper, Jérémie, et al.
Pubblicazione: (2025)
GTPred: Benchmarking MLLMs for Interpretable Geo-localization and Time-of-capture Prediction
di: Li, Jinnao, et al.
Pubblicazione: (2026)
di: Li, Jinnao, et al.
Pubblicazione: (2026)
Can MLLMs Reason About Visual Persuasion? Evaluating the Efficacy and Faithfulness of Reasoning
di: Lee, Naeun, et al.
Pubblicazione: (2026)
di: Lee, Naeun, et al.
Pubblicazione: (2026)
Video-R1: Reinforcing Video Reasoning in MLLMs
di: Feng, Kaituo, et al.
Pubblicazione: (2025)
di: Feng, Kaituo, et al.
Pubblicazione: (2025)
SpaceR: Reinforcing MLLMs in Video Spatial Reasoning
di: Ouyang, Kun, et al.
Pubblicazione: (2025)
di: Ouyang, Kun, et al.
Pubblicazione: (2025)
HiLM-D: Enhancing MLLMs with Multi-Scale High-Resolution Details for Autonomous Driving
di: Ding, Xinpeng, et al.
Pubblicazione: (2023)
di: Ding, Xinpeng, et al.
Pubblicazione: (2023)
Interpreting and Mitigating Hallucination in MLLMs through Multi-agent Debate
di: Lin, Zheng, et al.
Pubblicazione: (2024)
di: Lin, Zheng, et al.
Pubblicazione: (2024)
VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?
di: Liu, Yuanxin, et al.
Pubblicazione: (2025)
di: Liu, Yuanxin, et al.
Pubblicazione: (2025)
POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs
di: Wang, Haicheng, et al.
Pubblicazione: (2026)
di: Wang, Haicheng, et al.
Pubblicazione: (2026)
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs
di: Zhao, Jiahe, et al.
Pubblicazione: (2025)
di: Zhao, Jiahe, et al.
Pubblicazione: (2025)
Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs
di: Tong, Jintao, et al.
Pubblicazione: (2025)
di: Tong, Jintao, et al.
Pubblicazione: (2025)
Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression
di: Du, Yao, et al.
Pubblicazione: (2026)
di: Du, Yao, et al.
Pubblicazione: (2026)
Training-Free Reasoning and Reflection in MLLMs
di: Wei, Hongchen, et al.
Pubblicazione: (2025)
di: Wei, Hongchen, et al.
Pubblicazione: (2025)
VITAL: Visual-Semantic Dual Supervision for Enhanced and Interpretable Latent Reasoning in Medical MLLMs
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
di: Li, Qiaoru, et al.
Pubblicazione: (2026)
Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs
di: Ji, Yikun, et al.
Pubblicazione: (2025)
di: Ji, Yikun, et al.
Pubblicazione: (2025)
ContrastDiagnosis: Enhancing Interpretability in Lung Nodule Diagnosis Using Contrastive Learning
di: Wang, Chenglong, et al.
Pubblicazione: (2024)
di: Wang, Chenglong, et al.
Pubblicazione: (2024)
Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation
di: Han, Su Ho, et al.
Pubblicazione: (2025)
di: Han, Su Ho, et al.
Pubblicazione: (2025)
Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretations
di: Xu, Xinyue, et al.
Pubblicazione: (2024)
di: Xu, Xinyue, et al.
Pubblicazione: (2024)
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
di: Lai, Haoran, et al.
Pubblicazione: (2026)
di: Lai, Haoran, et al.
Pubblicazione: (2026)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
di: Wang, Haozhe, et al.
Pubblicazione: (2025)
di: Wang, Haozhe, et al.
Pubblicazione: (2025)
VER-Bench: Evaluating MLLMs on Reasoning with Fine-Grained Visual Evidence
di: Qiang, Chenhui, et al.
Pubblicazione: (2025)
di: Qiang, Chenhui, et al.
Pubblicazione: (2025)
Open Eyes, Then Reason: Fine-grained Visual Mathematical Understanding in MLLMs
di: Zhang, Shan, et al.
Pubblicazione: (2025)
di: Zhang, Shan, et al.
Pubblicazione: (2025)
AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs
di: Chang, Boyu, et al.
Pubblicazione: (2026)
di: Chang, Boyu, et al.
Pubblicazione: (2026)
Learning to Tune Like an Expert: Interpretable and Scene-Aware Navigation via MLLM Reasoning and CVAE-Based Adaptation
di: Wang, Yanbo, et al.
Pubblicazione: (2025)
di: Wang, Yanbo, et al.
Pubblicazione: (2025)
From Indoor to Open World: Revealing the Spatial Reasoning Gap in MLLMs
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
di: Wu, Mingrui, et al.
Pubblicazione: (2025)
Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios
di: Li, Xiaomin, et al.
Pubblicazione: (2026)
di: Li, Xiaomin, et al.
Pubblicazione: (2026)
EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs
di: Chen, Zhenghao, et al.
Pubblicazione: (2026)
di: Chen, Zhenghao, et al.
Pubblicazione: (2026)
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
di: Hao, Yunzhuo, et al.
Pubblicazione: (2025)
di: Hao, Yunzhuo, et al.
Pubblicazione: (2025)
Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning
di: Zhang, Bob, et al.
Pubblicazione: (2025)
di: Zhang, Bob, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Proactive Reasoning-with-Retrieval Framework for Medical Multimodal Large Language Models
di: Wang, Lehan, et al.
Pubblicazione: (2025) -
Neurons: Emulating the Human Visual Cortex Improves Fidelity and Interpretability in fMRI-to-Video Reconstruction
di: Wang, Haonan, et al.
Pubblicazione: (2025) -
Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks
di: Wang, Lehan, et al.
Pubblicazione: (2024) -
VideoRFT: Incentivizing Video Reasoning Capability in MLLMs via Reinforced Fine-Tuning
di: Wang, Qi, et al.
Pubblicazione: (2025) -
VAMPIRE: Uncovering Vessel Directional and Morphological Information from OCTA Images for Cardiovascular Disease Risk Factor Prediction
di: Wang, Lehan, et al.
Pubblicazione: (2025)