EchoVLM: Measurement-Grounded Multimodal Learning for Echocardiography
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yuheng, Zhang, Yue, Amadou, Abdoul Aziz, Lai, Yuxiang, Zhong, Jike, Passerini, Tiziano, Comaniciu, Dorin, Sharma, Puneet |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EchoApex: A General-Purpose Vision Foundation Model for Echocardiography
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence
por: She, Chaoyin, et al.
Publicado: (2025)
por: She, Chaoyin, et al.
Publicado: (2025)
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?
por: Lai, Yuxiang, et al.
Publicado: (2025)
por: Lai, Yuxiang, et al.
Publicado: (2025)
ConceptVAE: Self-Supervised Fine-Grained Concept Disentanglement from 2D Echocardiographies
por: Ciusdel, Costin F., et al.
Publicado: (2025)
por: Ciusdel, Costin F., et al.
Publicado: (2025)
Context Matters: Learning Global Semantics via Object-Centric Representation
por: Zhong, Jike, et al.
Publicado: (2025)
por: Zhong, Jike, et al.
Publicado: (2025)
EEE-Bench: A Comprehensive Multimodal Electrical And Electronics Engineering Benchmark
por: Li, Ming, et al.
Publicado: (2024)
por: Li, Ming, et al.
Publicado: (2024)
MedVista3D: Vision-Language Modeling for Reducing Diagnostic Errors in 3D CT Disease Detection, Understanding and Reporting
por: Li, Yuheng, et al.
Publicado: (2025)
por: Li, Yuheng, et al.
Publicado: (2025)
Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models
por: Lai, Yuxiang, et al.
Publicado: (2025)
por: Lai, Yuxiang, et al.
Publicado: (2025)
Towards a vision foundation model for comprehensive assessment of Cardiac MRI
por: Jacob, Athira J, et al.
Publicado: (2024)
por: Jacob, Athira J, et al.
Publicado: (2024)
Self-Supervised Learning for Interventional Image Analytics: Towards Robust Device Trackers
por: Islam, Saahil, et al.
Publicado: (2024)
por: Islam, Saahil, et al.
Publicado: (2024)
Patient-Specific Autoregressive Models for Organ Motion Prediction in Radiotherapy
por: Lai, Yuxiang, et al.
Publicado: (2025)
por: Lai, Yuxiang, et al.
Publicado: (2025)
A Novel Tracking Framework for Devices in X-ray Leveraging Supplementary Cue-Driven Self-Supervised Features
por: Islam, Saahil, et al.
Publicado: (2025)
por: Islam, Saahil, et al.
Publicado: (2025)
EchoWorld: Learning Motion-Aware World Models for Echocardiography Probe Guidance
por: Yue, Yang, et al.
Publicado: (2025)
por: Yue, Yang, et al.
Publicado: (2025)
EchoAgent: Guideline-Centric Reasoning Agent for Echocardiography Measurement and Interpretation
por: Daghyani, Matin, et al.
Publicado: (2025)
por: Daghyani, Matin, et al.
Publicado: (2025)
EchoTracker: Advancing Myocardial Point Tracking in Echocardiography
por: Azad, Md Abulkalam, et al.
Publicado: (2024)
por: Azad, Md Abulkalam, et al.
Publicado: (2024)
EchoAgent: Towards Reliable Echocardiography Interpretation with "Eyes","Hands" and "Minds"
por: Wang, Qin, et al.
Publicado: (2026)
por: Wang, Qin, et al.
Publicado: (2026)
Think or Not Think: A Study of Explicit Thinking in Rule-Based Visual Reinforcement Fine-Tuning
por: Li, Ming, et al.
Publicado: (2025)
por: Li, Ming, et al.
Publicado: (2025)
EchoXFlow: A Beamspace Echocardiography Dataset for Cardiac Motion, Flow, and Function
por: Stenhede, Elias, et al.
Publicado: (2026)
por: Stenhede, Elias, et al.
Publicado: (2026)
EchoJEPA: A Latent Predictive Foundation Model for Echocardiography
por: Munim, Alif, et al.
Publicado: (2026)
por: Munim, Alif, et al.
Publicado: (2026)
Goal-conditioned reinforcement learning for ultrasound navigation guidance
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
CoReEcho: Continuous Representation Learning for 2D+time Echocardiography Analysis
por: Maani, Fadillah Adamsyah, et al.
Publicado: (2024)
por: Maani, Fadillah Adamsyah, et al.
Publicado: (2024)
Concepts Worth Having: Refining VLM-Guided Concept Bottleneck Models with Minimal Annotations
por: Debole, Nicola, et al.
Publicado: (2026)
por: Debole, Nicola, et al.
Publicado: (2026)
TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning
por: Li, Ming, et al.
Publicado: (2025)
por: Li, Ming, et al.
Publicado: (2025)
Learning to Optimize Radiotherapy Plans via Fluence Maps Diffusion Model Generation and LSTM-based Optimization
por: Poles, Isabella, et al.
Publicado: (2026)
por: Poles, Isabella, et al.
Publicado: (2026)
MedDINOv3: How to adapt vision foundation models for medical image segmentation?
por: Li, Yuheng, et al.
Publicado: (2025)
por: Li, Yuheng, et al.
Publicado: (2025)
VRIQ: Benchmarking and Analyzing Visual-Reasoning IQ of VLMs
por: Khezresmaeilzadeh, Tina, et al.
Publicado: (2026)
por: Khezresmaeilzadeh, Tina, et al.
Publicado: (2026)
Self Pre-training with Adaptive Mask Autoencoders for Variable-Contrast 3D Medical Imaging
por: Das, Badhan Kumar, et al.
Publicado: (2025)
por: Das, Badhan Kumar, et al.
Publicado: (2025)
AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities
por: Das, Badhan Kumar, et al.
Publicado: (2025)
por: Das, Badhan Kumar, et al.
Publicado: (2025)
Multi-Plane Vision Transformer for Hemorrhage Classification Using Axial and Sagittal MRI Data
por: Das, Badhan Kumar, et al.
Publicado: (2025)
por: Das, Badhan Kumar, et al.
Publicado: (2025)
Echo4DIR: 4D Implicit Heart Reconstruction from 2D Echocardiography Videos
por: Liu, Yanan, et al.
Publicado: (2026)
por: Liu, Yanan, et al.
Publicado: (2026)
EchoPrime: A Multi-Video View-Informed Vision-Language Model for Comprehensive Echocardiography Interpretation
por: Vukadinovic, Milos, et al.
Publicado: (2024)
por: Vukadinovic, Milos, et al.
Publicado: (2024)
MObI: Multimodal Object Inpainting Using Diffusion Models
por: Buburuzan, Alexandru, et al.
Publicado: (2025)
por: Buburuzan, Alexandru, et al.
Publicado: (2025)
ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation
por: Peng, Cihang, et al.
Publicado: (2025)
por: Peng, Cihang, et al.
Publicado: (2025)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
por: Xu, Runsen, et al.
Publicado: (2024)
por: Xu, Runsen, et al.
Publicado: (2024)
VIViT: Variable-Input Vision Transformer Framework for 3D MR Image Segmentation
por: Das, Badhan Kumar, et al.
Publicado: (2025)
por: Das, Badhan Kumar, et al.
Publicado: (2025)
Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification
por: Liu, Han, et al.
Publicado: (2025)
por: Liu, Han, et al.
Publicado: (2025)
City-VLM: Towards Multidomain Perception Scene Understanding via Multimodal Incomplete Learning
por: Sun, Penglei, et al.
Publicado: (2025)
por: Sun, Penglei, et al.
Publicado: (2025)
Echo-CoPilot: A Multiple-Perspective Agentic Framework for Reliable Echocardiography Interpretation
por: Heidari, Moein, et al.
Publicado: (2025)
por: Heidari, Moein, et al.
Publicado: (2025)
Fake It Till You Make It: Using Synthetic Data and Domain Knowledge for Improved Text-Based Learning for LGE Detection
por: Jacob, Athira J, et al.
Publicado: (2025)
por: Jacob, Athira J, et al.
Publicado: (2025)
Demo: Generative AI helps Radiotherapy Planning with User Preference
por: Gao, Riqiang, et al.
Publicado: (2025)
por: Gao, Riqiang, et al.
Publicado: (2025)
Ejemplares similares
-
EchoApex: A General-Purpose Vision Foundation Model for Echocardiography
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024) -
EchoVLM: Dynamic Mixture-of-Experts Vision-Language Model for Universal Ultrasound Intelligence
por: She, Chaoyin, et al.
Publicado: (2025) -
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?
por: Lai, Yuxiang, et al.
Publicado: (2025) -
ConceptVAE: Self-Supervised Fine-Grained Concept Disentanglement from 2D Echocardiographies
por: Ciusdel, Costin F., et al.
Publicado: (2025) -
Context Matters: Learning Global Semantics via Object-Centric Representation
por: Zhong, Jike, et al.
Publicado: (2025)