OrthoDoc: Multimodal Large Language Model for Assisting Diagnosis in Computed Tomography
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Youzhu, Zhang, Yichen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OrthoInsight: Rib Fracture Diagnosis and Report Generation Based on Multi-Modal Large Models
von: Wu, Ningyong, et al.
Veröffentlicht: (2025)
von: Wu, Ningyong, et al.
Veröffentlicht: (2025)
Deep-Learning-Assisted Highly-Accurate COVID-19 Diagnosis on Lung Computed Tomography Images
von: Wang, Yinuo, et al.
Veröffentlicht: (2025)
von: Wang, Yinuo, et al.
Veröffentlicht: (2025)
Enhancing Low Dose Computed Tomography Images Using Consistency Training Techniques
von: Gokmen, Mahmut S., et al.
Veröffentlicht: (2024)
von: Gokmen, Mahmut S., et al.
Veröffentlicht: (2024)
DM4CT: Benchmarking Diffusion Models for Computed Tomography Reconstruction
von: Shi, Jiayang, et al.
Veröffentlicht: (2026)
von: Shi, Jiayang, et al.
Veröffentlicht: (2026)
Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts
von: Ge, Peixuan, et al.
Veröffentlicht: (2025)
von: Ge, Peixuan, et al.
Veröffentlicht: (2025)
FreeTumor: Large-Scale Generative Tumor Synthesis in Computed Tomography Images for Improving Tumor Recognition
von: Wu, Linshan, et al.
Veröffentlicht: (2025)
von: Wu, Linshan, et al.
Veröffentlicht: (2025)
A Versatile Pathology Co-pilot via Reasoning Enhanced Multimodal Large Language Model
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
von: Xu, Zhe, et al.
Veröffentlicht: (2025)
Potential of Multimodal Large Language Models for Data Mining of Medical Images and Free-text Reports
von: Zhang, Yutong, et al.
Veröffentlicht: (2024)
von: Zhang, Yutong, et al.
Veröffentlicht: (2024)
Brain-Adapter: Enhancing Neurological Disorder Analysis with Adapter-Tuning Multimodal Large Language Models
von: Zhang, Jing, et al.
Veröffentlicht: (2025)
von: Zhang, Jing, et al.
Veröffentlicht: (2025)
TCM-Tongue: A Standardized Tongue Image Dataset with Pathological Annotations for AI-Assisted TCM Diagnosis
von: Jin, Xuebo, et al.
Veröffentlicht: (2025)
von: Jin, Xuebo, et al.
Veröffentlicht: (2025)
Geospatial-Temporal Sensemaking of Remote Sensing Activity Detections with Multimodal Large Language Model
von: Ramirez, David F., et al.
Veröffentlicht: (2026)
von: Ramirez, David F., et al.
Veröffentlicht: (2026)
SCC-YOLO: An Improved Object Detector for Assisting in Brain Tumor Diagnosis
von: Bai, Runci, et al.
Veröffentlicht: (2025)
von: Bai, Runci, et al.
Veröffentlicht: (2025)
Evaluating the Diagnostic Classification Ability of Multimodal Large Language Models: Insights from the Osteoarthritis Initiative
von: Wang, Li, et al.
Veröffentlicht: (2026)
von: Wang, Li, et al.
Veröffentlicht: (2026)
A Hybrid Deep Learning CNN Model for Enhanced COVID-19 Detection from Computed Tomography (CT) Scan Images
von: Nettur, Suresh Babu, et al.
Veröffentlicht: (2025)
von: Nettur, Suresh Babu, et al.
Veröffentlicht: (2025)
MISC: Ultra-low Bitrate Image Semantic Compression Driven by Large Multimodal Model
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
von: Li, Chunyi, et al.
Veröffentlicht: (2024)
Automated Segmentation of Ischemic Stroke Lesions in Non-Contrast Computed Tomography Images for Enhanced Treatment and Prognosis
von: Musah, Toufiq, et al.
Veröffentlicht: (2024)
von: Musah, Toufiq, et al.
Veröffentlicht: (2024)
$ρ$-NeRF: Leveraging Attenuation Priors in Neural Radiance Field for 3D Computed Tomography Reconstruction
von: Zhou, Li, et al.
Veröffentlicht: (2024)
von: Zhou, Li, et al.
Veröffentlicht: (2024)
MedMimic: Physician-Inspired Multimodal Fusion for Early Diagnosis of Fever of Unknown Origin
von: Chen, Minrui, et al.
Veröffentlicht: (2025)
von: Chen, Minrui, et al.
Veröffentlicht: (2025)
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
Efficient Automated Diagnosis of Retinopathy of Prematurity by Customize CNN Models
von: Saeedi, Farzan, et al.
Veröffentlicht: (2025)
von: Saeedi, Farzan, et al.
Veröffentlicht: (2025)
Multimodal Fusion and Coherence Modeling for Video Topic Segmentation
von: Yu, Hai, et al.
Veröffentlicht: (2024)
von: Yu, Hai, et al.
Veröffentlicht: (2024)
Can Large Language Models Challenge CNNs in Medical Image Analysis?
von: Ahmed, Shibbir, et al.
Veröffentlicht: (2025)
von: Ahmed, Shibbir, et al.
Veröffentlicht: (2025)
Review and Recommendations for using Artificial Intelligence in Intracoronary Optical Coherence Tomography Analysis
von: Chen, Xu, et al.
Veröffentlicht: (2025)
von: Chen, Xu, et al.
Veröffentlicht: (2025)
MindVL: Towards Efficient and Effective Training of Multimodal Large Language Models on Ascend NPUs
von: Chen, Feilong, et al.
Veröffentlicht: (2025)
von: Chen, Feilong, et al.
Veröffentlicht: (2025)
Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2025)
von: Das, Anindya Bijoy, et al.
Veröffentlicht: (2025)
SurgWound-Bench: A Benchmark for Surgical Wound Diagnosis
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
von: Xu, Jiahao, et al.
Veröffentlicht: (2025)
Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition
von: Ramirez, David F., et al.
Veröffentlicht: (2026)
von: Ramirez, David F., et al.
Veröffentlicht: (2026)
Can Generalist Vision Language Models (VLMs) Rival Specialist Medical VLMs? Benchmarking and Strategic Insights
von: Zhong, Yuan, et al.
Veröffentlicht: (2025)
von: Zhong, Yuan, et al.
Veröffentlicht: (2025)
Mammo-CLIP: Leveraging Contrastive Language-Image Pre-training (CLIP) for Enhanced Breast Cancer Diagnosis with Multi-view Mammography
von: Chen, Xuxin, et al.
Veröffentlicht: (2024)
von: Chen, Xuxin, et al.
Veröffentlicht: (2024)
Dual Thinking and Logical Processing -- Are Multi-modal Large Language Models Closing the Gap with Human Vision ?
von: Dayanandan, Kailas, et al.
Veröffentlicht: (2024)
von: Dayanandan, Kailas, et al.
Veröffentlicht: (2024)
Beyond Textual Knowledge-Leveraging Multimodal Knowledge Bases for Enhancing Vision-and-Language Navigation
von: Yang, Dongsheng, et al.
Veröffentlicht: (2026)
von: Yang, Dongsheng, et al.
Veröffentlicht: (2026)
LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models
von: Bhatnagar, Shubhang, et al.
Veröffentlicht: (2025)
von: Bhatnagar, Shubhang, et al.
Veröffentlicht: (2025)
Comprehensive Evaluation of Multimodal AI Models in Medical Imaging Diagnosis: From Data Augmentation to Preference-Based Comparison
von: Ruan, Cailian, et al.
Veröffentlicht: (2024)
von: Ruan, Cailian, et al.
Veröffentlicht: (2024)
Multimodal Medical Image Binding via Shared Text Embeddings
von: Liu, Yunhao, et al.
Veröffentlicht: (2025)
von: Liu, Yunhao, et al.
Veröffentlicht: (2025)
Applying Conditional Generative Adversarial Networks for Imaging Diagnosis
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
Teach Multimodal LLMs to Comprehend Electrocardiographic Images
von: Liu, Ruoqi, et al.
Veröffentlicht: (2024)
von: Liu, Ruoqi, et al.
Veröffentlicht: (2024)
A Disease-Centric Vision-Language Foundation Model for Precision Oncology in Kidney Cancer
von: Tao, Yuhui, et al.
Veröffentlicht: (2025)
von: Tao, Yuhui, et al.
Veröffentlicht: (2025)
HA-HI: Synergising fMRI and DTI through Hierarchical Alignments and Hierarchical Interactions for Mild Cognitive Impairment Diagnosis
von: Shen, Xiongri, et al.
Veröffentlicht: (2024)
von: Shen, Xiongri, et al.
Veröffentlicht: (2024)
Deep Learning for Fetal Inflammatory Response Diagnosis in the Umbilical Cord
von: Ayad, Marina A., et al.
Veröffentlicht: (2024)
von: Ayad, Marina A., et al.
Veröffentlicht: (2024)
Survey of AI-Powered Approaches for Osteoporosis Diagnosis in Medical Imaging
von: Rahman, Abdul, et al.
Veröffentlicht: (2025)
von: Rahman, Abdul, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OrthoInsight: Rib Fracture Diagnosis and Report Generation Based on Multi-Modal Large Models
von: Wu, Ningyong, et al.
Veröffentlicht: (2025) -
Deep-Learning-Assisted Highly-Accurate COVID-19 Diagnosis on Lung Computed Tomography Images
von: Wang, Yinuo, et al.
Veröffentlicht: (2025) -
Enhancing Low Dose Computed Tomography Images Using Consistency Training Techniques
von: Gokmen, Mahmut S., et al.
Veröffentlicht: (2024) -
DM4CT: Benchmarking Diffusion Models for Computed Tomography Reconstruction
von: Shi, Jiayang, et al.
Veröffentlicht: (2026) -
Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts
von: Ge, Peixuan, et al.
Veröffentlicht: (2025)