Parameter-Efficient VLMs for Gastrointestinal Endoscopy: Medical Image Generation and Clinical Visual Question Answering
Fuente:
arXiv
Salvato in:
| Autori principali: | Peter, Ojonugwa Oluwafemi Ejiga, Ejiga, Frederick Akor, Khalifa, Fahmi, Rahman, Md Mahmudur |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Advancing AI-Powered Medical Image Synthesis: Insights from MedVQA-GI Challenge Using CLIP, Fine-Tuned Stable Diffusion, and Dream-Booth + LoRA
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
Synthetic Data-Driven Multi-Architecture Framework for Automated Polyp Segmentation Through Integrated Detection and Mask Generation
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
Transformer-Based Explainable Deep Learning for Breast Cancer Detection in Mammography: The MammoFormer Framework
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025)
LiteMedCoT-VL: Parameter-Efficient Adaptation for Medical Visual Question Answering
di: Ma, Runze, et al.
Pubblicazione: (2026)
di: Ma, Runze, et al.
Pubblicazione: (2026)
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
di: Gautam, Sushant, et al.
Pubblicazione: (2025)
di: Gautam, Sushant, et al.
Pubblicazione: (2025)
Visual Robustness Benchmark for Visual Question Answering (VQA)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2024)
Targeted Visual Prompting for Medical Visual Question Answering
di: Tascon-Morales, Sergio, et al.
Pubblicazione: (2024)
di: Tascon-Morales, Sergio, et al.
Pubblicazione: (2024)
ChitroJera: A Regionally Relevant Visual Question Answering Dataset for Bangla
di: Barua, Deeparghya Dutta, et al.
Pubblicazione: (2024)
di: Barua, Deeparghya Dutta, et al.
Pubblicazione: (2024)
How Good LLMs Are at Answering Bangla Medical Visual Questions? Dataset and Benchmarking
di: Ahmed, Rafid, et al.
Pubblicazione: (2026)
di: Ahmed, Rafid, et al.
Pubblicazione: (2026)
TPCL: Task Progressive Curriculum Learning for Robust Visual Question Answering
di: Akl, Ahmed, et al.
Pubblicazione: (2024)
di: Akl, Ahmed, et al.
Pubblicazione: (2024)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
di: Zhang, Xiaoman, et al.
Pubblicazione: (2023)
Elevating Visual Question Answering through Implicitly Learned Reasoning Pathways in LVLMs
di: Jing, Liu, et al.
Pubblicazione: (2025)
di: Jing, Liu, et al.
Pubblicazione: (2025)
Domain-Adaptive Pre-training of Self-Supervised Foundation Models for Medical Image Classification in Gastrointestinal Endoscopy
di: Roth, Marcel, et al.
Pubblicazione: (2024)
di: Roth, Marcel, et al.
Pubblicazione: (2024)
VIHD: Visual Intervention-based Hallucination Detection for Medical Visual Question Answering
di: Chen, Jiayi, et al.
Pubblicazione: (2026)
di: Chen, Jiayi, et al.
Pubblicazione: (2026)
Free Form Medical Visual Question Answering in Radiology
di: Narayanan, Abhishek, et al.
Pubblicazione: (2024)
di: Narayanan, Abhishek, et al.
Pubblicazione: (2024)
Saliency Guided Longitudinal Medical Visual Question Answering
di: Wu, Jialin, et al.
Pubblicazione: (2025)
di: Wu, Jialin, et al.
Pubblicazione: (2025)
Are Large Vision Language Models Truly Grounded in Medical Images? Evidence from Italian Clinical Visual Question Answering
di: Felizzi, Federico, et al.
Pubblicazione: (2025)
di: Felizzi, Federico, et al.
Pubblicazione: (2025)
From Scope to Script: An Automated Report Generation Model for Gastrointestinal Endoscopy
di: Kaklamanos, Evandros, et al.
Pubblicazione: (2025)
di: Kaklamanos, Evandros, et al.
Pubblicazione: (2025)
Hallucination Benchmark in Medical Visual Question Answering
di: Wu, Jinge, et al.
Pubblicazione: (2024)
di: Wu, Jinge, et al.
Pubblicazione: (2024)
Efficient Bilinear Attention-based Fusion for Medical Visual Question Answering
di: Zhang, Zhilin, et al.
Pubblicazione: (2024)
di: Zhang, Zhilin, et al.
Pubblicazione: (2024)
Prompt-based Personalized Federated Learning for Medical Visual Question Answering
di: Zhu, He, et al.
Pubblicazione: (2024)
di: Zhu, He, et al.
Pubblicazione: (2024)
Structure Causal Models and LLMs Integration in Medical Visual Question Answering
di: Xu, Zibo, et al.
Pubblicazione: (2025)
di: Xu, Zibo, et al.
Pubblicazione: (2025)
TM-PATHVQA:90000+ Textless Multilingual Questions for Medical Visual Question Answering
di: Rajkhowa, Tonmoy, et al.
Pubblicazione: (2024)
di: Rajkhowa, Tonmoy, et al.
Pubblicazione: (2024)
A Comprehensive Survey on Visual Question Answering Datasets and Algorithms
di: Kabir, Raihan, et al.
Pubblicazione: (2024)
di: Kabir, Raihan, et al.
Pubblicazione: (2024)
InViC: Intent-aware Visual Cues for Medical Visual Question Answering
di: Wang, Zhisong, et al.
Pubblicazione: (2026)
di: Wang, Zhisong, et al.
Pubblicazione: (2026)
Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering
di: Hagen, Luca, et al.
Pubblicazione: (2026)
di: Hagen, Luca, et al.
Pubblicazione: (2026)
Location-Aware Pretraining for Medical Difference Visual Question Answering
di: Musinguzi, Denis, et al.
Pubblicazione: (2026)
di: Musinguzi, Denis, et al.
Pubblicazione: (2026)
From Image to Language: A Critical Analysis of Visual Question Answering (VQA) Approaches, Challenges, and Opportunities
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2023)
di: Ishmam, Md Farhan, et al.
Pubblicazione: (2023)
Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review
di: Hartsock, Iryna, et al.
Pubblicazione: (2024)
di: Hartsock, Iryna, et al.
Pubblicazione: (2024)
Enhancing Generalization in Medical Visual Question Answering Tasks via Gradient-Guided Model Perturbation
di: Liu, Gang, et al.
Pubblicazione: (2024)
di: Liu, Gang, et al.
Pubblicazione: (2024)
MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering
di: Xi, Suyang, et al.
Pubblicazione: (2026)
di: Xi, Suyang, et al.
Pubblicazione: (2026)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
di: Nguyen, Hai-Dang, et al.
Pubblicazione: (2025)
di: Nguyen, Hai-Dang, et al.
Pubblicazione: (2025)
Q-FSRU: Quantum-Augmented Frequency-Spectral For Medical Visual Question Answering
di: Thakur, Rakesh, et al.
Pubblicazione: (2025)
di: Thakur, Rakesh, et al.
Pubblicazione: (2025)
V-Loop: Visual Logical Loop Verification for Hallucination Detection in Medical Visual Question Answering
di: Jin, Mengyuan, et al.
Pubblicazione: (2026)
di: Jin, Mengyuan, et al.
Pubblicazione: (2026)
Visual Question Answering on Multiple Remote Sensing Image Modalities
di: Boussaid, Hichem, et al.
Pubblicazione: (2025)
di: Boussaid, Hichem, et al.
Pubblicazione: (2025)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
di: Ahir, Param, et al.
Pubblicazione: (2023)
di: Ahir, Param, et al.
Pubblicazione: (2023)
QIRL: Boosting Visual Question Answering via Optimized Question-Image Relation Learning
di: Xu, Quanxing, et al.
Pubblicazione: (2025)
di: Xu, Quanxing, et al.
Pubblicazione: (2025)
Questioning the Stability of Visual Question Answering
di: Rosenfeld, Amir, et al.
Pubblicazione: (2025)
di: Rosenfeld, Amir, et al.
Pubblicazione: (2025)
Hierarchical Modeling for Medical Visual Question Answering with Cross-Attention Fusion
di: Zhang, Junkai, et al.
Pubblicazione: (2025)
di: Zhang, Junkai, et al.
Pubblicazione: (2025)
Are Vision Language Models Ready for Clinical Diagnosis? A 3D Medical Benchmark for Tumor-centric Visual Question Answering
di: Chen, Yixiong, et al.
Pubblicazione: (2025)
di: Chen, Yixiong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Advancing AI-Powered Medical Image Synthesis: Insights from MedVQA-GI Challenge Using CLIP, Fine-Tuned Stable Diffusion, and Dream-Booth + LoRA
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025) -
Synthetic Data-Driven Multi-Architecture Framework for Automated Polyp Segmentation Through Integrated Detection and Mask Generation
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025) -
Transformer-Based Explainable Deep Learning for Breast Cancer Detection in Mammography: The MammoFormer Framework
di: Peter, Ojonugwa Oluwafemi Ejiga, et al.
Pubblicazione: (2025) -
LiteMedCoT-VL: Parameter-Efficient Adaptation for Medical Visual Question Answering
di: Ma, Runze, et al.
Pubblicazione: (2026) -
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
di: Gautam, Sushant, et al.
Pubblicazione: (2025)