Improving Medical VQA through Trajectory-Aware Process Supervision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gulluk, Halil Ibrahim, Gevaert, Olivier |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
SemEnrich: Self-Supervised Semantic Enrichment of Radiology Reports for Vision-Language Learning
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation
von: Byun, Ji Young, et al.
Veröffentlicht: (2026)
von: Byun, Ji Young, et al.
Veröffentlicht: (2026)
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
Bridging the Semantic Gaps: Improving Medical VQA Consistency with LLM-Augmented Question Sets
von: Ma, Yongpei, et al.
Veröffentlicht: (2025)
von: Ma, Yongpei, et al.
Veröffentlicht: (2025)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
von: Wang, Jiayun, et al.
Veröffentlicht: (2024)
von: Wang, Jiayun, et al.
Veröffentlicht: (2024)
CAD: Confidence-Aware Adaptive Displacement for Semi-Supervised Medical Image Segmentation
von: Xiao, Wenbo, et al.
Veröffentlicht: (2025)
von: Xiao, Wenbo, et al.
Veröffentlicht: (2025)
Enhancing Vietnamese VQA through Curriculum Learning on Raw and Augmented Text Representations
von: Nguyen, Khoi Anh, et al.
Veröffentlicht: (2025)
von: Nguyen, Khoi Anh, et al.
Veröffentlicht: (2025)
Multimodal Machine Learning in Image-Based and Clinical Biomedicine: Survey and Prospects
von: Warner, Elisa, et al.
Veröffentlicht: (2023)
von: Warner, Elisa, et al.
Veröffentlicht: (2023)
SITUATE: Indoor Human Trajectory Prediction through Geometric Features and Self-Supervised Vision Representation
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2024)
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2024)
WildFireVQA: A Large-Scale Radiometric Thermal VQA Benchmark for Aerial Wildfire Monitoring
von: Habibpour, Mobin, et al.
Veröffentlicht: (2026)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2026)
CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space
von: Ding, Tianxingjian, et al.
Veröffentlicht: (2025)
von: Ding, Tianxingjian, et al.
Veröffentlicht: (2025)
Disentanglement-Based Equivariant Learning for Compositional VQA
von: Du, Zhou, et al.
Veröffentlicht: (2026)
von: Du, Zhou, et al.
Veröffentlicht: (2026)
Unexplored flaws in multiple-choice VQA evaluations
von: Rosenthal, Fabio, et al.
Veröffentlicht: (2025)
von: Rosenthal, Fabio, et al.
Veröffentlicht: (2025)
BERT-VQA: Visual Question Answering on Plots
von: Vu, Tai, et al.
Veröffentlicht: (2025)
von: Vu, Tai, et al.
Veröffentlicht: (2025)
RNNs, CNNs and Transformers in Human Action Recognition: A Survey and a Hybrid Model
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
VQA-Levels: A Hierarchical Approach for Classifying Questions in VQA
von: Madaka, Madhuri Latha, et al.
Veröffentlicht: (2025)
von: Madaka, Madhuri Latha, et al.
Veröffentlicht: (2025)
Synergy-Guided Regional Supervision of Pseudo Labels for Semi-Supervised Medical Image Segmentation
von: Wang, Tao, et al.
Veröffentlicht: (2024)
von: Wang, Tao, et al.
Veröffentlicht: (2024)
Refine and Align: Confidence Calibration through Multi-Agent Interaction in VQA
von: Pandey, Ayush, et al.
Veröffentlicht: (2025)
von: Pandey, Ayush, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
Improving Colorectal Cancer Screening and Risk Assessment through Predictive Modeling on Medical Images and Records
von: Jiang, Shuai, et al.
Veröffentlicht: (2024)
von: Jiang, Shuai, et al.
Veröffentlicht: (2024)
Supervised Anomaly Detection for Complex Industrial Images
von: Baitieva, Aimira, et al.
Veröffentlicht: (2024)
von: Baitieva, Aimira, et al.
Veröffentlicht: (2024)
Improving Automatic VQA Evaluation Using Large Language Models
von: Mañas, Oscar, et al.
Veröffentlicht: (2023)
von: Mañas, Oscar, et al.
Veröffentlicht: (2023)
Modality-Aware Infrared and Visible Image Fusion with Target-Aware Supervision
von: Sun, Tianyao, et al.
Veröffentlicht: (2025)
von: Sun, Tianyao, et al.
Veröffentlicht: (2025)
Learning Velocity and Acceleration: Self-Supervised Motion Consistency for Pedestrian Trajectory Prediction
von: Huang, Yizhou, et al.
Veröffentlicht: (2025)
von: Huang, Yizhou, et al.
Veröffentlicht: (2025)
FinePseudo: Improving Pseudo-Labelling through Temporal-Alignablity for Semi-Supervised Fine-Grained Action Recognition
von: Dave, Ishan Rajendrakumar, et al.
Veröffentlicht: (2024)
von: Dave, Ishan Rajendrakumar, et al.
Veröffentlicht: (2024)
HAMMR: HierArchical MultiModal React agents for generic VQA
von: Castrejon, Lluis, et al.
Veröffentlicht: (2024)
von: Castrejon, Lluis, et al.
Veröffentlicht: (2024)
AI-Derived Structural Building Intelligence for Urban Resilience: An Application in Saint Vincent and the Grenadines
von: Tingzon, Isabelle, et al.
Veröffentlicht: (2025)
von: Tingzon, Isabelle, et al.
Veröffentlicht: (2025)
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation
von: Rand, Amit, et al.
Veröffentlicht: (2025)
von: Rand, Amit, et al.
Veröffentlicht: (2025)
Attribute Diversity Determines the Systematicity Gap in VQA
von: Berlot-Attwell, Ian, et al.
Veröffentlicht: (2023)
von: Berlot-Attwell, Ian, et al.
Veröffentlicht: (2023)
WorldVQA: Measuring Atomic World Knowledge in Multimodal Large Language Models
von: Zhou, Runjie, et al.
Veröffentlicht: (2026)
von: Zhou, Runjie, et al.
Veröffentlicht: (2026)
Evaluating Feature Attribution Methods in the Image Domain
von: Gevaert, Arne, et al.
Veröffentlicht: (2022)
von: Gevaert, Arne, et al.
Veröffentlicht: (2022)
Capturing Context-Aware Route Choice Semantics for Trajectory Representation Learning
von: Cao, Ji, et al.
Veröffentlicht: (2025)
von: Cao, Ji, et al.
Veröffentlicht: (2025)
BloomVQA: Assessing Hierarchical Multi-modal Comprehension
von: Gong, Yunye, et al.
Veröffentlicht: (2023)
von: Gong, Yunye, et al.
Veröffentlicht: (2023)
Similarity Trajectories: Linking Sampling Process to Artifacts in Diffusion-Generated Images
von: Menn, Dennis, et al.
Veröffentlicht: (2024)
von: Menn, Dennis, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
von: Das, Aryan, et al.
Veröffentlicht: (2026)
von: Das, Aryan, et al.
Veröffentlicht: (2026)
Diffusion-Based Environment-Aware Trajectory Prediction
von: Westny, Theodor, et al.
Veröffentlicht: (2024)
von: Westny, Theodor, et al.
Veröffentlicht: (2024)
On Improving the Algorithm-, Model-, and Data- Efficiency of Self-Supervised Learning
von: Cao, Yun-Hao, et al.
Veröffentlicht: (2024)
von: Cao, Yun-Hao, et al.
Veröffentlicht: (2024)
Pseudo-label Refinement for Improving Self-Supervised Learning Systems
von: Zia-ur-Rehman, et al.
Veröffentlicht: (2024)
von: Zia-ur-Rehman, et al.
Veröffentlicht: (2024)
Can Generative Models Improve Self-Supervised Representation Learning?
von: Ayromlou, Sana, et al.
Veröffentlicht: (2024)
von: Ayromlou, Sana, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026) -
SemEnrich: Self-Supervised Semantic Enrichment of Radiology Reports for Vision-Language Learning
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026) -
Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation
von: Byun, Ji Young, et al.
Veröffentlicht: (2026) -
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024) -
Bridging the Semantic Gaps: Improving Medical VQA Consistency with LLM-Augmented Question Sets
von: Ma, Yongpei, et al.
Veröffentlicht: (2025)