ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xiwei, Li, Yulong, Zhuang, Xinlin, Li, Xuhui, Chen, Jianxu, Yang, Haolin, Razzak, Imran, Xie, Yutong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025)
by: Liu, Xiwei, et al.
Published: (2025)
DoAtlas-1: A Causal Compilation Paradigm for Clinical AI
by: Li, Yulong, et al.
Published: (2026)
by: Li, Yulong, et al.
Published: (2026)
Robust Atypical Mitosis Classification with DenseNet121: Stain-Aware Augmentation and Hybrid Loss for Domain Generalization
by: Dukre, Adinath, et al.
Published: (2025)
by: Dukre, Adinath, et al.
Published: (2025)
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
by: Liu, Xiwei, et al.
Published: (2026)
by: Liu, Xiwei, et al.
Published: (2026)
A Machine Learning Approach to Predict Biological Age and its Longitudinal Drivers
by: Dunbayeva, Nazira, et al.
Published: (2025)
by: Dunbayeva, Nazira, et al.
Published: (2025)
SAM-aware Test-time Adaptation for Universal Medical Image Segmentation
by: Wu, Jianghao, et al.
Published: (2025)
by: Wu, Jianghao, et al.
Published: (2025)
Towards Efficient Medical Reasoning with Minimal Fine-Tuning Data
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
CoMT: Chain-of-Medical-Thought Reduces Hallucination in Medical Report Generation
by: Jiang, Yue, et al.
Published: (2024)
by: Jiang, Yue, et al.
Published: (2024)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
by: Zhao, Qingqing, et al.
Published: (2025)
by: Zhao, Qingqing, et al.
Published: (2025)
DualCoT-VLA: Visual-Linguistic Chain of Thought via Parallel Reasoning for Vision-Language-Action Models
by: Zhong, Zhide, et al.
Published: (2026)
by: Zhong, Zhide, et al.
Published: (2026)
C-CoT: Counterfactual Chain-of-Thought with Vision-Language Models for Safe Autonomous Driving
by: Tian, Kefei, et al.
Published: (2026)
by: Tian, Kefei, et al.
Published: (2026)
AIM-CoT: Active Information-driven Multimodal Chain-of-Thought for Vision-Language Reasoning
by: Li, Xiping, et al.
Published: (2025)
by: Li, Xiping, et al.
Published: (2025)
APEX: Learning Adaptive Priorities for Multi-Objective Alignment in Vision-Language Generation
by: Chen, Dongliang, et al.
Published: (2026)
by: Chen, Dongliang, et al.
Published: (2026)
A Knowledge-driven Adaptive Collaboration of LLMs for Enhancing Medical Decision-making
by: Wu, Xiao, et al.
Published: (2025)
by: Wu, Xiao, et al.
Published: (2025)
CMSA-Net: Causal Multi-scale Aggregation with Adaptive Multi-source Reference for Video Polyp Segmentation
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
Step-CoT: Stepwise Visual Chain-of-Thought for Medical Visual Question Answering
by: Fan, Lin, et al.
Published: (2026)
by: Fan, Lin, et al.
Published: (2026)
ClinKD: Cross-Modal Clinical Knowledge Distiller For Multi-Task Medical Images
by: Ge, Hongyu, et al.
Published: (2025)
by: Ge, Hongyu, et al.
Published: (2025)
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
CoFFT: Chain of Foresight-Focus Thought for Visual Language Models
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment
by: Azeez, Mohammad Anas, et al.
Published: (2026)
by: Azeez, Mohammad Anas, et al.
Published: (2026)
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning
by: Wang, Yifan, et al.
Published: (2026)
by: Wang, Yifan, et al.
Published: (2026)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
by: Kolli, Govinda, et al.
Published: (2026)
by: Kolli, Govinda, et al.
Published: (2026)
MedCoT: Medical Chain of Thought via Hierarchical Expert
by: Liu, Jiaxiang, et al.
Published: (2024)
by: Liu, Jiaxiang, et al.
Published: (2024)
CHIPS: Efficient CLIP Adaptation via Curvature-aware Hybrid Influence-based Data Selection
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
MINT-CoT: Enabling Interleaved Visual Tokens in Mathematical Chain-of-Thought Reasoning
by: Chen, Xinyan, et al.
Published: (2025)
by: Chen, Xinyan, et al.
Published: (2025)
CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models
by: Cheng, Zihui, et al.
Published: (2024)
by: Cheng, Zihui, et al.
Published: (2024)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
by: Wang, Tianqi, et al.
Published: (2024)
by: Wang, Tianqi, et al.
Published: (2024)
Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
by: Shao, Hao, et al.
Published: (2024)
by: Shao, Hao, et al.
Published: (2024)
Phenome-Wide Multi-Omics Integration Uncovers Distinct Archetypes of Human Aging
by: Li, Huifa, et al.
Published: (2025)
by: Li, Huifa, et al.
Published: (2025)
CoTBox-TTT: Grounding Medical VQA with Visual Chain-of-Thought Boxes During Test-time Training
by: Qian, Jiahe, et al.
Published: (2025)
by: Qian, Jiahe, et al.
Published: (2025)
CoT4Det: A Chain-of-Thought Framework for Perception-Oriented Vision-Language Tasks
by: Qi, Yu, et al.
Published: (2025)
by: Qi, Yu, et al.
Published: (2025)
MedMO: Grounding and Understanding Multimodal Large Language Model for Medical Images
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Boosting Multi-modal Keyphrase Prediction with Dynamic Chain-of-Thought in Vision-Language Models
by: Ma, Qihang, et al.
Published: (2025)
by: Ma, Qihang, et al.
Published: (2025)
Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine
by: Wu, Yuan, et al.
Published: (2026)
by: Wu, Yuan, et al.
Published: (2026)
Cause-Aware Empathetic Response Generation via Chain-of-Thought Fine-Tuning
by: Chen, Xinhao, et al.
Published: (2024)
by: Chen, Xinhao, et al.
Published: (2024)
S-Chain: Structured Visual Chain-of-Thought For Medicine
by: Le-Duc, Khai, et al.
Published: (2025)
by: Le-Duc, Khai, et al.
Published: (2025)
Chain-of-Anomaly Thoughts with Large Vision-Language Models
by: Domingos, Pedro, et al.
Published: (2025)
by: Domingos, Pedro, et al.
Published: (2025)
Retinal Lipidomics Associations as Candidate Biomarkers for Cardiovascular Health
by: Inamullah, et al.
Published: (2025)
by: Inamullah, et al.
Published: (2025)
scAGC: Learning Adaptive Cell Graphs with Contrastive Guidance for Single-Cell Clustering
by: Li, Huifa, et al.
Published: (2025)
by: Li, Huifa, et al.
Published: (2025)
Similar Items
-
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025) -
DoAtlas-1: A Causal Compilation Paradigm for Clinical AI
by: Li, Yulong, et al.
Published: (2026) -
Robust Atypical Mitosis Classification with DenseNet121: Stain-Aware Augmentation and Hybrid Loss for Domain Generalization
by: Dukre, Adinath, et al.
Published: (2025) -
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
by: Liu, Xiwei, et al.
Published: (2026) -
A Machine Learning Approach to Predict Biological Age and its Longitudinal Drivers
by: Dunbayeva, Nazira, et al.
Published: (2025)