M3CoTBench: Benchmark Chain-of-Thought of MLLMs in Medical Image Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jiang, Juntao, Zhang, Jiangning, Bi, Yali, Bai, Jinsheng, Liu, Weixuan, Jin, Weiwei, Xue, Zhucun, Liu, Yong, Hu, Xiaobin, Yan, Shuicheng |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
RWKV-UNet: Improving UNet with Long-Range Cooperation for Effective Medical Image Segmentation
par: Jiang, Juntao, et autres
Publié: (2025)
par: Jiang, Juntao, et autres
Publié: (2025)
CASC: Condition-Aware Semantic Communication with Latent Diffusion Models
par: Chen, Weixuan, et autres
Publié: (2024)
par: Chen, Weixuan, et autres
Publié: (2024)
LV-UNet: A Lightweight and Vanilla Model for Medical Image Segmentation
par: Jiang, Juntao, et autres
Publié: (2024)
par: Jiang, Juntao, et autres
Publié: (2024)
Multi-Prototype Embedding Refinement for Semi-Supervised Medical Image Segmentation
par: Bi, Yali, et autres
Publié: (2025)
par: Bi, Yali, et autres
Publié: (2025)
Q-Agent: Quality-Driven Chain-of-Thought Image Restoration Agent through Robust Multimodal Large Language Model
par: Zhou, Yingjie, et autres
Publié: (2025)
par: Zhou, Yingjie, et autres
Publié: (2025)
Enhancing Image Privacy in Semantic Communication over Wiretap Channels leveraging Differential Privacy
par: Chen, Weixuan, et autres
Publié: (2024)
par: Chen, Weixuan, et autres
Publié: (2024)
Mutual Evidential Deep Learning for Medical Image Segmentation
par: He, Yuanpeng, et autres
Publié: (2025)
par: He, Yuanpeng, et autres
Publié: (2025)
Entropy-and-Channel-Aware Adaptive-Rate Semantic Communication with MLLM-Aided Feature Compensation
par: Chen, Weixuan, et autres
Publié: (2025)
par: Chen, Weixuan, et autres
Publié: (2025)
VoCo: A Simple-yet-Effective Volume Contrastive Learning Framework for 3D Medical Image Analysis
par: Wu, Linshan, et autres
Publié: (2024)
par: Wu, Linshan, et autres
Publié: (2024)
A Survey on Medical Image Compression: From Traditional to Learning-Based Approaches
par: Tong, Guofeng, et autres
Publié: (2025)
par: Tong, Guofeng, et autres
Publié: (2025)
Adaptive Regularized Low-Rank Tensor Decomposition for Hyperspectral Image Denoising and Destriping
par: Li, Dongyi, et autres
Publié: (2024)
par: Li, Dongyi, et autres
Publié: (2024)
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
par: Jin, Yili, et autres
Publié: (2025)
par: Jin, Yili, et autres
Publié: (2025)
Collaborative Learning for Unsupervised Multimodal Remote Sensing Image Registration: Integrating Self-Supervision and MIM-Guided Diffusion-Based Image Translation
par: Wei, Xiaochen, et autres
Publié: (2025)
par: Wei, Xiaochen, et autres
Publié: (2025)
A Low-Complexity View Synthesis Distortion Estimation Method for 3D Video with Large Baseline Considerations
par: Bi, Chongyuan, et autres
Publié: (2025)
par: Bi, Chongyuan, et autres
Publié: (2025)
MediViSTA: Medical Video Segmentation via Temporal Fusion SAM Adaptation for Echocardiography
par: Kim, Sekeun, et autres
Publié: (2023)
par: Kim, Sekeun, et autres
Publié: (2023)
Human Gaze-based Dual Teacher Guidance Learning for Semi-Supervised Medical Image Segmentation
par: Ge, Rongjun, et autres
Publié: (2026)
par: Ge, Rongjun, et autres
Publié: (2026)
PrismAudio: Decomposed Chain-of-Thoughts and Multi-dimensional Rewards for Video-to-Audio Generation
par: Liu, Huadai, et autres
Publié: (2025)
par: Liu, Huadai, et autres
Publié: (2025)
A Superposition Code-Based Semantic Communication Approach with Quantifiable and Controllable Security
par: Chen, Weixuan, et autres
Publié: (2024)
par: Chen, Weixuan, et autres
Publié: (2024)
Enhancing Privacy in Semantic Communication over Wiretap Channels leveraging Differential Privacy
par: Chen, Weixuan, et autres
Publié: (2025)
par: Chen, Weixuan, et autres
Publié: (2025)
Unified Medical Image Tokenizer for Autoregressive Synthesis and Understanding
par: Ma, Chenglong, et autres
Publié: (2025)
par: Ma, Chenglong, et autres
Publié: (2025)
Optimizing Prompt Strategies for SAM: Advancing lesion Segmentation Across Diverse Medical Imaging Modalities
par: Wang, Yuli, et autres
Publié: (2024)
par: Wang, Yuli, et autres
Publié: (2024)
Dataset and Benchmark for Enhancing Critical Retained Foreign Object Detection
par: Wang, Yuli, et autres
Publié: (2025)
par: Wang, Yuli, et autres
Publié: (2025)
Robust Grounding with MLLMs Against Occlusion and Small Objects via Language-Guided Semantic Cues
par: Park, Beomchan, et autres
Publié: (2026)
par: Park, Beomchan, et autres
Publié: (2026)
General Intelligent Imaging and Uncertainty Quantification by Deterministic Diffusion Model
par: Fan, Weiru, et autres
Publié: (2024)
par: Fan, Weiru, et autres
Publié: (2024)
Active Learning on Medical Image
par: Biswas, Angona, et autres
Publié: (2023)
par: Biswas, Angona, et autres
Publié: (2023)
Unsupervised Learning of Multi-modal Affine Registration for PET/CT
par: Chen, Junyu, et autres
Publié: (2024)
par: Chen, Junyu, et autres
Publié: (2024)
AstMatch: Adversarial Self-training Consistency Framework for Semi-Supervised Medical Image Segmentation
par: Zhu, Guanghao, et autres
Publié: (2024)
par: Zhu, Guanghao, et autres
Publié: (2024)
BMAD: Benchmarks for Medical Anomaly Detection
par: Bao, Jinan, et autres
Publié: (2023)
par: Bao, Jinan, et autres
Publié: (2023)
Can Knowledge Improve Security? A Coding-Enhanced Jamming Approach for Semantic Communication
par: Chen, Weixuan, et autres
Publié: (2025)
par: Chen, Weixuan, et autres
Publié: (2025)
Privacy-Preserving Semantic Communication over Wiretap Channels with Learnable Differential Privacy
par: Chen, Weixuan, et autres
Publié: (2025)
par: Chen, Weixuan, et autres
Publié: (2025)
MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs
par: Shi, Baorong, et autres
Publié: (2026)
par: Shi, Baorong, et autres
Publié: (2026)
Context Adaptive Extended Chain Coding for Semantic Map Compression
par: Yang, Runyu, et autres
Publié: (2026)
par: Yang, Runyu, et autres
Publié: (2026)
IMIL: Interactive Medical Image Learning Framework
par: Rao, Adrit, et autres
Publié: (2024)
par: Rao, Adrit, et autres
Publié: (2024)
SVFR: A Unified Framework for Generalized Video Face Restoration
par: Wang, Zhiyao, et autres
Publié: (2025)
par: Wang, Zhiyao, et autres
Publié: (2025)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
par: Jin, Yili, et autres
Publié: (2025)
par: Jin, Yili, et autres
Publié: (2025)
TransVFC: A Transformable Video Feature Compression Framework for Machines
par: Sun, Yuxiao, et autres
Publié: (2025)
par: Sun, Yuxiao, et autres
Publié: (2025)
Correlation Ratio for Unsupervised Learning of Multi-modal Deformable Registration
par: Chen, Xiaojian, et autres
Publié: (2025)
par: Chen, Xiaojian, et autres
Publié: (2025)
Quantitative Metrics for Benchmarking Medical Image Harmonization
par: Parida, Abhijeet, et autres
Publié: (2024)
par: Parida, Abhijeet, et autres
Publié: (2024)
MAD: Meta Adversarial Defense Benchmark
par: Peng, X., et autres
Publié: (2023)
par: Peng, X., et autres
Publié: (2023)
Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image Registration
par: Meng, Mingyuan, et autres
Publié: (2024)
par: Meng, Mingyuan, et autres
Publié: (2024)
Documents similaires
-
RWKV-UNet: Improving UNet with Long-Range Cooperation for Effective Medical Image Segmentation
par: Jiang, Juntao, et autres
Publié: (2025) -
CASC: Condition-Aware Semantic Communication with Latent Diffusion Models
par: Chen, Weixuan, et autres
Publié: (2024) -
LV-UNet: A Lightweight and Vanilla Model for Medical Image Segmentation
par: Jiang, Juntao, et autres
Publié: (2024) -
Multi-Prototype Embedding Refinement for Semi-Supervised Medical Image Segmentation
par: Bi, Yali, et autres
Publié: (2025) -
Q-Agent: Quality-Driven Chain-of-Thought Image Restoration Agent through Robust Multimodal Large Language Model
par: Zhou, Yingjie, et autres
Publié: (2025)