DiffVP: Differential Visual Semantic Prompting for LLM-Based CT Report Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Tian, Yuhe, Zhang, Kun, Ma, Haoran, Yan, Rui, Li, Yingtai, Wang, Rongsheng, Zhou, Shaohua Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedReason-R1: Learning to Reason for CT Diagnosis with Reinforcement Learning and Local Zoom
by: Li, Yifan, et al.
Published: (2025)
by: Li, Yifan, et al.
Published: (2025)
More performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM era
by: Li, Yingtai, et al.
Published: (2025)
by: Li, Yingtai, et al.
Published: (2025)
ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training
by: Wang, Rongsheng, et al.
Published: (2026)
by: Wang, Rongsheng, et al.
Published: (2026)
GreenRFM: Toward a resource-efficient radiology foundation model
by: Li, Yingtai, et al.
Published: (2026)
by: Li, Yingtai, et al.
Published: (2026)
VP-MEL: Visual Prompts Guided Multimodal Entity Linking
by: Mi, Hongze, et al.
Published: (2024)
by: Mi, Hongze, et al.
Published: (2024)
DiffPrompter: Differentiable Implicit Visual Prompts for Semantic-Segmentation in Adverse Conditions
by: Kalwar, Sanket, et al.
Published: (2023)
by: Kalwar, Sanket, et al.
Published: (2023)
VP Lab: a PEFT-Enabled Visual Prompting Laboratory for Semantic Segmentation
by: Avogaro, Niccolo, et al.
Published: (2025)
by: Avogaro, Niccolo, et al.
Published: (2025)
VP-NTK: Exploring the Benefits of Visual Prompting in Differentially Private Data Synthesis
by: Hsu, Chia-Yi, et al.
Published: (2025)
by: Hsu, Chia-Yi, et al.
Published: (2025)
3DGR-CT: Sparse-View CT Reconstruction with a 3D Gaussian Representation
by: Li, Yingtai, et al.
Published: (2023)
by: Li, Yingtai, et al.
Published: (2023)
VP-Bench: A Comprehensive Benchmark for Visual Prompting in Multimodal Large Language Models
by: Xu, Mingjie, et al.
Published: (2025)
by: Xu, Mingjie, et al.
Published: (2025)
AutoVP: An Automated Visual Prompting Framework and Benchmark
by: Tsao, Hsi-Ai, et al.
Published: (2023)
by: Tsao, Hsi-Ai, et al.
Published: (2023)
Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation
by: Tian, Yuanhe, et al.
Published: (2025)
by: Tian, Yuanhe, et al.
Published: (2025)
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
by: Lai, Haoran, et al.
Published: (2026)
by: Lai, Haoran, et al.
Published: (2026)
A General Knowledge Injection Framework for ICD Coding
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Concept-to-Pixel: Prompt-Free Universal Medical Image Segmentation
by: Chen, Haoyun, et al.
Published: (2026)
by: Chen, Haoyun, et al.
Published: (2026)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
SimCroP: Radiograph Representation Learning with Similarity-driven Cross-granularity Pre-training
by: Wang, Rongsheng, et al.
Published: (2025)
by: Wang, Rongsheng, et al.
Published: (2025)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
by: Yan, Weicai, et al.
Published: (2025)
by: Yan, Weicai, et al.
Published: (2025)
OT-VP: Optimal Transport-guided Visual Prompting for Test-Time Adaptation
by: Zhang, Yunbei, et al.
Published: (2024)
by: Zhang, Yunbei, et al.
Published: (2024)
LoR-VP: Low-Rank Visual Prompting for Efficient Vision Model Adaptation
by: Jin, Can, et al.
Published: (2025)
by: Jin, Can, et al.
Published: (2025)
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic Reasoning
by: Liu, Weimin, et al.
Published: (2026)
by: Liu, Weimin, et al.
Published: (2026)
Bridged Semantic Alignment for Zero-shot 3D Medical Image Diagnosis
by: Lai, Haoran, et al.
Published: (2025)
by: Lai, Haoran, et al.
Published: (2025)
GenVP: Generating Visual Puzzles with Contrastive Hierarchical VAEs
by: Basioti, Kalliopi, et al.
Published: (2025)
by: Basioti, Kalliopi, et al.
Published: (2025)
Diff3DS: Generating View-Consistent 3D Sketch via Differentiable Curve Rendering
by: Zhang, Yibo, et al.
Published: (2024)
by: Zhang, Yibo, et al.
Published: (2024)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
by: Sellam, Abdellah Zakaria, et al.
Published: (2026)
by: Sellam, Abdellah Zakaria, et al.
Published: (2026)
SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation
by: Tian, Yuanhe, et al.
Published: (2026)
by: Tian, Yuanhe, et al.
Published: (2026)
QCAgent: An agentic framework for quality-controllable pathology report generation from whole slide image
by: Wang, Rundong, et al.
Published: (2026)
by: Wang, Rundong, et al.
Published: (2026)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
by: Ma, Wenxin, et al.
Published: (2025)
by: Ma, Wenxin, et al.
Published: (2025)
P-Flow: Prompting Visual Effects Generation
by: Zhao, Rui, et al.
Published: (2026)
by: Zhao, Rui, et al.
Published: (2026)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
by: Chen, Hongzhou, et al.
Published: (2024)
by: Chen, Hongzhou, et al.
Published: (2024)
CARZero: Cross-Attention Alignment for Radiology Zero-Shot Classification
by: Lai, Haoran, et al.
Published: (2024)
by: Lai, Haoran, et al.
Published: (2024)
ECAMP: Entity-centered Context-aware Medical Vision Language Pre-training
by: Wang, Rongsheng, et al.
Published: (2023)
by: Wang, Rongsheng, et al.
Published: (2023)
Visual and Semantic Prompt Collaboration for Generalized Zero-Shot Learning
by: Jiang, Huajie, et al.
Published: (2025)
by: Jiang, Huajie, et al.
Published: (2025)
Diff-Restorer: Unleashing Visual Prompts for Diffusion-based Universal Image Restoration
by: Zhang, Yuhong, et al.
Published: (2024)
by: Zhang, Yuhong, et al.
Published: (2024)
Vision-Language Models for Automated 3D PET/CT Report Generation
by: Jiao, Wenpei, et al.
Published: (2025)
by: Jiao, Wenpei, et al.
Published: (2025)
Flexible Control of 3D CT Generation via Text and Semantically-Defined Segmentation Prompts
by: Dai, Weicheng, et al.
Published: (2026)
by: Dai, Weicheng, et al.
Published: (2026)
E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model
by: Lai, Haoran, et al.
Published: (2024)
by: Lai, Haoran, et al.
Published: (2024)
DiffPop: Plausibility-Guided Object Placement Diffusion for Image Composition
by: Liu, Jiacheng, et al.
Published: (2024)
by: Liu, Jiacheng, et al.
Published: (2024)
DiffCap: Diffusion-based Real-time Human Motion Capture using Sparse IMUs and a Monocular Camera
by: Pan, Shaohua, et al.
Published: (2025)
by: Pan, Shaohua, et al.
Published: (2025)
ZeroDiff: Solidified Visual-Semantic Correlation in Zero-Shot Learning
by: Ye, Zihan, et al.
Published: (2024)
by: Ye, Zihan, et al.
Published: (2024)
Similar Items
-
MedReason-R1: Learning to Reason for CT Diagnosis with Reinforcement Learning and Local Zoom
by: Li, Yifan, et al.
Published: (2025) -
More performant and scalable: Rethinking contrastive vision-language pre-training of radiology in the LLM era
by: Li, Yingtai, et al.
Published: (2025) -
ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training
by: Wang, Rongsheng, et al.
Published: (2026) -
GreenRFM: Toward a resource-efficient radiology foundation model
by: Li, Yingtai, et al.
Published: (2026) -
VP-MEL: Visual Prompts Guided Multimodal Entity Linking
by: Mi, Hongze, et al.
Published: (2024)