See Detail Say Clear: Towards Brain CT Report Generation via Pathological Clue-driven Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Chengxin, Ji, Junzhong, Shi, Yanzhao, Zhang, Xiaodan, Qu, Liangqiong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MEPNet: Medical Entity-balanced Prompting Network for Brain CT Report Generation
by: Zhang, Xiaodan, et al.
Published: (2025)
by: Zhang, Xiaodan, et al.
Published: (2025)
HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding
by: Shi, Yanzhao, et al.
Published: (2025)
by: Shi, Yanzhao, et al.
Published: (2025)
SeeClear: Reliable Transparent Object Depth Estimation via Generative Opacification
by: Wang, Xiaoying, et al.
Published: (2026)
by: Wang, Xiaoying, et al.
Published: (2026)
Say Cheese! Detail-Preserving Portrait Collection Generation via Natural Language Edits
by: Sun, Zelong, et al.
Published: (2026)
by: Sun, Zelong, et al.
Published: (2026)
Towards Unified Molecule-Enhanced Pathology Image Representation Learning via Integrating Spatial Transcriptomics
by: Han, Minghao, et al.
Published: (2024)
by: Han, Minghao, et al.
Published: (2024)
Unleashing the Potential of Large Language Models for Text-to-Image Generation through Autoregressive Representation Alignment
by: Xie, Xing, et al.
Published: (2025)
by: Xie, Xing, et al.
Published: (2025)
Unveiling the Ignorance of MLLMs: Seeing Clearly, Answering Incorrectly
by: Liu, Yexin, et al.
Published: (2024)
by: Liu, Yexin, et al.
Published: (2024)
DVG-Diffusion: Dual-View Guided Diffusion Model for CT Reconstruction from X-Rays
by: Xie, Xing, et al.
Published: (2025)
by: Xie, Xing, et al.
Published: (2025)
Pathology Report Generation and Multimodal Representation Learning for Cutaneous Melanocytic Lesions
by: Lucassen, Ruben T., et al.
Published: (2025)
by: Lucassen, Ruben T., et al.
Published: (2025)
On the Importance of Text Preprocessing for Multimodal Representation Learning and Pathology Report Generation
by: Lucassen, Ruben T., et al.
Published: (2025)
by: Lucassen, Ruben T., et al.
Published: (2025)
StyleTailor: Towards Personalized Fashion Styling via Hierarchical Negative Feedback
by: Ma, Hongbo, et al.
Published: (2025)
by: Ma, Hongbo, et al.
Published: (2025)
See it. Say it. Sorted: Agentic System for Compositional Diagram Generation
by: Zhang, Hantao, et al.
Published: (2025)
by: Zhang, Hantao, et al.
Published: (2025)
Anatomy-Guided Radiology Report Generation with Pathology-Aware Regional Prompts
by: Gao, Yijian, et al.
Published: (2024)
by: Gao, Yijian, et al.
Published: (2024)
See Further When Clear: Curriculum Consistency Model
by: Liu, Yunpeng, et al.
Published: (2024)
by: Liu, Yunpeng, et al.
Published: (2024)
Dia-LLaMA: Towards Large Language Model-driven CT Report Generation
by: Chen, Zhixuan, et al.
Published: (2024)
by: Chen, Zhixuan, et al.
Published: (2024)
Are VLMs Seeing or Just Saying? Uncovering the Illusion of Visual Re-examination
by: Shi, Chufan, et al.
Published: (2026)
by: Shi, Chufan, et al.
Published: (2026)
Seeing What You Say: Expressive Image Generation from Speech
by: Lee, Jiyoung, et al.
Published: (2025)
by: Lee, Jiyoung, et al.
Published: (2025)
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors
by: Wang, Xiangchen, et al.
Published: (2025)
by: Wang, Xiangchen, et al.
Published: (2025)
Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation
by: Chang, Chun-Peng, et al.
Published: (2025)
by: Chang, Chun-Peng, et al.
Published: (2025)
MOT FCG++: Enhanced Representation of Spatio-temporal Motion and Appearance Features
by: Fang, Yanzhao
Published: (2024)
by: Fang, Yanzhao
Published: (2024)
TRRG: Towards Truthful Radiology Report Generation With Cross-modal Disease Clue Enhanced Large Language Model
by: Wang, Yuhao, et al.
Published: (2024)
by: Wang, Yuhao, et al.
Published: (2024)
Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow
by: Liu, Chengxin, et al.
Published: (2026)
by: Liu, Chengxin, et al.
Published: (2026)
The Devil is in the Details: Simple Remedies for Image-to-LiDAR Representation Learning
by: Jo, Wonjun, et al.
Published: (2025)
by: Jo, Wonjun, et al.
Published: (2025)
FedVLMBench: Benchmarking Federated Fine-Tuning of Vision-Language Models
by: Zheng, Weiying, et al.
Published: (2025)
by: Zheng, Weiying, et al.
Published: (2025)
Uncovering Latent Pathological Signatures in Pulmonary CT via Cross-Window Knowledge Distillation
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
Unleashing the Potential of SAM for Medical Adaptation via Hierarchical Decoding
by: Cheng, Zhiheng, et al.
Published: (2024)
by: Cheng, Zhiheng, et al.
Published: (2024)
Aligning What EEG Can See: Structural Representations for Brain-Vision Matching
by: Tang, Jingyi, et al.
Published: (2026)
by: Tang, Jingyi, et al.
Published: (2026)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
by: Tang, Feilong, et al.
Published: (2025)
by: Tang, Feilong, et al.
Published: (2025)
SeeClear: Semantic Distillation Enhances Pixel Condensation for Video Super-Resolution
by: Tang, Qi, et al.
Published: (2024)
by: Tang, Qi, et al.
Published: (2024)
Seeing Clearly without Training: Mitigating Hallucinations in Multimodal LLMs for Remote Sensing
by: Liu, Yi, et al.
Published: (2026)
by: Liu, Yi, et al.
Published: (2026)
Towards Spatial Transcriptomics-driven Pathology Foundation Models
by: Hemker, Konstantin, et al.
Published: (2026)
by: Hemker, Konstantin, et al.
Published: (2026)
Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation
by: Yuan, Qianhao, et al.
Published: (2026)
by: Yuan, Qianhao, et al.
Published: (2026)
DetailFlow: 1D Coarse-to-Fine Autoregressive Image Generation via Next-Detail Prediction
by: Liu, Yiheng, et al.
Published: (2025)
by: Liu, Yiheng, et al.
Published: (2025)
Can Multimodal LLMs See Materials Clearly? A Multimodal Benchmark on Materials Characterization
by: Lai, Zhengzhao, et al.
Published: (2025)
by: Lai, Zhengzhao, et al.
Published: (2025)
ATSTrack: Enhancing Visual-Language Tracking by Aligning Temporal and Spatial Scales
by: Zhen, Yihao, et al.
Published: (2025)
by: Zhen, Yihao, et al.
Published: (2025)
Seeing through Imagination: Learning Scene Geometry via Implicit Spatial World Modeling
by: Cao, Meng, et al.
Published: (2025)
by: Cao, Meng, et al.
Published: (2025)
Improving Hierarchical Representations of Vectorized HD Maps with Perspective Clues
by: Zhang, Chi, et al.
Published: (2024)
by: Zhang, Chi, et al.
Published: (2024)
Towards Imperceptible JPEG Image Hiding: Multi-range Representations-driven Adversarial Stego Generation
by: Yang, Junxue, et al.
Published: (2025)
by: Yang, Junxue, et al.
Published: (2025)
Learning Brain Representation with Hierarchical Visual Embeddings
by: Zheng, Jiawen, et al.
Published: (2026)
by: Zheng, Jiawen, et al.
Published: (2026)
Prompt Guiding Multi-Scale Adaptive Sparse Representation-driven Network for Low-Dose CT MAR
by: Shi, Baoshun, et al.
Published: (2025)
by: Shi, Baoshun, et al.
Published: (2025)
Similar Items
-
MEPNet: Medical Entity-balanced Prompting Network for Brain CT Report Generation
by: Zhang, Xiaodan, et al.
Published: (2025) -
HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding
by: Shi, Yanzhao, et al.
Published: (2025) -
SeeClear: Reliable Transparent Object Depth Estimation via Generative Opacification
by: Wang, Xiaoying, et al.
Published: (2026) -
Say Cheese! Detail-Preserving Portrait Collection Generation via Natural Language Edits
by: Sun, Zelong, et al.
Published: (2026) -
Towards Unified Molecule-Enhanced Pathology Image Representation Learning via Integrating Spatial Transcriptomics
by: Han, Minghao, et al.
Published: (2024)