Memory-based Cross-modal Semantic Alignment Network for Radiology Report Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Tao, Yitian, Ma, Liyan, Yu, Jing, Zhang, Han |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RIHA: Report-Image Hierarchical Alignment for Radiology Report Generation
by: Chen, Yucheng, et al.
Published: (2026)
by: Chen, Yucheng, et al.
Published: (2026)
Online Iterative Self-Alignment for Radiology Report Generation
by: Xiao, Ting, et al.
Published: (2025)
by: Xiao, Ting, et al.
Published: (2025)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
Semantically Informed Salient Regions Guided Radiology Report Generation
by: Hou, Zeyi, et al.
Published: (2025)
by: Hou, Zeyi, et al.
Published: (2025)
R2GenKG: Hierarchical Multi-modal Knowledge Graph for LLM-based Radiology Report Generation
by: Wang, Futian, et al.
Published: (2025)
by: Wang, Futian, et al.
Published: (2025)
LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
by: Zhu, Bin, et al.
Published: (2023)
by: Zhu, Bin, et al.
Published: (2023)
EMRRG: Efficient Fine-Tuning Pre-trained X-ray Mamba Networks for Radiology Report Generation
by: Zhang, Mingzheng, et al.
Published: (2025)
by: Zhang, Mingzheng, et al.
Published: (2025)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
Designing a Robust Radiology Report Generation System
by: Singh, Sonit
Published: (2024)
by: Singh, Sonit
Published: (2024)
Single-Sample Black-Box Membership Inference Attack against Vision-Language Models via Cross-modal Semantic Alignment
by: Li, Jiaqing, et al.
Published: (2026)
by: Li, Jiaqing, et al.
Published: (2026)
DDaTR: Dynamic Difference-aware Temporal Residual Network for Longitudinal Radiology Report Generation
by: Song, Shanshan, et al.
Published: (2025)
by: Song, Shanshan, et al.
Published: (2025)
e5-omni: Explicit Cross-modal Alignment for Omni-modal Embeddings
by: Chen, Haonan, et al.
Published: (2026)
by: Chen, Haonan, et al.
Published: (2026)
Uncovering Knowledge Gaps in Radiology Report Generation Models through Knowledge Graphs
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
A Novel Spike Transformer Network for Depth Estimation from Event Cameras via Cross-modality Knowledge Distillation
by: Zhang, Xin, et al.
Published: (2024)
by: Zhang, Xin, et al.
Published: (2024)
AlignMamba: Enhancing Multimodal Mamba with Local and Global Cross-modal Alignment
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
Spectral Discrepancy and Cross-modal Semantic Consistency Learning for Object Detection in Hyperspectral Image
by: He, Xiao, et al.
Published: (2025)
by: He, Xiao, et al.
Published: (2025)
Enhancing Radiology Report Generation and Visual Grounding using Reinforcement Learning
by: Gundersen, Benjamin, et al.
Published: (2025)
by: Gundersen, Benjamin, et al.
Published: (2025)
MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation
by: Lin, Yi, et al.
Published: (2026)
by: Lin, Yi, et al.
Published: (2026)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
by: Wang, Ziyi, et al.
Published: (2024)
by: Wang, Ziyi, et al.
Published: (2024)
ReXrank: A Public Leaderboard for AI-Powered Radiology Report Generation
by: Zhang, Xiaoman, et al.
Published: (2024)
by: Zhang, Xiaoman, et al.
Published: (2024)
Argus: Benchmarking and Enhancing Vision-Language Models for 3D Radiology Report Generation
by: Liu, Che, et al.
Published: (2024)
by: Liu, Che, et al.
Published: (2024)
TACFN: Transformer-based Adaptive Cross-modal Fusion Network for Multimodal Emotion Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
by: Lei, Youbo, et al.
Published: (2023)
by: Lei, Youbo, et al.
Published: (2023)
AnomalyControl: Learning Cross-modal Semantic Features for Controllable Anomaly Synthesis
by: He, Shidan, et al.
Published: (2024)
by: He, Shidan, et al.
Published: (2024)
Historical Report Guided Bi-modal Concurrent Learning for Pathology Report Generation
by: Zhang, Ling, et al.
Published: (2025)
by: Zhang, Ling, et al.
Published: (2025)
Focus on Focus: Focus-oriented Representation Learning and Multi-view Cross-modal Alignment for Glioma Grading
by: Pan, Li, et al.
Published: (2024)
by: Pan, Li, et al.
Published: (2024)
Multi-granularity Contrastive Cross-modal Collaborative Generation for End-to-End Long-term Video Question Answering
by: Yu, Ting, et al.
Published: (2024)
by: Yu, Ting, et al.
Published: (2024)
Quality Control for Radiology Report Generation Models via Auxiliary Auditing Components
by: Warr, Hermione, et al.
Published: (2024)
by: Warr, Hermione, et al.
Published: (2024)
CRRG-CLIP: Automatic Generation of Chest Radiology Reports and Classification of Chest Radiographs
by: Xu, Jianfei, et al.
Published: (2024)
by: Xu, Jianfei, et al.
Published: (2024)
3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Masked Contrastive Reconstruction for Cross-modal Medical Image-Report Retrieval
by: Wei, Zeqiang, et al.
Published: (2023)
by: Wei, Zeqiang, et al.
Published: (2023)
ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking
by: Ge, Jiawei, et al.
Published: (2026)
by: Ge, Jiawei, et al.
Published: (2026)
KARGEN: Knowledge-enhanced Automated Radiology Report Generation Using Large Language Models
by: Li, Yingshu, et al.
Published: (2024)
by: Li, Yingshu, et al.
Published: (2024)
An X-Ray Is Worth 15 Features: Sparse Autoencoders for Interpretable Radiology Report Generation
by: Abdulaal, Ahmed, et al.
Published: (2024)
by: Abdulaal, Ahmed, et al.
Published: (2024)
CSMCIR: CoT-Enhanced Symmetric Alignment with Memory Bank for Composed Image Retrieval
by: Qian, Zhipeng, et al.
Published: (2026)
by: Qian, Zhipeng, et al.
Published: (2026)
A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension
by: Rehman, Mohammad Zia Ur, et al.
Published: (2025)
by: Rehman, Mohammad Zia Ur, et al.
Published: (2025)
Memory-guided Network with Uncertainty-based Feature Augmentation for Few-shot Semantic Segmentation
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
by: Molino, Daniele, et al.
Published: (2025)
by: Molino, Daniele, et al.
Published: (2025)
R2GenCSR: Mining Contextual and Residual Information for LLMs-based Radiology Report Generation
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
Fusion-Mamba for Cross-modality Object Detection
by: Dong, Wenhao, et al.
Published: (2024)
by: Dong, Wenhao, et al.
Published: (2024)
Similar Items
-
RIHA: Report-Image Hierarchical Alignment for Radiology Report Generation
by: Chen, Yucheng, et al.
Published: (2026) -
Online Iterative Self-Alignment for Radiology Report Generation
by: Xiao, Ting, et al.
Published: (2025) -
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025) -
Semantically Informed Salient Regions Guided Radiology Report Generation
by: Hou, Zeyi, et al.
Published: (2025) -
R2GenKG: Hierarchical Multi-modal Knowledge Graph for LLM-based Radiology Report Generation
by: Wang, Futian, et al.
Published: (2025)