Prototype-Based Knowledge Guidance for Fine-Grained Structured Radiology Reporting
Fuente:
arXiv
Saved in:
| Main Authors: | Pellegrini, Chantal, Delchev, Adrian, Özsoy, Ege, Navab, Nassir, Keicher, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ORacle: Large Vision-Language Models for Knowledge-Guided Holistic OR Domain Modeling
by: Özsoy, Ege, et al.
Published: (2024)
by: Özsoy, Ege, et al.
Published: (2024)
RaDialog: A Large Vision-Language Model for Radiology Report Generation and Conversational Assistance
by: Pellegrini, Chantal, et al.
Published: (2023)
by: Pellegrini, Chantal, et al.
Published: (2023)
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Specialized Foundation Models for Intelligent Operating Rooms
by: Özsoy, Ege, et al.
Published: (2025)
by: Özsoy, Ege, et al.
Published: (2025)
CAT-SG: A Large Dynamic Scene Graph Dataset for Fine-Grained Understanding of Cataract Surgery
by: Holm, Felix, et al.
Published: (2025)
by: Holm, Felix, et al.
Published: (2025)
EHR2Path: Scalable Modeling of Longitudinal Patient Pathways from Multimodal Electronic Health Records
by: Pellegrini, Chantal, et al.
Published: (2025)
by: Pellegrini, Chantal, et al.
Published: (2025)
MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments
by: Özsoy, Ege, et al.
Published: (2025)
by: Özsoy, Ege, et al.
Published: (2025)
Rewarding Doubt: A Reinforcement Learning Approach to Calibrated Confidence Expression of Large Language Models
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Location-Free Scene Graph Generation
by: Özsoy, Ege, et al.
Published: (2023)
by: Özsoy, Ege, et al.
Published: (2023)
EgoExOR: An Ego-Exo-Centric Operating Room Dataset for Surgical Activity Understanding
by: Özsoy, Ege, et al.
Published: (2025)
by: Özsoy, Ege, et al.
Published: (2025)
PanORama: Multiview Consistent Panoptic Segmentation in Operating Rooms
by: Gürbüz, Tuna, et al.
Published: (2026)
by: Gürbüz, Tuna, et al.
Published: (2026)
Semantic Scene Graph for Ultrasound Image Explanation and Scanning Guidance
by: Li, Xuesong, et al.
Published: (2025)
by: Li, Xuesong, et al.
Published: (2025)
ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes
by: Holm, Felix, et al.
Published: (2025)
by: Holm, Felix, et al.
Published: (2025)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
ProtoVQA: An Adaptable Prototypical Framework for Explainable Fine-Grained Visual Question Answering
by: Diao, Xingjian, et al.
Published: (2025)
by: Diao, Xingjian, et al.
Published: (2025)
Adversarial Wear and Tear: Exploiting Natural Damage for Generating Physical-World Adversarial Examples
by: Irshad, Samra, et al.
Published: (2025)
by: Irshad, Samra, et al.
Published: (2025)
Stress-Aware Resilient Neural Training
by: Shakarami, Ashkan, et al.
Published: (2025)
by: Shakarami, Ashkan, et al.
Published: (2025)
VeLU: Variance-enhanced Learning Unit for Deep Neural Networks
by: Shakarami, Ashkan, et al.
Published: (2025)
by: Shakarami, Ashkan, et al.
Published: (2025)
Foundation Visual Encoders Are Secretly Few-Shot Anomaly Detectors
by: Zhai, Guangyao, et al.
Published: (2025)
by: Zhai, Guangyao, et al.
Published: (2025)
Advancing Surgical VQA with Scene Graph Knowledge
by: Yuan, Kun, et al.
Published: (2023)
by: Yuan, Kun, et al.
Published: (2023)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
by: Stilz, Florian, et al.
Published: (2026)
by: Stilz, Florian, et al.
Published: (2026)
VISAGE: Video Synthesis using Action Graphs for Surgery
by: Yeganeh, Yousef, et al.
Published: (2024)
by: Yeganeh, Yousef, et al.
Published: (2024)
R2GenKG: Hierarchical Multi-modal Knowledge Graph for LLM-based Radiology Report Generation
by: Wang, Futian, et al.
Published: (2025)
by: Wang, Futian, et al.
Published: (2025)
Fine-Grained Knowledge Structuring and Retrieval for Visual Question Answering
by: Zhang, Zhengxuan, et al.
Published: (2025)
by: Zhang, Zhengxuan, et al.
Published: (2025)
From Linear Probing to Joint-Weighted Token Hierarchy: A Foundation Model Bridging Global and Cellular Representations in Biomarker Detection
by: Liu, Jingsong, et al.
Published: (2025)
by: Liu, Jingsong, et al.
Published: (2025)
Understanding the Fine-Grained Knowledge Capabilities of Vision-Language Models
by: Ghosh, Dhruba, et al.
Published: (2026)
by: Ghosh, Dhruba, et al.
Published: (2026)
Boosting Fine-Grained Visual Anomaly Detection with Coarse-Knowledge-Aware Adversarial Learning
by: Fang, Qingqing, et al.
Published: (2024)
by: Fang, Qingqing, et al.
Published: (2024)
Learning Part Knowledge to Facilitate Category Understanding for Fine-Grained Generalized Category Discovery
by: Wang, Enguang, et al.
Published: (2025)
by: Wang, Enguang, et al.
Published: (2025)
Evaluating Automated Radiology Report Quality through Fine-Grained Phrasal Grounding of Clinical Findings
by: Mahmood, Razi, et al.
Published: (2024)
by: Mahmood, Razi, et al.
Published: (2024)
Speckle2Self: Self-Supervised Ultrasound Speckle Reduction Without Clean Data
by: Li, Xuesong, et al.
Published: (2025)
by: Li, Xuesong, et al.
Published: (2025)
SANGRIA: Surgical Video Scene Graph Optimization for Surgical Workflow Prediction
by: Köksal, Çağhan, et al.
Published: (2024)
by: Köksal, Çağhan, et al.
Published: (2024)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
by: Chen, Tingxuan, et al.
Published: (2025)
by: Chen, Tingxuan, et al.
Published: (2025)
PhenoKG: Knowledge Graph-Driven Gene Discovery and Patient Insights from Phenotypes Alone
by: Zaripova, Kamilia, et al.
Published: (2025)
by: Zaripova, Kamilia, et al.
Published: (2025)
Unit-Based Histopathology Tissue Segmentation via Multi-Level Feature Representation
by: Shakarami, Ashkan, et al.
Published: (2025)
by: Shakarami, Ashkan, et al.
Published: (2025)
Task Prototype-Based Knowledge Retrieval for Multi-Task Learning from Partially Annotated Data
by: Oh, Youngmin, et al.
Published: (2026)
by: Oh, Youngmin, et al.
Published: (2026)
ProtoQuant: Quantization of Prototypical Parts For General and Fine-Grained Image Classification
by: Janusz, Mikołaj, et al.
Published: (2026)
by: Janusz, Mikołaj, et al.
Published: (2026)
Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation
by: Baba, Kaito, et al.
Published: (2026)
by: Baba, Kaito, et al.
Published: (2026)
EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2024)
by: Zhai, Guangyao, et al.
Published: (2024)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
by: Gu, Difei, et al.
Published: (2025)
by: Gu, Difei, et al.
Published: (2025)
Similar Items
-
ORacle: Large Vision-Language Models for Knowledge-Guided Holistic OR Domain Modeling
by: Özsoy, Ege, et al.
Published: (2024) -
RaDialog: A Large Vision-Language Model for Radiology Report Generation and Conversational Assistance
by: Pellegrini, Chantal, et al.
Published: (2023) -
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
by: Bani-Harouni, David, et al.
Published: (2025) -
Specialized Foundation Models for Intelligent Operating Rooms
by: Özsoy, Ege, et al.
Published: (2025) -
CAT-SG: A Large Dynamic Scene Graph Dataset for Fine-Grained Understanding of Cataract Surgery
by: Holm, Felix, et al.
Published: (2025)