Visual Prompt Engineering for Vision Language Models in Radiology
Fuente:
arXiv
Guardado en:
| Autores principales: | Denner, Stefan, Bujotzek, Markus, Bounias, Dimitrios, Zimmerer, David, Stock, Raphael, Maier-Hein, Klaus |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Leveraging Foundation Models for Content-Based Image Retrieval in Radiology
por: Denner, Stefan, et al.
Publicado: (2024)
por: Denner, Stefan, et al.
Publicado: (2024)
nnLandmark: A Self-Configuring Method for 3D Medical Landmark Detection
por: Ertl, Alexandra, et al.
Publicado: (2025)
por: Ertl, Alexandra, et al.
Publicado: (2025)
Kaapana: A Comprehensive Open-Source Platform for Integrating AI in Medical Imaging Research Environments
por: Akünal, Ünal, et al.
Publicado: (2025)
por: Akünal, Ünal, et al.
Publicado: (2025)
Decoupling Semantic Similarity from Spatial Alignment for Neural Networks
por: Wald, Tassilo, et al.
Publicado: (2024)
por: Wald, Tassilo, et al.
Publicado: (2024)
Real-World Federated Learning in Radiology: Hurdles to overcome and Benefits to gain
por: Bujotzek, Markus R., et al.
Publicado: (2024)
por: Bujotzek, Markus R., et al.
Publicado: (2024)
Comparative Benchmarking of Failure Detection Methods in Medical Image Segmentation: Unveiling the Role of Confidence Aggregation
por: Zenk, Maximilian, et al.
Publicado: (2024)
por: Zenk, Maximilian, et al.
Publicado: (2024)
Lost in the Folds: When Cross-Validation Is Not a Deep Ensemble for Uncertainty Estimation
por: Kirscher, Tristan, et al.
Publicado: (2026)
por: Kirscher, Tristan, et al.
Publicado: (2026)
Temporal Flow Matching for Learning Spatio-Temporal Trajectories in 4D Longitudinal Medical Imaging
por: Disch, Nico Albert, et al.
Publicado: (2025)
por: Disch, Nico Albert, et al.
Publicado: (2025)
The Missing Piece: A Case for Pre-Training in 3D Medical Object Detection
por: Eckstein, Katharina, et al.
Publicado: (2025)
por: Eckstein, Katharina, et al.
Publicado: (2025)
Visual Alignment of Medical Vision-Language Models for Grounded Radiology Report Generation
por: Bose, Sarosij, et al.
Publicado: (2025)
por: Bose, Sarosij, et al.
Publicado: (2025)
Adapting Lightweight Vision Language Models for Radiological Visual Question Answering
por: Shourya, Aditya, et al.
Publicado: (2025)
por: Shourya, Aditya, et al.
Publicado: (2025)
Primus: Enforcing Attention Usage for 3D Medical Image Segmentation
por: Wald, Tassilo, et al.
Publicado: (2025)
por: Wald, Tassilo, et al.
Publicado: (2025)
Revisiting 3D Medical Scribble Supervision: Benchmarking Beyond Cardiac Segmentation
por: Gotkowski, Karol, et al.
Publicado: (2024)
por: Gotkowski, Karol, et al.
Publicado: (2024)
Towards Interactive Lesion Segmentation in Whole-Body PET/CT with Promptable Models
por: Rokuss, Maximilian, et al.
Publicado: (2025)
por: Rokuss, Maximilian, et al.
Publicado: (2025)
RadioActive: 3D Radiological Interactive Segmentation Benchmark
por: Ulrich, Constantin, et al.
Publicado: (2024)
por: Ulrich, Constantin, et al.
Publicado: (2024)
Automatic classification of prostate MR series type using image content and metadata
por: Krishnaswamy, Deepa, et al.
Publicado: (2024)
por: Krishnaswamy, Deepa, et al.
Publicado: (2024)
MeisenMeister: A Simple Two Stage Pipeline for Breast Cancer Classification on MRI
por: Hamm, Benjamin, et al.
Publicado: (2025)
por: Hamm, Benjamin, et al.
Publicado: (2025)
FedMVP: Federated Multimodal Visual Prompt Tuning for Vision-Language Models
por: Singha, Mainak, et al.
Publicado: (2025)
por: Singha, Mainak, et al.
Publicado: (2025)
Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy
por: Wienholt, Patrick, et al.
Publicado: (2025)
por: Wienholt, Patrick, et al.
Publicado: (2025)
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models
por: Woo, Sangmin, et al.
Publicado: (2025)
por: Woo, Sangmin, et al.
Publicado: (2025)
Beyond Knowledge Silos: Task Fingerprinting for Democratization of Medical Imaging AI
por: Godau, Patrick, et al.
Publicado: (2024)
por: Godau, Patrick, et al.
Publicado: (2024)
VPTracker: Global Vision-Language Tracking via Visual Prompt
por: Wang, Jingchao, et al.
Publicado: (2025)
por: Wang, Jingchao, et al.
Publicado: (2025)
Targeted Visual Prompting for Medical Visual Question Answering
por: Tascon-Morales, Sergio, et al.
Publicado: (2024)
por: Tascon-Morales, Sergio, et al.
Publicado: (2024)
Large Scale Supervised Pretraining For Traumatic Brain Injury Segmentation
por: Ulrich, Constantin, et al.
Publicado: (2025)
por: Ulrich, Constantin, et al.
Publicado: (2025)
Modeling Variants of Prompts for Vision-Language Models
por: Li, Ao, et al.
Publicado: (2025)
por: Li, Ao, et al.
Publicado: (2025)
Evaluating Vision Language Model Adaptations for Radiology Report Generation in Low-Resource Languages
por: Salmè, Marco, et al.
Publicado: (2025)
por: Salmè, Marco, et al.
Publicado: (2025)
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models
por: Li, Zheng, et al.
Publicado: (2024)
por: Li, Zheng, et al.
Publicado: (2024)
6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models
por: Mayer, Leon, et al.
Publicado: (2025)
por: Mayer, Leon, et al.
Publicado: (2025)
In the Era of Prompt Learning with Vision-Language Models
por: Jha, Ankit
Publicado: (2024)
por: Jha, Ankit
Publicado: (2024)
Revisiting Prompt Pretraining of Vision-Language Models
por: Chen, Zhenyuan, et al.
Publicado: (2024)
por: Chen, Zhenyuan, et al.
Publicado: (2024)
Generalizable Prompt Tuning for Vision-Language Models
por: Zhang, Qian
Publicado: (2024)
por: Zhang, Qian
Publicado: (2024)
Mixture of Prompt Learning for Vision Language Models
por: Du, Yu, et al.
Publicado: (2024)
por: Du, Yu, et al.
Publicado: (2024)
Active Prompt Learning in Vision Language Models
por: Bang, Jihwan, et al.
Publicado: (2023)
por: Bang, Jihwan, et al.
Publicado: (2023)
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations
por: Zhu, Kangyu, et al.
Publicado: (2025)
por: Zhu, Kangyu, et al.
Publicado: (2025)
A Multi-Stage Fine-Tuning and Ensembling Strategy for Pancreatic Tumor Segmentation in Diagnostic and Therapeutic MRI
por: Durugol, Omer Faruk, et al.
Publicado: (2025)
por: Durugol, Omer Faruk, et al.
Publicado: (2025)
Challenging Vision-Language Models with Surgical Data: A New Dataset and Broad Benchmarking Study
por: Mayer, Leon, et al.
Publicado: (2025)
por: Mayer, Leon, et al.
Publicado: (2025)
Divide and Conquer: A Large-Scale Dataset and Model for Left-Right Breast MRI Segmentation
por: Rokuss, Maximilian, et al.
Publicado: (2025)
por: Rokuss, Maximilian, et al.
Publicado: (2025)
VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference
por: Zhu, Hao, et al.
Publicado: (2026)
por: Zhu, Hao, et al.
Publicado: (2026)
Illusion-Aware Visual Preprocessing and Anti-Illusion Prompting for Classic Illusion Understanding in Vision-Language Models
por: Zha, Junli, et al.
Publicado: (2026)
por: Zha, Junli, et al.
Publicado: (2026)
Leveraging Vision-Language Models for Visual Grounding and Analysis of Automotive UI
por: Ernhofer, Benjamin Raphael, et al.
Publicado: (2025)
por: Ernhofer, Benjamin Raphael, et al.
Publicado: (2025)
Ejemplares similares
-
Leveraging Foundation Models for Content-Based Image Retrieval in Radiology
por: Denner, Stefan, et al.
Publicado: (2024) -
nnLandmark: A Self-Configuring Method for 3D Medical Landmark Detection
por: Ertl, Alexandra, et al.
Publicado: (2025) -
Kaapana: A Comprehensive Open-Source Platform for Integrating AI in Medical Imaging Research Environments
por: Akünal, Ünal, et al.
Publicado: (2025) -
Decoupling Semantic Similarity from Spatial Alignment for Neural Networks
por: Wald, Tassilo, et al.
Publicado: (2024) -
Real-World Federated Learning in Radiology: Hurdles to overcome and Benefits to gain
por: Bujotzek, Markus R., et al.
Publicado: (2024)