DualPrompt-MedCap: A Dual-Prompt Enhanced Approach for Medical Image Captioning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yining, Braytee, Ali, Prasad, Mukesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
From eye to AI: studying rodent social behavior in the era of machine Learning
von: Chindemi, Giuseppe, et al.
Veröffentlicht: (2025)
von: Chindemi, Giuseppe, et al.
Veröffentlicht: (2025)
Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery
von: Henriques, Angelo, et al.
Veröffentlicht: (2025)
von: Henriques, Angelo, et al.
Veröffentlicht: (2025)
SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition
von: Oh, Sukju, et al.
Veröffentlicht: (2026)
von: Oh, Sukju, et al.
Veröffentlicht: (2026)
Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
Towards Accurate and Efficient Waste Image Classification: A Hybrid Deep Learning and Machine Learning Approach
von: Nguyen, Ngoc-Bao-Quang, et al.
Veröffentlicht: (2025)
von: Nguyen, Ngoc-Bao-Quang, et al.
Veröffentlicht: (2025)
Dense Motion Captioning
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
Dynamic Arthroscopic Navigation System for Anterior Cruciate Ligament Reconstruction Based on Multi-level Memory Architecture
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
Prompt to Polyp: Medical Text-Conditioned Image Synthesis with Diffusion Models
von: Chaichuk, Mikhail, et al.
Veröffentlicht: (2025)
von: Chaichuk, Mikhail, et al.
Veröffentlicht: (2025)
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
von: Quy, Nguyen Lam Phu, et al.
Veröffentlicht: (2025)
von: Quy, Nguyen Lam Phu, et al.
Veröffentlicht: (2025)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
VLM-VPI: A Vision-Language Reasoning Framework for Improving Automated Vehicle-Pedestrian Interactions
von: Pu, Qingwen, et al.
Veröffentlicht: (2026)
von: Pu, Qingwen, et al.
Veröffentlicht: (2026)
Semi supervised GAN for smart microscopy, fast and data efficient cell cycle classification
von: Manick, Rajeev, et al.
Veröffentlicht: (2026)
von: Manick, Rajeev, et al.
Veröffentlicht: (2026)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
von: Chen, Pei-Chi, et al.
Veröffentlicht: (2025)
von: Chen, Pei-Chi, et al.
Veröffentlicht: (2025)
Image-Based Leopard Seal Recognition: Approaches and Challenges in Current Automated Systems
von: Salazar, Jorge Yero, et al.
Veröffentlicht: (2024)
von: Salazar, Jorge Yero, et al.
Veröffentlicht: (2024)
From Cheap to Pro: A Learning-based Adaptive Camera Parameter Network for Professional-Style Imaging
von: Li, Fuchen, et al.
Veröffentlicht: (2025)
von: Li, Fuchen, et al.
Veröffentlicht: (2025)
Domain-Adaptive Transformer for Data-Efficient Glioma Segmentation in Sub-Saharan MRI
von: Abolade, Ilerioluwakiiye, et al.
Veröffentlicht: (2025)
von: Abolade, Ilerioluwakiiye, et al.
Veröffentlicht: (2025)
Explainable vertebral fracture analysis with uncertainty estimation using differentiable rule-based classification
von: Skärström, Victor Wåhlstrand, et al.
Veröffentlicht: (2024)
von: Skärström, Victor Wåhlstrand, et al.
Veröffentlicht: (2024)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
Reducing Object Hallucination in LVLMs via Emphasizing Image-negative Tokens
von: Shen, Meng, et al.
Veröffentlicht: (2026)
von: Shen, Meng, et al.
Veröffentlicht: (2026)
SCA-Net: Spatial-Contextual Aggregation Network for Enhanced Small Building and Road Change Detection
von: Gholibeigi, Emad, et al.
Veröffentlicht: (2026)
von: Gholibeigi, Emad, et al.
Veröffentlicht: (2026)
FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes
von: Büsching, Marcel, et al.
Veröffentlicht: (2023)
von: Büsching, Marcel, et al.
Veröffentlicht: (2023)
From Dead Pixels to Editable Slides: Infographic Reconstruction into Native Google Slides via Vision-Language Region Understanding
von: Gonzalez, Leonardo
Veröffentlicht: (2026)
von: Gonzalez, Leonardo
Veröffentlicht: (2026)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
IMASHRIMP: Automatic White Shrimp (Penaeus vannamei) Biometrical Analysis from Laboratory Images Using Computer Vision and Deep Learning
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
von: Panek, Vojtech, et al.
Veröffentlicht: (2026)
von: Panek, Vojtech, et al.
Veröffentlicht: (2026)
SERA-H: Beyond Native Sentinel Spatial Limits for High-Resolution Canopy Height Mapping
von: Boudras, Thomas, et al.
Veröffentlicht: (2025)
von: Boudras, Thomas, et al.
Veröffentlicht: (2025)
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
von: Deng, Pei, et al.
Veröffentlicht: (2025)
von: Deng, Pei, et al.
Veröffentlicht: (2025)
Visible Iris Area as a Quality Metric for Reliable Iris Recognition Under Pupil Dilation and Eyelid Occlusion
von: Pessaud, Jack, et al.
Veröffentlicht: (2025)
von: Pessaud, Jack, et al.
Veröffentlicht: (2025)
Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion
von: Zhu, Yu, et al.
Veröffentlicht: (2025)
von: Zhu, Yu, et al.
Veröffentlicht: (2025)
Experimenting active and sequential learning in a medieval music manuscript
von: Sharma, Sachin, et al.
Veröffentlicht: (2025)
von: Sharma, Sachin, et al.
Veröffentlicht: (2025)
NeuroGaze-Distill: Brain-informed Distillation and Depression-Inspired Geometric Priors for Robust Facial Emotion Recognition
von: Li, Zilin, et al.
Veröffentlicht: (2025)
von: Li, Zilin, et al.
Veröffentlicht: (2025)
DRIFT open dataset: A drone-derived intelligence for traffic analysis in urban environment
von: Lee, Hyejin, et al.
Veröffentlicht: (2025)
von: Lee, Hyejin, et al.
Veröffentlicht: (2025)
OmniAcc: Personalized Accessibility Assistant Using Generative AI
von: Karki, Siddhant, et al.
Veröffentlicht: (2025)
von: Karki, Siddhant, et al.
Veröffentlicht: (2025)
From Photons to Physics: Autonomous Indoor Drones and the Future of Objective Property Assessment
von: Teikari, Petteri, et al.
Veröffentlicht: (2025)
von: Teikari, Petteri, et al.
Veröffentlicht: (2025)
SSD-GS: Scattering and Shadow Decomposition for Relightable 3D Gaussian Splatting
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
MSGS: Multispectral 3D Gaussian Splatting
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
von: Zheng, Iris, et al.
Veröffentlicht: (2026)
From Gaze to Insight: Bridging Human Visual Attention and Vision Language Model Explanation for Weakly-Supervised Medical Image Segmentation
von: Chen, Jingkun, et al.
Veröffentlicht: (2025)
von: Chen, Jingkun, et al.
Veröffentlicht: (2025)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025) -
From eye to AI: studying rodent social behavior in the era of machine Learning
von: Chindemi, Giuseppe, et al.
Veröffentlicht: (2025) -
Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery
von: Henriques, Angelo, et al.
Veröffentlicht: (2025) -
SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition
von: Oh, Sukju, et al.
Veröffentlicht: (2026) -
Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)