Prompt to Polyp: Medical Text-Conditioned Image Synthesis with Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chaichuk, Mikhail, Gautam, Sushant, Hicks, Steven, Tutubalina, Elena |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Continuous Receive Apodization Weights via Implicit Neural Representation for Ultrafast ICE Ultrasound Imaging
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025)
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025)
Modulated INR with Prior Embeddings for Ultrasound Imaging Reconstruction
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025)
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025)
Technical Report: Automated Optical Inspection of Surgical Instruments
von: Shafqat, Zunaira, et al.
Veröffentlicht: (2026)
von: Shafqat, Zunaira, et al.
Veröffentlicht: (2026)
Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays
von: Ko, Hanbin, et al.
Veröffentlicht: (2025)
von: Ko, Hanbin, et al.
Veröffentlicht: (2025)
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
MATEX: Multi-scale Attention and Text-guided Explainability of Medical Vision-Language Models
von: Imran, Muhammad, et al.
Veröffentlicht: (2026)
von: Imran, Muhammad, et al.
Veröffentlicht: (2026)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
Beyond RGB: Leveraging Vision Transformers for Thermal Weapon Segmentation
von: Kambhatla, Akhila, et al.
Veröffentlicht: (2025)
von: Kambhatla, Akhila, et al.
Veröffentlicht: (2025)
QSilk: Micrograin Stabilization and Adaptive Quantile Clipping for Detail-Friendly Latent Diffusion
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
von: Rychkovskiy, Denis
Veröffentlicht: (2025)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
Multi-AD: Cross-Domain Unsupervised Anomaly Detection for Medical and Industrial Applications
von: Rahmaniar, Wahyu, et al.
Veröffentlicht: (2026)
von: Rahmaniar, Wahyu, et al.
Veröffentlicht: (2026)
Dual-sensing driving detection model
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
Transforming faces into video stories -- VideoFace2.0
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
AQFusionNet: Multimodal Deep Learning for Air Quality Index Prediction with Imagery and Sensor Data
von: Kushal, Koushik Ahmed, et al.
Veröffentlicht: (2025)
von: Kushal, Koushik Ahmed, et al.
Veröffentlicht: (2025)
ARTPS: Depth-Enhanced Hybrid Anomaly Detection and Learnable Curiosity Score for Autonomous Rover Target Prioritization
von: Baydemir, Poyraz
Veröffentlicht: (2025)
von: Baydemir, Poyraz
Veröffentlicht: (2025)
TRACES: Temporal Recall with Contextual Embeddings for Real-Time Video Anomaly Detection
von: Siddiqui, Yousuf Ahmed, et al.
Veröffentlicht: (2025)
von: Siddiqui, Yousuf Ahmed, et al.
Veröffentlicht: (2025)
Data Augmentation and Resolution Enhancement using GANs and Diffusion Models for Tree Segmentation
von: Ferreira, Alessandro dos Santos, et al.
Veröffentlicht: (2025)
von: Ferreira, Alessandro dos Santos, et al.
Veröffentlicht: (2025)
Real Time Human Detection by Unmanned Aerial Vehicles
von: Guettala, Walid, et al.
Veröffentlicht: (2024)
von: Guettala, Walid, et al.
Veröffentlicht: (2024)
Automated Defect Detection for Mass-Produced Electronic Components Based on YOLO Object Detection Models
von: Mao, Wei-Lung, et al.
Veröffentlicht: (2025)
von: Mao, Wei-Lung, et al.
Veröffentlicht: (2025)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
A Novel Approach to Breast Cancer Segmentation using U-Net Model with Attention Mechanisms and FedProx
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
A large-scale, physically-based synthetic dataset for satellite pose estimation
von: Velkei, Szabolcs, et al.
Veröffentlicht: (2025)
von: Velkei, Szabolcs, et al.
Veröffentlicht: (2025)
OpenFusion++: An Open-vocabulary Real-time Scene Understanding System
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment
von: Zhao, Weiyi, et al.
Veröffentlicht: (2025)
von: Zhao, Weiyi, et al.
Veröffentlicht: (2025)
BreastDCEDL: A Comprehensive Breast Cancer DCE-MRI Dataset and Transformer Implementation for Treatment Response Prediction
von: Fridman, Naomi, et al.
Veröffentlicht: (2025)
von: Fridman, Naomi, et al.
Veröffentlicht: (2025)
Autoregressive Medical Image Segmentation via Next-Scale Mask Prediction
von: Chen, Tao, et al.
Veröffentlicht: (2025)
von: Chen, Tao, et al.
Veröffentlicht: (2025)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
Comparison of Neural Models for X-ray Image Classification in COVID-19 Detection
von: Togni, Jimi, et al.
Veröffentlicht: (2025)
von: Togni, Jimi, et al.
Veröffentlicht: (2025)
Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
von: Kang, Xueyang, et al.
Veröffentlicht: (2026)
von: Kang, Xueyang, et al.
Veröffentlicht: (2026)
Wafer Map Defect Classification Using Autoencoder-Based Data Augmentation and Convolutional Neural Network
von: Bao, Yin-Yin, et al.
Veröffentlicht: (2024)
von: Bao, Yin-Yin, et al.
Veröffentlicht: (2024)
EventFlow: Real-Time Neuromorphic Event-Driven Classification of Two-Phase Boiling Flow Regimes
von: Chang, Sanghyeon, et al.
Veröffentlicht: (2025)
von: Chang, Sanghyeon, et al.
Veröffentlicht: (2025)
Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification
von: Dong, Haohua, et al.
Veröffentlicht: (2025)
von: Dong, Haohua, et al.
Veröffentlicht: (2025)
Breast Cell Segmentation Under Extreme Data Constraints: Quantum Enhancement Meets Adaptive Loss Stabilization
von: Dasoju, Varun Kumar, et al.
Veröffentlicht: (2025)
von: Dasoju, Varun Kumar, et al.
Veröffentlicht: (2025)
VALE: A Multimodal Visual and Language Explanation Framework for Image Classifiers using eXplainable AI and Language Models
von: Natarajan, Purushothaman, et al.
Veröffentlicht: (2024)
von: Natarajan, Purushothaman, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Continuous Receive Apodization Weights via Implicit Neural Representation for Ultrafast ICE Ultrasound Imaging
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025) -
Modulated INR with Prior Embeddings for Ultrasound Imaging Reconstruction
von: Delaunay, Rémi, et al.
Veröffentlicht: (2025) -
Technical Report: Automated Optical Inspection of Surgical Instruments
von: Shafqat, Zunaira, et al.
Veröffentlicht: (2026) -
Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays
von: Ko, Hanbin, et al.
Veröffentlicht: (2025) -
CADE 2.5 - ZeResFDG: Frequency-Decoupled, Rescaled and Zero-Projected Guidance for SD/SDXL Latent Diffusion Models
von: Rychkovskiy, Denis
Veröffentlicht: (2025)