Contrastive Language Prompting to Ease False Positives in Medical Anomaly Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, YeongHyeon, Kim, Myung Jin, Kim, Hyeong Seok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Feature Attenuation of Defective Representation Can Resolve Incomplete Masking on Anomaly Detection
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024)
Empirical Analysis of Anomaly Detection on Hyperspectral Imaging Using Dimension Reduction Methods
von: Kim, Dongeon, et al.
Veröffentlicht: (2024)
von: Kim, Dongeon, et al.
Veröffentlicht: (2024)
Anomaly Detection by Effectively Leveraging Synthetic Images
von: Kang, Sungho, et al.
Veröffentlicht: (2025)
von: Kang, Sungho, et al.
Veröffentlicht: (2025)
Bayesian Principles Improve Prompt Learning In Vision-Language Models
von: Kim, Mingyu, et al.
Veröffentlicht: (2025)
von: Kim, Mingyu, et al.
Veröffentlicht: (2025)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)
MAGIC: Few-Shot Mask-Guided Anomaly Inpainting with Prompt Perturbation, Spatially Adaptive Guidance, and Context Awareness
von: Choi, JaeHyuck, et al.
Veröffentlicht: (2025)
von: Choi, JaeHyuck, et al.
Veröffentlicht: (2025)
FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language Alignment
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Myunsoo, et al.
Veröffentlicht: (2025)
EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
Resource-Efficient Medical Report Generation using Large Language Models
von: Abdullah, et al.
Veröffentlicht: (2024)
von: Abdullah, et al.
Veröffentlicht: (2024)
Homeomorphism Prior for False Positive and Negative Problem in Medical Image Dense Contrastive Representation Learning
von: He, Yuting, et al.
Veröffentlicht: (2025)
von: He, Yuting, et al.
Veröffentlicht: (2025)
Scene Depth Estimation from Traditional Oriental Landscape Paintings
von: Kang, Sungho, et al.
Veröffentlicht: (2024)
von: Kang, Sungho, et al.
Veröffentlicht: (2024)
Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models
von: Kim, Donghoon, et al.
Veröffentlicht: (2025)
von: Kim, Donghoon, et al.
Veröffentlicht: (2025)
Positive-Augmented Contrastive Learning for Vision-and-Language Evaluation and Training
von: Sarto, Sara, et al.
Veröffentlicht: (2024)
von: Sarto, Sara, et al.
Veröffentlicht: (2024)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
MoECLIP: Patch-Specialized Experts for Zero-shot Anomaly Detection
von: Park, Jun Yeong, et al.
Veröffentlicht: (2026)
von: Park, Jun Yeong, et al.
Veröffentlicht: (2026)
Hallucination Benchmark in Medical Visual Question Answering
von: Wu, Jinge, et al.
Veröffentlicht: (2024)
von: Wu, Jinge, et al.
Veröffentlicht: (2024)
Evaluating Multimodal Generative AI with Korean Educational Standards
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
von: Park, Sanghee, et al.
Veröffentlicht: (2025)
Prompt-Driven Contrastive Learning for Transferable Adversarial Attacks
von: Yang, Hunmin, et al.
Veröffentlicht: (2024)
von: Yang, Hunmin, et al.
Veröffentlicht: (2024)
SyncVSR: Data-Efficient Visual Speech Recognition with End-to-End Crossmodal Audio Token Synchronization
von: Ahn, Young Jin, et al.
Veröffentlicht: (2024)
von: Ahn, Young Jin, et al.
Veröffentlicht: (2024)
Large Language Models can Share Images, Too!
von: Lee, Young-Jun, et al.
Veröffentlicht: (2023)
von: Lee, Young-Jun, et al.
Veröffentlicht: (2023)
VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
von: Park, Seongheon, et al.
Veröffentlicht: (2026)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
von: Kim, Jeongho, et al.
Veröffentlicht: (2024)
M4CXR: Exploring Multi-task Potentials of Multi-modal Large Language Models for Chest X-ray Interpretation
von: Park, Jonggwon, et al.
Veröffentlicht: (2024)
von: Park, Jonggwon, et al.
Veröffentlicht: (2024)
LINGO-Space: Language-Conditioned Incremental Grounding for Space
von: Kim, Dohyun, et al.
Veröffentlicht: (2024)
von: Kim, Dohyun, et al.
Veröffentlicht: (2024)
Towards Visual-Prompt Temporal Answering Grounding in Medical Instructional Video
von: Li, Bin, et al.
Veröffentlicht: (2022)
von: Li, Bin, et al.
Veröffentlicht: (2022)
KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
von: Kim, Yoonshik, et al.
Veröffentlicht: (2025)
SIMPLOT: Enhancing Chart Question Answering by Distilling Essentials
von: Kim, Wonjoong, et al.
Veröffentlicht: (2024)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2024)
MAFA: Managing False Negatives for Vision-Language Pre-training
von: Byun, Jaeseok, et al.
Veröffentlicht: (2023)
von: Byun, Jaeseok, et al.
Veröffentlicht: (2023)
VLind-Bench: Measuring Language Priors in Large Vision-Language Models
von: Lee, Kang-il, et al.
Veröffentlicht: (2024)
von: Lee, Kang-il, et al.
Veröffentlicht: (2024)
Enhancing Contrastive Learning with Efficient Combinatorial Positive Pairing
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
von: Kim, Jaeill, et al.
Veröffentlicht: (2024)
CANVAS: A Benchmark for Vision-Language Models on Tool-Based User Interface Design
von: Jeong, Daeheon, et al.
Veröffentlicht: (2025)
von: Jeong, Daeheon, et al.
Veröffentlicht: (2025)
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
von: Kim, Youngmin, et al.
Veröffentlicht: (2025)
von: Kim, Youngmin, et al.
Veröffentlicht: (2025)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
von: Chen, Wenting, et al.
Veröffentlicht: (2024)
von: Chen, Wenting, et al.
Veröffentlicht: (2024)
Real-Time Multimodal Cognitive Assistant for Emergency Medical Services
von: Weerasinghe, Keshara, et al.
Veröffentlicht: (2024)
von: Weerasinghe, Keshara, et al.
Veröffentlicht: (2024)
Text Change Detection in Multilingual Documents Using Image Comparison
von: Park, Doyoung, et al.
Veröffentlicht: (2024)
von: Park, Doyoung, et al.
Veröffentlicht: (2024)
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
Background-Aware Defect Generation for Robust Industrial Anomaly Detection
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2025)
LEAP:D -- A Novel Prompt-based Approach for Domain-Generalized Aerial Object Detection
von: Park, Chanyeong, et al.
Veröffentlicht: (2024)
von: Park, Chanyeong, et al.
Veröffentlicht: (2024)
MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models
von: Paik, Gio, et al.
Veröffentlicht: (2025)
von: Paik, Gio, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Feature Attenuation of Defective Representation Can Resolve Incomplete Masking on Anomaly Detection
von: Park, YeongHyeon, et al.
Veröffentlicht: (2024) -
Empirical Analysis of Anomaly Detection on Hyperspectral Imaging Using Dimension Reduction Methods
von: Kim, Dongeon, et al.
Veröffentlicht: (2024) -
Anomaly Detection by Effectively Leveraging Synthetic Images
von: Kang, Sungho, et al.
Veröffentlicht: (2025) -
Bayesian Principles Improve Prompt Learning In Vision-Language Models
von: Kim, Mingyu, et al.
Veröffentlicht: (2025) -
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
von: Kalluri, Tarun, et al.
Veröffentlicht: (2024)