PromptMAD: Cross-Modal Prompting for Multi-Class Visual Anomaly Localization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | McCain, Duncan, Kashiani, Hossein, Afghah, Fatemeh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ROADS: Robust Prompt-driven Multi-Class Anomaly Detection under Domain Shift
von: Kashiani, Hossein, et al.
Veröffentlicht: (2024)
von: Kashiani, Hossein, et al.
Veröffentlicht: (2024)
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
DiSa: Directional Saliency-Aware Prompt Learning for Generalizable Vision-Language Models
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2025)
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2025)
FreqDebias: Towards Generalizable Deepfake Detection via Consistency-Driven Frequency Debiasing
von: Kashiani, Hossein, et al.
Veröffentlicht: (2025)
von: Kashiani, Hossein, et al.
Veröffentlicht: (2025)
CATFace: Cross-Attribute-Guided Transformer with Self-Attention Distillation for Low-Quality Face Recognition
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
Modality-Aware SAM: Sharpness-Aware-Minimization Driven Gradient Modulation for Harmonized Multimodal Learning
von: Nowdeh, Hossein R., et al.
Veröffentlicht: (2025)
von: Nowdeh, Hossein R., et al.
Veröffentlicht: (2025)
FlameFinder: Illuminating Obscured Fire through Smoke with Attentive Deep Metric Learning
von: Rajoli, Hossein, et al.
Veröffentlicht: (2024)
von: Rajoli, Hossein, et al.
Veröffentlicht: (2024)
Audio-Guided Visual Editing with Complex Multi-Modal Prompts
von: Kim, Hyeonyu, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonyu, et al.
Veröffentlicht: (2025)
Multi-Prompt with Depth Partitioned Cross-Modal Learning
von: Tian, Yingjie, et al.
Veröffentlicht: (2023)
von: Tian, Yingjie, et al.
Veröffentlicht: (2023)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
von: Li, Xu, et al.
Veröffentlicht: (2025)
von: Li, Xu, et al.
Veröffentlicht: (2025)
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts
von: De Marinis, Pasquale, et al.
Veröffentlicht: (2024)
von: De Marinis, Pasquale, et al.
Veröffentlicht: (2024)
Thermal Image Calibration and Correction using Unpaired Cycle-Consistent Adversarial Networks
von: Rajoli, Hossein, et al.
Veröffentlicht: (2024)
von: Rajoli, Hossein, et al.
Veröffentlicht: (2024)
Federated Cross-Modal Style-Aware Prompt Generation
von: Prasad, Suraj, et al.
Veröffentlicht: (2025)
von: Prasad, Suraj, et al.
Veröffentlicht: (2025)
Deep Correlated Prompting for Visual Recognition with Missing Modalities
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
Synergistic Prompting for Robust Visual Recognition with Missing Modalities
von: Zhang, Zhihui, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihui, et al.
Veröffentlicht: (2025)
DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
PromptMoE: Generalizable Zero-Shot Anomaly Detection via Visually-Guided Prompt Mixtures
von: Shao, Yuheng, et al.
Veröffentlicht: (2025)
von: Shao, Yuheng, et al.
Veröffentlicht: (2025)
KAnoCLIP: Zero-Shot Anomaly Detection through Knowledge-Driven Prompt Learning and Enhanced Cross-Modal Integration
von: Li, Chengyuan, et al.
Veröffentlicht: (2025)
von: Li, Chengyuan, et al.
Veröffentlicht: (2025)
Bidirectional Cross-Modal Prompting for Event-Frame Asymmetric Stereo
von: Xu, Ninghui, et al.
Veröffentlicht: (2026)
von: Xu, Ninghui, et al.
Veröffentlicht: (2026)
Manipulating Multimodal Agents via Cross-Modal Prompt Injection
von: Wang, Le, et al.
Veröffentlicht: (2025)
von: Wang, Le, et al.
Veröffentlicht: (2025)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
Unified Open-World Segmentation with Multi-Modal Prompts
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
User-Friendly Customized Generation with Multi-Modal Prompts
von: Zhong, Linhao, et al.
Veröffentlicht: (2024)
von: Zhong, Linhao, et al.
Veröffentlicht: (2024)
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2025)
Robust RGB-T Tracking via Learnable Visual Fourier Prompt Fine-tuning and Modality Fusion Prompt Generation
von: Yang, Hongtao, et al.
Veröffentlicht: (2025)
von: Yang, Hongtao, et al.
Veröffentlicht: (2025)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
von: Li, Jiachen, et al.
Veröffentlicht: (2026)
von: Li, Jiachen, et al.
Veröffentlicht: (2026)
LMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition
von: Xia, Peng, et al.
Veröffentlicht: (2023)
von: Xia, Peng, et al.
Veröffentlicht: (2023)
Align3D-AD: Cross-Modal Feature Alignment and Dual-Prompt Learning for Zero-shot 3D Anomaly Detection
von: Bai, Letian, et al.
Veröffentlicht: (2026)
von: Bai, Letian, et al.
Veröffentlicht: (2026)
ModalPrompt: Towards Efficient Multimodal Continual Instruction Tuning with Dual-Modality Guided Prompt
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
Agentic AI in Remote Sensing: Foundations, Taxonomy, and Emerging Systems
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2026)
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2026)
CVPT: Cross Visual Prompt Tuning
von: Huang, Lingyun, et al.
Veröffentlicht: (2024)
von: Huang, Lingyun, et al.
Veröffentlicht: (2024)
PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
von: Luo, Tianci, et al.
Veröffentlicht: (2026)
von: Luo, Tianci, et al.
Veröffentlicht: (2026)
Multi-Modal Prompt Learning on Blind Image Quality Assessment
von: Pan, Wensheng, et al.
Veröffentlicht: (2024)
von: Pan, Wensheng, et al.
Veröffentlicht: (2024)
Prompt Highlighter: Interactive Control for Multi-Modal LLMs
von: Zhang, Yuechen, et al.
Veröffentlicht: (2023)
von: Zhang, Yuechen, et al.
Veröffentlicht: (2023)
TCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
von: Yao, Hantao, et al.
Veröffentlicht: (2023)
von: Yao, Hantao, et al.
Veröffentlicht: (2023)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
Human-Free Automated Prompting for Vision-Language Anomaly Detection: Prompt Optimization with Meta-guiding Prompt Scheme
von: Chen, Pi-Wei, et al.
Veröffentlicht: (2024)
von: Chen, Pi-Wei, et al.
Veröffentlicht: (2024)
Exploring Interpretability for Visual Prompt Tuning with Cross-layer Concepts
von: Wang, Yubin, et al.
Veröffentlicht: (2025)
von: Wang, Yubin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ROADS: Robust Prompt-driven Multi-Class Anomaly Detection under Domain Shift
von: Kashiani, Hossein, et al.
Veröffentlicht: (2024) -
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024) -
DiSa: Directional Saliency-Aware Prompt Learning for Generalizable Vision-Language Models
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2025) -
FreqDebias: Towards Generalizable Deepfake Detection via Consistency-Driven Frequency Debiasing
von: Kashiani, Hossein, et al.
Veröffentlicht: (2025) -
CATFace: Cross-Attribute-Guided Transformer with Self-Attention Distillation for Low-Quality Face Recognition
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)