AG-VAS: Anchor-Guided Zero-Shot Visual Anomaly Segmentation with Large Multimodal Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Qu, Zhen, Tao, Xian, Bao, Xiaoyi, Wang, Dingrong, Qu, ShiChen, Zhang, Zhengtao, Wang, Xingang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DictAS: A Framework for Class-Generalizable Few-Shot Anomaly Segmentation via Dictionary Lookup
di: Qu, Zhen, et al.
Pubblicazione: (2025)
di: Qu, Zhen, et al.
Pubblicazione: (2025)
Bayesian Prompt Flow Learning for Zero-Shot Anomaly Detection
di: Qu, Zhen, et al.
Pubblicazione: (2025)
di: Qu, Zhen, et al.
Pubblicazione: (2025)
ALMRR: Anomaly Localization Mamba on Industrial Textured Surface with Feature Reconstruction and Refinement
di: Qu, Shichen, et al.
Pubblicazione: (2024)
di: Qu, Shichen, et al.
Pubblicazione: (2024)
CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection
di: Chen, Qiyu, et al.
Pubblicazione: (2025)
di: Chen, Qiyu, et al.
Pubblicazione: (2025)
MRAD: Zero-Shot Anomaly Detection with Memory-Driven Retrieval
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
di: Xu, Chaoran, et al.
Pubblicazione: (2026)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
di: Deng, Huilin, et al.
Pubblicazione: (2024)
di: Deng, Huilin, et al.
Pubblicazione: (2024)
Center-aware Residual Anomaly Synthesis for Multi-class Industrial Anomaly Detection
di: Chen, Qiyu, et al.
Pubblicazione: (2025)
di: Chen, Qiyu, et al.
Pubblicazione: (2025)
DeltaDeno: Zero-Shot Anomaly Generation via Delta-Denoising Attribution
di: Xu, Chaoran, et al.
Pubblicazione: (2025)
di: Xu, Chaoran, et al.
Pubblicazione: (2025)
PromptMoE: Generalizable Zero-Shot Anomaly Detection via Visually-Guided Prompt Mixtures
di: Shao, Yuheng, et al.
Pubblicazione: (2025)
di: Shao, Yuheng, et al.
Pubblicazione: (2025)
VCP-CLIP: A visual context prompting model for zero-shot anomaly segmentation
di: Qu, Zhen, et al.
Pubblicazione: (2024)
di: Qu, Zhen, et al.
Pubblicazione: (2024)
A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis
di: Lin, Dongheng, et al.
Pubblicazione: (2025)
di: Lin, Dongheng, et al.
Pubblicazione: (2025)
AnomalyPainter: Vision-Language-Diffusion Synergy for Zero-Shot Realistic and Diverse Industrial Anomaly Synthesis
di: Lai, Zhangyu, et al.
Pubblicazione: (2025)
di: Lai, Zhangyu, et al.
Pubblicazione: (2025)
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation
di: Li, Shengze, et al.
Pubblicazione: (2024)
di: Li, Shengze, et al.
Pubblicazione: (2024)
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding
di: Bao, Xiaoyi, et al.
Pubblicazione: (2025)
di: Bao, Xiaoyi, et al.
Pubblicazione: (2025)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
di: Sarkar, Anindya, et al.
Pubblicazione: (2026)
di: Sarkar, Anindya, et al.
Pubblicazione: (2026)
Universal Features Guided Zero-Shot Category-Level Object Pose Estimation
di: Qu, Wentian, et al.
Pubblicazione: (2025)
di: Qu, Wentian, et al.
Pubblicazione: (2025)
Few-Shot Medical Image Segmentation with Large Kernel Attention
di: Wu, Xiaoxiao, et al.
Pubblicazione: (2024)
di: Wu, Xiaoxiao, et al.
Pubblicazione: (2024)
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
di: Xu, Jiacong, et al.
Pubblicazione: (2025)
di: Xu, Jiacong, et al.
Pubblicazione: (2025)
CoReS: Orchestrating the Dance of Reasoning and Segmentation
di: Bao, Xiaoyi, et al.
Pubblicazione: (2024)
di: Bao, Xiaoyi, et al.
Pubblicazione: (2024)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
di: Lou, Yuke, et al.
Pubblicazione: (2025)
di: Lou, Yuke, et al.
Pubblicazione: (2025)
Dual-Image Enhanced CLIP for Zero-Shot Anomaly Detection
di: Zhang, Zhaoxiang, et al.
Pubblicazione: (2024)
di: Zhang, Zhaoxiang, et al.
Pubblicazione: (2024)
Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial Inspection
di: Lu, Yilin, et al.
Pubblicazione: (2025)
di: Lu, Yilin, et al.
Pubblicazione: (2025)
CLIP-FSAC++: Few-Shot Anomaly Classification with Anomaly Descriptor Based on CLIP
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
di: Zuo, Zuo, et al.
Pubblicazione: (2024)
Can Large Language Models Grasp Event Signals? Exploring Pure Zero-Shot Event-based Recognition
di: Yu, Zongyou, et al.
Pubblicazione: (2024)
di: Yu, Zongyou, et al.
Pubblicazione: (2024)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
di: Hou, Yanning, et al.
Pubblicazione: (2026)
di: Hou, Yanning, et al.
Pubblicazione: (2026)
MuSc-V2: Zero-Shot Multimodal Industrial Anomaly Classification and Segmentation with Mutual Scoring of Unlabeled Samples
di: Li, Xurui, et al.
Pubblicazione: (2025)
di: Li, Xurui, et al.
Pubblicazione: (2025)
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion
di: Viola, Massimiliano, et al.
Pubblicazione: (2024)
di: Viola, Massimiliano, et al.
Pubblicazione: (2024)
Visual-Semantic Decomposition and Partial Alignment for Document-based Zero-Shot Learning
di: Qu, Xiangyan, et al.
Pubblicazione: (2024)
di: Qu, Xiangyan, et al.
Pubblicazione: (2024)
Zero-Shot Aerial Object Detection with Visual Description Regularization
di: Zang, Zhengqing, et al.
Pubblicazione: (2024)
di: Zang, Zhengqing, et al.
Pubblicazione: (2024)
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
di: Lin, Jiawen, et al.
Pubblicazione: (2025)
di: Lin, Jiawen, et al.
Pubblicazione: (2025)
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
di: Yang, Chengxu, et al.
Pubblicazione: (2026)
di: Yang, Chengxu, et al.
Pubblicazione: (2026)
Progressive Boundary Guided Anomaly Synthesis for Industrial Anomaly Detection
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
di: Chen, Qiyu, et al.
Pubblicazione: (2024)
Visual Anchors Are Strong Information Aggregators For Multimodal Large Language Model
di: Liu, Haogeng, et al.
Pubblicazione: (2024)
di: Liu, Haogeng, et al.
Pubblicazione: (2024)
SPG: Sparse-Projected Guides with Sparse Autoencoders for Zero-Shot Anomaly Detection
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning
di: Peng, Zhong, et al.
Pubblicazione: (2025)
di: Peng, Zhong, et al.
Pubblicazione: (2025)
TPCap: Unlocking Zero-Shot Image Captioning with Trigger-Augmented and Multi-Modal Purification Modules
di: Zhang, Ruoyu, et al.
Pubblicazione: (2025)
di: Zhang, Ruoyu, et al.
Pubblicazione: (2025)
Zero-Shot Anomaly Detection with Dual-Branch Prompt Selection
di: Wang, Zihan, et al.
Pubblicazione: (2025)
di: Wang, Zihan, et al.
Pubblicazione: (2025)
Investigating the Semantic Robustness of CLIP-based Zero-Shot Anomaly Segmentation
di: Stangl, Kevin, et al.
Pubblicazione: (2024)
di: Stangl, Kevin, et al.
Pubblicazione: (2024)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
di: Park, SoYoung, et al.
Pubblicazione: (2025)
di: Park, SoYoung, et al.
Pubblicazione: (2025)
Training Free Zero-Shot Visual Anomaly Localization via Diffusion Inversion
di: Hicsonmez, Samet, et al.
Pubblicazione: (2026)
di: Hicsonmez, Samet, et al.
Pubblicazione: (2026)
Documenti analoghi
-
DictAS: A Framework for Class-Generalizable Few-Shot Anomaly Segmentation via Dictionary Lookup
di: Qu, Zhen, et al.
Pubblicazione: (2025) -
Bayesian Prompt Flow Learning for Zero-Shot Anomaly Detection
di: Qu, Zhen, et al.
Pubblicazione: (2025) -
ALMRR: Anomaly Localization Mamba on Industrial Textured Surface with Feature Reconstruction and Refinement
di: Qu, Shichen, et al.
Pubblicazione: (2024) -
CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection
di: Chen, Qiyu, et al.
Pubblicazione: (2025) -
MRAD: Zero-Shot Anomaly Detection with Memory-Driven Retrieval
di: Xu, Chaoran, et al.
Pubblicazione: (2026)