Saved in:
| Main Authors: | Hu, Zongxiang, Zhang, Zhaosheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.03634 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adapting Visual-Language Models for Generalizable Anomaly Detection in Medical Images
by: Huang, Chaoqin, et al.
Published: (2024)
by: Huang, Chaoqin, et al.
Published: (2024)
AdaptCLIP: Adapting CLIP for Universal Visual Anomaly Detection
by: Gao, Bin-Bin, et al.
Published: (2025)
by: Gao, Bin-Bin, et al.
Published: (2025)
Anomaly Detection by Adapting a pre-trained Vision Language Model
by: Cai, Yuxuan, et al.
Published: (2024)
by: Cai, Yuxuan, et al.
Published: (2024)
Language Models Meet Anomaly Detection for Better Interpretability and Generalizability
by: Li, Jun, et al.
Published: (2024)
by: Li, Jun, et al.
Published: (2024)
Advancing Generalizable Tumor Segmentation with Anomaly-Aware Open-Vocabulary Attention Maps and Frozen Foundation Diffusion Models
by: Jiang, Yankai, et al.
Published: (2025)
by: Jiang, Yankai, et al.
Published: (2025)
Hierarchical Attention for Sparse Volumetric Anomaly Detection in Subclinical Keratoconus
by: Kandakji, Lynn, et al.
Published: (2025)
by: Kandakji, Lynn, et al.
Published: (2025)
Self-Adapting Large Visual-Language Models to Edge Devices across Visual Modalities
by: Cai, Kaiwen, et al.
Published: (2024)
by: Cai, Kaiwen, et al.
Published: (2024)
From CNN to CNN + RNN: Adapting Visualization Techniques for Time-Series Anomaly Detection
by: Poirier, Fabien
Published: (2024)
by: Poirier, Fabien
Published: (2024)
Beyond Text: Frozen Large Language Models in Visual Signal Comprehension
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
Anomize: Better Open Vocabulary Video Anomaly Detection
by: Li, Fei, et al.
Published: (2025)
by: Li, Fei, et al.
Published: (2025)
MonoSOWA: Scalable monocular 3D Object detector Without human Annotations
by: Skvrna, Jan, et al.
Published: (2025)
by: Skvrna, Jan, et al.
Published: (2025)
MIRAGE: Model-agnostic Industrial Realistic Anomaly Generation and Evaluation for Visual Anomaly Detection
by: Hu, Jinwei, et al.
Published: (2026)
by: Hu, Jinwei, et al.
Published: (2026)
Cross-level Attention with Overlapped Windows for Camouflaged Object Detection
by: Li, Jiepan, et al.
Published: (2023)
by: Li, Jiepan, et al.
Published: (2023)
Customizing Visual-Language Foundation Models for Multi-modal Anomaly Detection and Reasoning
by: Xu, Xiaohao, et al.
Published: (2024)
by: Xu, Xiaohao, et al.
Published: (2024)
Steering and Rectifying Latent Representation Manifolds in Frozen Multi-modal LLMs for Video Anomaly Detection
by: Cai, Zhaolin, et al.
Published: (2026)
by: Cai, Zhaolin, et al.
Published: (2026)
AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly Detection
by: Gu, Zhaopeng, et al.
Published: (2025)
by: Gu, Zhaopeng, et al.
Published: (2025)
MediCLIP: Adapting CLIP for Few-shot Medical Image Anomaly Detection
by: Zhang, Ximiao, et al.
Published: (2024)
by: Zhang, Ximiao, et al.
Published: (2024)
Chat-CBM: Towards Interactive Concept Bottleneck Models with Frozen Large Language Models
by: He, Hangzhou, et al.
Published: (2025)
by: He, Hangzhou, et al.
Published: (2025)
Window Token Concatenation for Efficient Visual Large Language Models
by: Li, Yifan, et al.
Published: (2025)
by: Li, Yifan, et al.
Published: (2025)
Breaking the Bias: Recalibrating the Attention of Industrial Anomaly Detection
by: Chen, Xin, et al.
Published: (2024)
by: Chen, Xin, et al.
Published: (2024)
CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision
by: Khadka, Puskal, et al.
Published: (2025)
by: Khadka, Puskal, et al.
Published: (2025)
Referring Camouflaged Object Detection With Multi-Context Overlapped Windows Cross-Attention
by: Wen, Yu, et al.
Published: (2025)
by: Wen, Yu, et al.
Published: (2025)
Hierarchical Windowed Graph Attention Network and a Large Scale Dataset for Isolated Indian Sign Language Recognition
by: Patra, Suvajit, et al.
Published: (2024)
by: Patra, Suvajit, et al.
Published: (2024)
YOLO-FDA: Integrating Hierarchical Attention and Detail Enhancement for Surface Defect Detection
by: Hu, Jiawei
Published: (2025)
by: Hu, Jiawei
Published: (2025)
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
by: Cao, Yunkang, et al.
Published: (2024)
by: Cao, Yunkang, et al.
Published: (2024)
Self-Supervised Training with Autoencoders for Visual Anomaly Detection
by: Bauer, Alexander, et al.
Published: (2022)
by: Bauer, Alexander, et al.
Published: (2022)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
by: Deng, Huilin, et al.
Published: (2024)
by: Deng, Huilin, et al.
Published: (2024)
GATE-AD: Graph Attention Network Encoding For Few-Shot Industrial Visual Anomaly Detection
by: Psiris, Aggelos, et al.
Published: (2026)
by: Psiris, Aggelos, et al.
Published: (2026)
AtrousMamaba: An Atrous-Window Scanning Visual State Space Model for Remote Sensing Change Detection
by: Wang, Tao, et al.
Published: (2025)
by: Wang, Tao, et al.
Published: (2025)
GLAD: Towards Better Reconstruction with Global and Local Adaptive Diffusion Models for Unsupervised Anomaly Detection
by: Yao, Hang, et al.
Published: (2024)
by: Yao, Hang, et al.
Published: (2024)
Attention Fusion Reverse Distillation for Multi-Lighting Image Anomaly Detection
by: Zhang, Yiheng, et al.
Published: (2024)
by: Zhang, Yiheng, et al.
Published: (2024)
Topo-R1: Detecting Topological Anomalies via Vision-Language Models
by: Xu, Meilong, et al.
Published: (2026)
by: Xu, Meilong, et al.
Published: (2026)
TAU-R1: Visual Language Model for Traffic Anomaly Understanding
by: Lin, Yuqiang, et al.
Published: (2026)
by: Lin, Yuqiang, et al.
Published: (2026)
DINO-AD: Unsupervised Anomaly Detection with Frozen DINO-V3 Features
by: Huo, Jiayu, et al.
Published: (2026)
by: Huo, Jiayu, et al.
Published: (2026)
Improving Anomaly Detection with Foundation-Model Synthesis and Wavelet-Domain Attention
by: Wu, Wensheng, et al.
Published: (2026)
by: Wu, Wensheng, et al.
Published: (2026)
IAD-GPT: Advancing Visual Knowledge in Multimodal Large Language Model for Industrial Anomaly Detection
by: Li, Zewen, et al.
Published: (2025)
by: Li, Zewen, et al.
Published: (2025)
Frozen Transformers in Language Models Are Effective Visual Encoder Layers
by: Pang, Ziqi, et al.
Published: (2023)
by: Pang, Ziqi, et al.
Published: (2023)
Seeing Is Believing? A Benchmark for Multimodal Large Language Models on Visual Illusions and Anomalies
by: Hou, Wenjin, et al.
Published: (2026)
by: Hou, Wenjin, et al.
Published: (2026)
Hierarchical Gaussian Mixture Normalizing Flow Modeling for Unified Anomaly Detection
by: Yao, Xincheng, et al.
Published: (2024)
by: Yao, Xincheng, et al.
Published: (2024)
FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Similar Items
-
Adapting Visual-Language Models for Generalizable Anomaly Detection in Medical Images
by: Huang, Chaoqin, et al.
Published: (2024) -
AdaptCLIP: Adapting CLIP for Universal Visual Anomaly Detection
by: Gao, Bin-Bin, et al.
Published: (2025) -
Anomaly Detection by Adapting a pre-trained Vision Language Model
by: Cai, Yuxuan, et al.
Published: (2024) -
Language Models Meet Anomaly Detection for Better Interpretability and Generalizability
by: Li, Jun, et al.
Published: (2024) -
Advancing Generalizable Tumor Segmentation with Anomaly-Aware Open-Vocabulary Attention Maps and Frozen Foundation Diffusion Models
by: Jiang, Yankai, et al.
Published: (2025)