Text-Guided Multimodal Unified Industrial Anomaly Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Zewen, Ye, Shuo, Yu, Zitong, Xie, Weicheng, Shen, Linlin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
IAD-GPT: Advancing Visual Knowledge in Multimodal Large Language Model for Industrial Anomaly Detection
por: Li, Zewen, et al.
Publicado: (2025)
por: Li, Zewen, et al.
Publicado: (2025)
BIG-MoE: Bypass Isolated Gating MoE for Generalized Multimodal Face Anti-Spoofing
por: Ma, Yingjie, et al.
Publicado: (2024)
por: Ma, Yingjie, et al.
Publicado: (2024)
SFDA-rPPG: Source-Free Domain Adaptive Remote Physiological Measurement with Spatio-Temporal Consistency
por: Xie, Yiping, et al.
Publicado: (2024)
por: Xie, Yiping, et al.
Publicado: (2024)
Denoising and Alignment: Rethinking Domain Generalization for Multimodal Face Anti-Spoofing
por: Ma, Yingjie, et al.
Publicado: (2025)
por: Ma, Yingjie, et al.
Publicado: (2025)
Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling
por: Niu, Zenghao, et al.
Publicado: (2025)
por: Niu, Zenghao, et al.
Publicado: (2025)
Distilled Transformers with Locally Enhanced Global Representations for Face Forgery Detection
por: Zhang, Yaning, et al.
Publicado: (2024)
por: Zhang, Yaning, et al.
Publicado: (2024)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
por: Zhang, Yaning, et al.
Publicado: (2026)
por: Zhang, Yaning, et al.
Publicado: (2026)
Tuned Reverse Distillation: Enhancing Multimodal Industrial Anomaly Detection with Crossmodal Tuners
por: Liu, Xinyue, et al.
Publicado: (2024)
por: Liu, Xinyue, et al.
Publicado: (2024)
BridgeNet: A Unified Multimodal Framework for Bridging 2D and 3D Industrial Anomaly Detection
por: Xiang, An, et al.
Publicado: (2025)
por: Xiang, An, et al.
Publicado: (2025)
Multimodal Industrial Anomaly Detection via Geometric Prior
por: Li, Min, et al.
Publicado: (2026)
por: Li, Min, et al.
Publicado: (2026)
Can Multimodal Large Language Models be Guided to Improve Industrial Anomaly Detection?
por: Chen, Zhiling, et al.
Publicado: (2025)
por: Chen, Zhiling, et al.
Publicado: (2025)
Multimodal Fake News Detection: MFND Dataset and Shallow-Deep Multitask Learning
por: Zhu, Ye, et al.
Publicado: (2025)
por: Zhu, Ye, et al.
Publicado: (2025)
PhysLLM: Harnessing Large Language Models for Cross-Modal Remote Physiological Sensing
por: Xie, Yiping, et al.
Publicado: (2025)
por: Xie, Yiping, et al.
Publicado: (2025)
GM-DF: Generalized Multi-Scenario Deepfake Detection
por: Lai, Yingxin, et al.
Publicado: (2024)
por: Lai, Yingxin, et al.
Publicado: (2024)
PA-FAS: Towards Interpretable and Generalizable Multimodal Face Anti-Spoofing via Path-Augmented Reinforcement Learning
por: Ma, Yingjie, et al.
Publicado: (2025)
por: Ma, Yingjie, et al.
Publicado: (2025)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
por: Lee, Mingyu, et al.
Publicado: (2024)
por: Lee, Mingyu, et al.
Publicado: (2024)
Towards Explainable Industrial Anomaly Detection via Knowledge-Guided Latent Reasoning
por: Chen, Peng, et al.
Publicado: (2026)
por: Chen, Peng, et al.
Publicado: (2026)
Multimodal Industrial Anomaly Detection by Crossmodal Feature Mapping
por: Costanzino, Alex, et al.
Publicado: (2023)
por: Costanzino, Alex, et al.
Publicado: (2023)
A Unified Anomaly Synthesis Strategy with Gradient Ascent for Industrial Anomaly Detection and Localization
por: Chen, Qiyu, et al.
Publicado: (2024)
por: Chen, Qiyu, et al.
Publicado: (2024)
CA-Edit: Causality-Aware Condition Adapter for High-Fidelity Local Facial Attribute Editing
por: Xian, Xiaole, et al.
Publicado: (2024)
por: Xian, Xiaole, et al.
Publicado: (2024)
Self-Navigated Residual Mamba for Universal Industrial Anomaly Detection
por: Li, Hanxi, et al.
Publicado: (2025)
por: Li, Hanxi, et al.
Publicado: (2025)
UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning
por: Li, Hongrui, et al.
Publicado: (2026)
por: Li, Hongrui, et al.
Publicado: (2026)
YUV20K: A Complexity-Driven Benchmark and Trajectory-Aware Alignment Model for Video Camouflaged Object Detection
por: Liu, Yiyu, et al.
Publicado: (2026)
por: Liu, Yiyu, et al.
Publicado: (2026)
Towards an Incremental Unified Multimodal Anomaly Detection: Augmenting Multimodal Denoising From an Information Bottleneck Perspective
por: Long, Kaifang, et al.
Publicado: (2026)
por: Long, Kaifang, et al.
Publicado: (2026)
IADGPT: Unified LVLM for Few-Shot Industrial Anomaly Detection, Localization, and Reasoning via In-Context Learning
por: Zhao, Mengyang, et al.
Publicado: (2025)
por: Zhao, Mengyang, et al.
Publicado: (2025)
Not All Regions Are Equal: Attention-Guided Perturbation Network for Industrial Anomaly Detection
por: Huang, Tingfeng, et al.
Publicado: (2024)
por: Huang, Tingfeng, et al.
Publicado: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
por: Zhang, Yaning, et al.
Publicado: (2024)
por: Zhang, Yaning, et al.
Publicado: (2024)
Agent4FaceForgery: Multi-Agent LLM Framework for Realistic Face Forgery Detection
por: Lai, Yingxin, et al.
Publicado: (2025)
por: Lai, Yingxin, et al.
Publicado: (2025)
Progressive Boundary Guided Anomaly Synthesis for Industrial Anomaly Detection
por: Chen, Qiyu, et al.
Publicado: (2024)
por: Chen, Qiyu, et al.
Publicado: (2024)
Answering Diverse Questions via Text Attached with Key Audio-Visual Clues
por: Ye, Qilang, et al.
Publicado: (2024)
por: Ye, Qilang, et al.
Publicado: (2024)
Deep Industrial Image Anomaly Detection: A Survey
por: Liu, Jiaqi, et al.
Publicado: (2023)
por: Liu, Jiaqi, et al.
Publicado: (2023)
PatchEAD: Unifying Industrial Visual Prompting Frameworks for Patch-Exclusive Anomaly Detection
por: Huang, Po-Han, et al.
Publicado: (2025)
por: Huang, Po-Han, et al.
Publicado: (2025)
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
por: Li, Yuanze, et al.
Publicado: (2023)
por: Li, Yuanze, et al.
Publicado: (2023)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
por: Zhao, Shifang, et al.
Publicado: (2025)
por: Zhao, Shifang, et al.
Publicado: (2025)
Incomplete Multimodal Industrial Anomaly Detection via Cross-Modal Distillation
por: Sui, Wenbo, et al.
Publicado: (2024)
por: Sui, Wenbo, et al.
Publicado: (2024)
VTFusion: A Vision-Text Multimodal Fusion Network for Few-Shot Anomaly Detection
por: Jiang, Yuxin, et al.
Publicado: (2026)
por: Jiang, Yuxin, et al.
Publicado: (2026)
SPF-Portrait: Towards Pure Text-to-Portrait Customization with Semantic Pollution-Free Fine-Tuning
por: Xian, Xiaole, et al.
Publicado: (2025)
por: Xian, Xiaole, et al.
Publicado: (2025)
Towards Efficient Pixel Labeling for Industrial Anomaly Detection and Localization
por: Li, Hanxi, et al.
Publicado: (2024)
por: Li, Hanxi, et al.
Publicado: (2024)
Distribution-Specific Learning for Joint Salient and Camouflaged Object Detection
por: Hao, Chao, et al.
Publicado: (2025)
por: Hao, Chao, et al.
Publicado: (2025)
AnoRefiner: Anomaly-Aware Group-Wise Refinement for Zero-Shot Industrial Anomaly Detection
por: Huang, Dayou, et al.
Publicado: (2025)
por: Huang, Dayou, et al.
Publicado: (2025)
Ejemplares similares
-
IAD-GPT: Advancing Visual Knowledge in Multimodal Large Language Model for Industrial Anomaly Detection
por: Li, Zewen, et al.
Publicado: (2025) -
BIG-MoE: Bypass Isolated Gating MoE for Generalized Multimodal Face Anti-Spoofing
por: Ma, Yingjie, et al.
Publicado: (2024) -
SFDA-rPPG: Source-Free Domain Adaptive Remote Physiological Measurement with Spatio-Temporal Consistency
por: Xie, Yiping, et al.
Publicado: (2024) -
Denoising and Alignment: Rethinking Domain Generalization for Multimodal Face Anti-Spoofing
por: Ma, Yingjie, et al.
Publicado: (2025) -
Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling
por: Niu, Zenghao, et al.
Publicado: (2025)