OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Shifang, Lin, Yiheng, Han, Lu, Zhao, Yao, Wei, Yunchao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AlignGen: Boosting Personalized Image Generation with Cross-Modality Prior Alignment
by: Lin, Yiheng, et al.
Published: (2025)
by: Lin, Yiheng, et al.
Published: (2025)
Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
by: Tan, Chuangchuang, et al.
Published: (2025)
by: Tan, Chuangchuang, et al.
Published: (2025)
CutClaw: Agentic Hours-Long Video Editing via Music Synchronization
by: Zhao, Shifang, et al.
Published: (2026)
by: Zhao, Shifang, et al.
Published: (2026)
AD-FM: Multimodal LLMs for Anomaly Detection via Multi-Stage Reasoning and Fine-Grained Reward Optimization
by: Liao, Jingyi, et al.
Published: (2025)
by: Liao, Jingyi, et al.
Published: (2025)
Multimodal Industrial Anomaly Detection via Geometric Prior
by: Li, Min, et al.
Published: (2026)
by: Li, Min, et al.
Published: (2026)
Diffusion for Natural Image Matting
by: Hu, Yihan, et al.
Published: (2023)
by: Hu, Yihan, et al.
Published: (2023)
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
by: Yang, Qize, et al.
Published: (2025)
by: Yang, Qize, et al.
Published: (2025)
ResAD++: Towards Class Agnostic Anomaly Detection via Residual Feature Learning
by: Yao, Xincheng, et al.
Published: (2025)
by: Yao, Xincheng, et al.
Published: (2025)
PixelLM: Pixel Reasoning with Large Multimodal Model
by: Ren, Zhongwei, et al.
Published: (2023)
by: Ren, Zhongwei, et al.
Published: (2023)
Omni-AD: Learning to Reconstruct Global and Local Features for Multi-class Anomaly Detection
by: Quan, Jiajie, et al.
Published: (2025)
by: Quan, Jiajie, et al.
Published: (2025)
A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis
by: Lin, Dongheng, et al.
Published: (2025)
by: Lin, Dongheng, et al.
Published: (2025)
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
by: Tan, Chuangchuang, et al.
Published: (2025)
by: Tan, Chuangchuang, et al.
Published: (2025)
DCEdit: Dual-Level Controlled Image Editing via Precisely Localized Semantics
by: Hu, Yihan, et al.
Published: (2025)
by: Hu, Yihan, et al.
Published: (2025)
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
by: Wang, Zehan, et al.
Published: (2024)
by: Wang, Zehan, et al.
Published: (2024)
ADPretrain: Advancing Industrial Anomaly Detection via Anomaly Representation Pretraining
by: Yao, Xincheng, et al.
Published: (2025)
by: Yao, Xincheng, et al.
Published: (2025)
VarAD: Lightweight High-Resolution Image Anomaly Detection via Visual Autoregressive Modeling
by: Cao, Yunkang, et al.
Published: (2024)
by: Cao, Yunkang, et al.
Published: (2024)
IADGPT: Unified LVLM for Few-Shot Industrial Anomaly Detection, Localization, and Reasoning via In-Context Learning
by: Zhao, Mengyang, et al.
Published: (2025)
by: Zhao, Mengyang, et al.
Published: (2025)
Collaborative Feature-Logits Contrastive Learning for Open-Set Semi-Supervised Object Detection
by: Zhong, Xinhao, et al.
Published: (2024)
by: Zhong, Xinhao, et al.
Published: (2024)
CXR-AD: Component X-ray Image Dataset for Industrial Anomaly Detection
by: Bai, Haoyu, et al.
Published: (2025)
by: Bai, Haoyu, et al.
Published: (2025)
ToCoAD: Two-Stage Contrastive Learning for Industrial Anomaly Detection
by: Liang, Yun, et al.
Published: (2024)
by: Liang, Yun, et al.
Published: (2024)
Towards Explainable Industrial Anomaly Detection via Knowledge-Guided Latent Reasoning
by: Chen, Peng, et al.
Published: (2026)
by: Chen, Peng, et al.
Published: (2026)
AD-Copilot: A Vision-Language Assistant for Industrial Anomaly Detection via Visual In-context Comparison
by: Jiang, Xi, et al.
Published: (2026)
by: Jiang, Xi, et al.
Published: (2026)
Incomplete Multimodal Industrial Anomaly Detection via Cross-Modal Distillation
by: Sui, Wenbo, et al.
Published: (2024)
by: Sui, Wenbo, et al.
Published: (2024)
LAD-Reasoner: Tiny Multimodal Models are Good Reasoners for Logical Anomaly Detection
by: Li, Weijia, et al.
Published: (2025)
by: Li, Weijia, et al.
Published: (2025)
Template-Based Feature Aggregation Network for Industrial Anomaly Detection
by: Luo, Wei, et al.
Published: (2026)
by: Luo, Wei, et al.
Published: (2026)
DreamLCM: Towards High-Quality Text-to-3D Generation via Latent Consistency Model
by: Zhong, Yiming, et al.
Published: (2024)
by: Zhong, Yiming, et al.
Published: (2024)
Myriad: Large Multimodal Model by Applying Vision Experts for Industrial Anomaly Detection
by: Li, Yuanze, et al.
Published: (2023)
by: Li, Yuanze, et al.
Published: (2023)
Component-aware Unsupervised Logical Anomaly Generation for Industrial Anomaly Detection
by: Tong, Xuan, et al.
Published: (2025)
by: Tong, Xuan, et al.
Published: (2025)
KairosAD: A SAM-Based Model for Industrial Anomaly Detection on Embedded Devices
by: Khan, Uzair, et al.
Published: (2025)
by: Khan, Uzair, et al.
Published: (2025)
Memory Efficient Matting with Adaptive Token Routing
by: Lin, Yiheng, et al.
Published: (2024)
by: Lin, Yiheng, et al.
Published: (2024)
Text-Guided Multimodal Unified Industrial Anomaly Detection
by: Li, Zewen, et al.
Published: (2026)
by: Li, Zewen, et al.
Published: (2026)
Multimodal Industrial Anomaly Detection by Crossmodal Feature Mapping
by: Costanzino, Alex, et al.
Published: (2023)
by: Costanzino, Alex, et al.
Published: (2023)
AD-Reasoning: Multimodal Guideline-Guided Reasoning for Alzheimer's Disease Diagnosis
by: Chen, Qiuhui, et al.
Published: (2026)
by: Chen, Qiuhui, et al.
Published: (2026)
R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
by: Zhao, Jiaxing, et al.
Published: (2025)
by: Zhao, Jiaxing, et al.
Published: (2025)
Center-aware Residual Anomaly Synthesis for Multi-class Industrial Anomaly Detection
by: Chen, Qiyu, et al.
Published: (2025)
by: Chen, Qiyu, et al.
Published: (2025)
GATE-AD: Graph Attention Network Encoding For Few-Shot Industrial Visual Anomaly Detection
by: Psiris, Aggelos, et al.
Published: (2026)
by: Psiris, Aggelos, et al.
Published: (2026)
ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and Understanding
by: Ye, Junliang, et al.
Published: (2025)
by: Ye, Junliang, et al.
Published: (2025)
OmniScene: Attention-Augmented Multimodal 4D Scene Understanding for Autonomous Driving
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
Weakly Supervised Video Anomaly Detection with Anomaly-Connected Components and Intention Reasoning
by: Wang, Yu, et al.
Published: (2026)
by: Wang, Yu, et al.
Published: (2026)
IPAD: Industrial Process Anomaly Detection Dataset
by: Liu, Jinfan, et al.
Published: (2024)
by: Liu, Jinfan, et al.
Published: (2024)
Similar Items
-
AlignGen: Boosting Personalized Image Generation with Cross-Modality Prior Alignment
by: Lin, Yiheng, et al.
Published: (2025) -
Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
by: Tan, Chuangchuang, et al.
Published: (2025) -
CutClaw: Agentic Hours-Long Video Editing via Music Synchronization
by: Zhao, Shifang, et al.
Published: (2026) -
AD-FM: Multimodal LLMs for Anomaly Detection via Multi-Stage Reasoning and Fine-Grained Reward Optimization
by: Liao, Jingyi, et al.
Published: (2025) -
Multimodal Industrial Anomaly Detection via Geometric Prior
by: Li, Min, et al.
Published: (2026)