IADGPT: Unified LVLM for Few-Shot Industrial Anomaly Detection, Localization, and Reasoning via In-Context Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Mengyang, Fu, Teng, Yu, Haiyang, Niu, Ke, Li, Bin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CReFT-CAD: Boosting Orthographic Projection Reasoning for CAD via Reinforcement Fine-Tuning
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
Provoking Multi-modal Few-Shot LVLM via Exploration-Exploitation In-Context Learning
by: Chen, Cheng, et al.
Published: (2025)
by: Chen, Cheng, et al.
Published: (2025)
From Intent to Execution: Multimodal Chain-of-Thought Reinforcement Learning for Precise CAD Code Generation
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
by: Ahn, Sunghyun, et al.
Published: (2025)
by: Ahn, Sunghyun, et al.
Published: (2025)
Interpretable Oracle Bone Script Decipherment through Radical and Pictographic Analysis with LVLMs
by: Peng, Kaixin, et al.
Published: (2025)
by: Peng, Kaixin, et al.
Published: (2025)
ChatReID: Open-ended Interactive Person Retrieval via Hierarchical Progressive Tuning for Vision Language Models
by: Niu, Ke, et al.
Published: (2025)
by: Niu, Ke, et al.
Published: (2025)
OmniPT: Unleashing the Potential of Large Vision Language Models for Pedestrian Tracking and Understanding
by: Fu, Teng, et al.
Published: (2025)
by: Fu, Teng, et al.
Published: (2025)
UMIT: Unifying Medical Imaging Tasks via Vision-Language Models
by: Yu, Haiyang, et al.
Published: (2025)
by: Yu, Haiyang, et al.
Published: (2025)
FastRef:Fast Prototype Refinement for Few-Shot Industrial Anomaly Detection
by: Tian, Long, et al.
Published: (2025)
by: Tian, Long, et al.
Published: (2025)
Prototypical Learning Guided Context-Aware Segmentation Network for Few-Shot Anomaly Detection
by: Jiang, Yuxin, et al.
Published: (2025)
by: Jiang, Yuxin, et al.
Published: (2025)
Few-Shot Anomaly Detection via Category-Agnostic Registration Learning
by: Huang, Chaoqin, et al.
Published: (2024)
by: Huang, Chaoqin, et al.
Published: (2024)
Kernel-Aware Graph Prompt Learning for Few-Shot Anomaly Detection
by: Tao, Fenfang, et al.
Published: (2024)
by: Tao, Fenfang, et al.
Published: (2024)
AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection
by: Zhang, Yi, et al.
Published: (2026)
by: Zhang, Yi, et al.
Published: (2026)
InCTRLv2: Generalist Residual Models for Few-Shot Anomaly Detection and Segmentation
by: Zhu, Jiawen, et al.
Published: (2026)
by: Zhu, Jiawen, et al.
Published: (2026)
GATE-AD: Graph Attention Network Encoding For Few-Shot Industrial Visual Anomaly Detection
by: Psiris, Aggelos, et al.
Published: (2026)
by: Psiris, Aggelos, et al.
Published: (2026)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
by: Zhao, Shifang, et al.
Published: (2025)
by: Zhao, Shifang, et al.
Published: (2025)
A Unified Anomaly Synthesis Strategy with Gradient Ascent for Industrial Anomaly Detection and Localization
by: Chen, Qiyu, et al.
Published: (2024)
by: Chen, Qiyu, et al.
Published: (2024)
Towards Fine-Grained Vision-Language Alignment for Few-Shot Anomaly Detection
by: Fan, Yuanting, et al.
Published: (2025)
by: Fan, Yuanting, et al.
Published: (2025)
Commonality in Few: Few-Shot Multimodal Anomaly Detection via Hypergraph-Enhanced Memory
by: Lin, Yuxuan, et al.
Published: (2025)
by: Lin, Yuxuan, et al.
Published: (2025)
Dual Distillation for Few-Shot Anomaly Detection
by: Dong, Le, et al.
Published: (2026)
by: Dong, Le, et al.
Published: (2026)
Transformer Based Self-Context Aware Prediction for Few-Shot Anomaly Detection in Videos
by: Pillai, Gargi V., et al.
Published: (2025)
by: Pillai, Gargi V., et al.
Published: (2025)
Unify the Views: View-Consistent Prototype Learning for Few-Shot Segmentation
by: Liu, Hongli, et al.
Published: (2026)
by: Liu, Hongli, et al.
Published: (2026)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
by: Hu, Teng, et al.
Published: (2023)
by: Hu, Teng, et al.
Published: (2023)
Generate Aligned Anomaly: Region-Guided Few-Shot Anomaly Image-Mask Pair Synthesis for Industrial Inspection
by: Lu, Yilin, et al.
Published: (2025)
by: Lu, Yilin, et al.
Published: (2025)
DocCogito: Aligning Layout Cognition and Step-Level Grounded Reasoning for Document Understanding
by: Wu, Yuchuan, et al.
Published: (2026)
by: Wu, Yuchuan, et al.
Published: (2026)
Few-Shot Anomaly-Driven Generation for Anomaly Classification and Segmentation
by: Gui, Guan, et al.
Published: (2025)
by: Gui, Guan, et al.
Published: (2025)
Text-Guided Multimodal Unified Industrial Anomaly Detection
by: Li, Zewen, et al.
Published: (2026)
by: Li, Zewen, et al.
Published: (2026)
ABounD: Adversarial Boundary-Driven Few-Shot Learning for Multi-Class Anomaly Detection
by: Deng, Runzhi, et al.
Published: (2025)
by: Deng, Runzhi, et al.
Published: (2025)
Synthesizing Efficient Data with Diffusion Models for Person Re-Identification Pre-Training
by: Niu, Ke, et al.
Published: (2024)
by: Niu, Ke, et al.
Published: (2024)
StackCLIP: Clustering-Driven Stacked Prompt in Zero-Shot Industrial Anomaly Detection
by: Hou, Yanning, et al.
Published: (2025)
by: Hou, Yanning, et al.
Published: (2025)
Few Shot Part Segmentation Reveals Compositional Logic for Industrial Anomaly Detection
by: Kim, Soopil, et al.
Published: (2023)
by: Kim, Soopil, et al.
Published: (2023)
On the Problem of Consistent Anomalies in Zero-Shot Industrial Anomaly Detection
by: Le-Gia, Tai, et al.
Published: (2025)
by: Le-Gia, Tai, et al.
Published: (2025)
AnoRefiner: Anomaly-Aware Group-Wise Refinement for Zero-Shot Industrial Anomaly Detection
by: Huang, Dayou, et al.
Published: (2025)
by: Huang, Dayou, et al.
Published: (2025)
Few-shot Human Action Anomaly Detection via a Unified Contrastive Learning Framework
by: Kamide, Koichiro, et al.
Published: (2025)
by: Kamide, Koichiro, et al.
Published: (2025)
Beyond Normal References: Discriminative Few-Shot Anomaly Detection
by: Wang, Huan, et al.
Published: (2026)
by: Wang, Huan, et al.
Published: (2026)
POPEN: Preference-Based Optimization and Ensemble for LVLM-Based Reasoning Segmentation
by: Zhu, Lanyun, et al.
Published: (2025)
by: Zhu, Lanyun, et al.
Published: (2025)
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
by: Gu, Zhaopeng, et al.
Published: (2025)
by: Gu, Zhaopeng, et al.
Published: (2025)
Few-Shot Object Detection with Sparse Context Transformers
by: Mei, Jie, et al.
Published: (2024)
by: Mei, Jie, et al.
Published: (2024)
PromptAD: Learning Prompts with only Normal Samples for Few-Shot Anomaly Detection
by: Li, Xiaofan, et al.
Published: (2024)
by: Li, Xiaofan, et al.
Published: (2024)
EVE: Towards End-to-End Video Subtitle Extraction with Vision-Language Models
by: Yu, Haiyang, et al.
Published: (2025)
by: Yu, Haiyang, et al.
Published: (2025)
Similar Items
-
CReFT-CAD: Boosting Orthographic Projection Reasoning for CAD via Reinforcement Fine-Tuning
by: Niu, Ke, et al.
Published: (2025) -
Provoking Multi-modal Few-Shot LVLM via Exploration-Exploitation In-Context Learning
by: Chen, Cheng, et al.
Published: (2025) -
From Intent to Execution: Multimodal Chain-of-Thought Reinforcement Learning for Precise CAD Code Generation
by: Niu, Ke, et al.
Published: (2025) -
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
by: Ahn, Sunghyun, et al.
Published: (2025) -
Interpretable Oracle Bone Script Decipherment through Radical and Pictographic Analysis with LVLMs
by: Peng, Kaixin, et al.
Published: (2025)