Quality-Aware Language-Conditioned Local Auto-Regressive Anomaly Synthesis and Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Qian, Long, Zhu, Bingke, Chen, Yingying, Tang, Ming, Wang, Jinqiao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MathPhys-Guided Coarse-to-Fine Anomaly Synthesis with SQE-Driven Bi-Level Optimization for Anomaly Detection
por: Qian, Long, et al.
Publicado: (2025)
por: Qian, Long, et al.
Publicado: (2025)
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
por: Gu, Zhaopeng, et al.
Publicado: (2025)
por: Gu, Zhaopeng, et al.
Publicado: (2025)
FiLo: Zero-Shot Anomaly Detection by Fine-Grained Description and High-Quality Localization
por: Gu, Zhaopeng, et al.
Publicado: (2024)
por: Gu, Zhaopeng, et al.
Publicado: (2024)
AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly Detection
por: Gu, Zhaopeng, et al.
Publicado: (2025)
por: Gu, Zhaopeng, et al.
Publicado: (2025)
UniVAD: A Training-free Unified Model for Few-shot Visual Anomaly Detection
por: Gu, Zhaopeng, et al.
Publicado: (2024)
por: Gu, Zhaopeng, et al.
Publicado: (2024)
Optimization of Prompt Learning via Multi-Knowledge Representation for Vision-Language Models
por: Zhang, Enming, et al.
Publicado: (2024)
por: Zhang, Enming, et al.
Publicado: (2024)
LINK: Adaptive Modality Interaction for Audio-Visual Video Parsing
por: Wang, Langyu, et al.
Publicado: (2024)
por: Wang, Langyu, et al.
Publicado: (2024)
MROVSeg: Breaking the Resolution Curse of Vision-Language Models in Open-Vocabulary Image Segmentation
por: Zhu, Yuanbing, et al.
Publicado: (2024)
por: Zhu, Yuanbing, et al.
Publicado: (2024)
MUG: Pseudo Labeling Augmented Audio-Visual Mamba Network for Audio-Visual Video Parsing
por: Wang, Langyu, et al.
Publicado: (2025)
por: Wang, Langyu, et al.
Publicado: (2025)
Semantic Noise Reduction via Teacher-Guided Dual-Path Audio-Visual Representation Learning
por: Wang, Linge, et al.
Publicado: (2026)
por: Wang, Linge, et al.
Publicado: (2026)
Monocular Lane Detection Based on Deep Learning: A Survey
por: He, Xin, et al.
Publicado: (2024)
por: He, Xin, et al.
Publicado: (2024)
AAformer: Auto-Aligned Transformer for Person Re-Identification
por: Zhu, Kuan, et al.
Publicado: (2021)
por: Zhu, Kuan, et al.
Publicado: (2021)
FAIR: Frequency-aware Image Restoration for Industrial Visual Anomaly Detection
por: Liu, Tongkun, et al.
Publicado: (2023)
por: Liu, Tongkun, et al.
Publicado: (2023)
TraceVision: Trajectory-Aware Vision-Language Model for Human-Like Spatial Understanding
por: Yang, Fan, et al.
Publicado: (2026)
por: Yang, Fan, et al.
Publicado: (2026)
Auto DragGAN: Editing the Generative Image Manifold in an Autoregressive Manner
por: Cai, Pengxiang, et al.
Publicado: (2024)
por: Cai, Pengxiang, et al.
Publicado: (2024)
Friend or Foe? Harnessing Controllable Overfitting for Anomaly Detection
por: Qian, Long, et al.
Publicado: (2024)
por: Qian, Long, et al.
Publicado: (2024)
Griffon-G: Bridging Vision-Language and Vision-Centric Tasks via Large Multimodal Models
por: Zhan, Yufei, et al.
Publicado: (2024)
por: Zhan, Yufei, et al.
Publicado: (2024)
A Unified Anomaly Synthesis Strategy with Gradient Ascent for Industrial Anomaly Detection and Localization
por: Chen, Qiyu, et al.
Publicado: (2024)
por: Chen, Qiyu, et al.
Publicado: (2024)
CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection
por: Chen, Qiyu, et al.
Publicado: (2025)
por: Chen, Qiyu, et al.
Publicado: (2025)
PA-CLIP: Enhancing Zero-Shot Anomaly Detection through Pseudo-Anomaly Awareness
por: Pan, Yurui, et al.
Publicado: (2025)
por: Pan, Yurui, et al.
Publicado: (2025)
Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models
por: Ma, Kexin, et al.
Publicado: (2026)
por: Ma, Kexin, et al.
Publicado: (2026)
Griffon: Spelling out All Object Locations at Any Granularity with Large Language Models
por: Zhan, Yufei, et al.
Publicado: (2023)
por: Zhan, Yufei, et al.
Publicado: (2023)
HaloAE: An HaloNet based Local Transformer Auto-Encoder for Anomaly Detection and Localization
por: Mathian, E., et al.
Publicado: (2022)
por: Mathian, E., et al.
Publicado: (2022)
EventVAD: Training-Free Event-Aware Video Anomaly Detection
por: Shao, Yihua, et al.
Publicado: (2025)
por: Shao, Yihua, et al.
Publicado: (2025)
From Horizontal to Rotated: Cross-View Object Geo-Localization with Orientation Awareness
por: Fu, Chenlin, et al.
Publicado: (2026)
por: Fu, Chenlin, et al.
Publicado: (2026)
Boosting Global-Local Feature Matching via Anomaly Synthesis for Multi-Class Point Cloud Anomaly Detection
por: Cheng, Yuqi, et al.
Publicado: (2025)
por: Cheng, Yuqi, et al.
Publicado: (2025)
DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
por: Zhao, Ruowen, et al.
Publicado: (2025)
por: Zhao, Ruowen, et al.
Publicado: (2025)
MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis
por: He, Wanggui, et al.
Publicado: (2024)
por: He, Wanggui, et al.
Publicado: (2024)
ASBench: Image Anomalies Synthesis Benchmark for Anomaly Detection
por: Zhang, Qunyi, et al.
Publicado: (2025)
por: Zhang, Qunyi, et al.
Publicado: (2025)
From Seeing to Predicting: A Vision-Language Framework for Trajectory Forecasting and Controlled Video Generation
por: Yang, Fan, et al.
Publicado: (2025)
por: Yang, Fan, et al.
Publicado: (2025)
Exploring Large Vision-Language Models for Robust and Efficient Industrial Anomaly Detection
por: Qian, Kun, et al.
Publicado: (2024)
por: Qian, Kun, et al.
Publicado: (2024)
Semantic-Deviation-Anchored Multi-Branch Fusion for Unsupervised Anomaly Detection and Localization in Unstructured Conveyor-Belt Coal Scenes
por: Jin, Wenping, et al.
Publicado: (2026)
por: Jin, Wenping, et al.
Publicado: (2026)
FOCUS: Unified Vision-Language Modeling for Interactive Editing Driven by Referential Segmentation
por: Yang, Fan, et al.
Publicado: (2025)
por: Yang, Fan, et al.
Publicado: (2025)
ARCON: Advancing Auto-Regressive Continuation for Driving Videos
por: Ming, Ruibo, et al.
Publicado: (2024)
por: Ming, Ruibo, et al.
Publicado: (2024)
FOCUS: Fine-grained Optimization with Semantic Guided Understanding for Pedestrian Attributes Recognition
por: An, Hongyan, et al.
Publicado: (2025)
por: An, Hongyan, et al.
Publicado: (2025)
AR-Diffusion: Asynchronous Video Generation with Auto-Regressive Diffusion
por: Sun, Mingzhen, et al.
Publicado: (2025)
por: Sun, Mingzhen, et al.
Publicado: (2025)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
por: Han, Jian, et al.
Publicado: (2024)
por: Han, Jian, et al.
Publicado: (2024)
Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
por: Chen, Zigeng, et al.
Publicado: (2024)
por: Chen, Zigeng, et al.
Publicado: (2024)
Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction
por: Chen, Junyi, et al.
Publicado: (2024)
por: Chen, Junyi, et al.
Publicado: (2024)
VAR-CLIP: Text-to-Image Generator with Visual Auto-Regressive Modeling
por: Zhang, Qian, et al.
Publicado: (2024)
por: Zhang, Qian, et al.
Publicado: (2024)
Ejemplares similares
-
MathPhys-Guided Coarse-to-Fine Anomaly Synthesis with SQE-Driven Bi-Level Optimization for Anomaly Detection
por: Qian, Long, et al.
Publicado: (2025) -
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
por: Gu, Zhaopeng, et al.
Publicado: (2025) -
FiLo: Zero-Shot Anomaly Detection by Fine-Grained Description and High-Quality Localization
por: Gu, Zhaopeng, et al.
Publicado: (2024) -
AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly Detection
por: Gu, Zhaopeng, et al.
Publicado: (2025) -
UniVAD: A Training-free Unified Model for Few-shot Visual Anomaly Detection
por: Gu, Zhaopeng, et al.
Publicado: (2024)