Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
Fuente:
arXiv
Guardado en:
| Autores principales: | Tan, Chuangchuang, Ming, Xiang, Wang, Jinglu, Tao, Renshuai, Li, Bin, Wei, Yunchao, Zhao, Yao, Lu, Yan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
por: Tan, Chuangchuang, et al.
Publicado: (2025)
por: Tan, Chuangchuang, et al.
Publicado: (2025)
C2P-CLIP: Injecting Category Common Prompt in CLIP to Enhance Generalization in Deepfake Detection
por: Tan, Chuangchuang, et al.
Publicado: (2024)
por: Tan, Chuangchuang, et al.
Publicado: (2024)
Can a Second-View Image Be a Language? Geometric and Semantic Cross-Modal Reasoning for X-ray Prohibited Item Detection
por: Peng, Chuang, et al.
Publicado: (2025)
por: Peng, Chuang, et al.
Publicado: (2025)
ODDN: Addressing Unpaired Data Challenges in Open-World Deepfake Detection on Online Social Networks
por: Tao, Renshuai, et al.
Publicado: (2024)
por: Tao, Renshuai, et al.
Publicado: (2024)
Pay Less Attention to Deceptive Artifacts: Robust Detection of Compressed Deepfakes on Online Social Networks
por: Li, Manyi, et al.
Publicado: (2025)
por: Li, Manyi, et al.
Publicado: (2025)
PAD-F: Prior-Aware Debiasing Framework for Long-Tailed X-ray Prohibited Item Detection
por: Wang, Haoyu, et al.
Publicado: (2024)
por: Wang, Haoyu, et al.
Publicado: (2024)
Dual-view X-ray Detection: Can AI Detect Prohibited Items from Dual-view X-ray Images like Humans?
por: Tao, Renshuai, et al.
Publicado: (2024)
por: Tao, Renshuai, et al.
Publicado: (2024)
OmniAD: Detect and Understand Industrial Anomaly via Multimodal Reasoning
por: Zhao, Shifang, et al.
Publicado: (2025)
por: Zhao, Shifang, et al.
Publicado: (2025)
DCI: Dual-Conditional Inversion for Boosting Diffusion-Based Image Editing
por: Li, Zixiang, et al.
Publicado: (2025)
por: Li, Zixiang, et al.
Publicado: (2025)
Frequency-Aware Deepfake Detection: Improving Generalizability through Frequency Space Learning
por: Tan, Chuangchuang, et al.
Publicado: (2024)
por: Tan, Chuangchuang, et al.
Publicado: (2024)
BGM: Background Mixup for X-ray Prohibited Items Detection
por: Liu, Weizhe, et al.
Publicado: (2024)
por: Liu, Weizhe, et al.
Publicado: (2024)
Data-Independent Operator: A Training-Free Artifact Representation Extractor for Generalizable Deepfake Detection
por: Tan, Chuangchuang, et al.
Publicado: (2024)
por: Tan, Chuangchuang, et al.
Publicado: (2024)
Taming Generative Synthetic Data for X-ray Prohibited Item Detection
por: Sun, Jialong, et al.
Publicado: (2025)
por: Sun, Jialong, et al.
Publicado: (2025)
CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer
por: Nie, Wenbo, et al.
Publicado: (2026)
por: Nie, Wenbo, et al.
Publicado: (2026)
IPSeg: Image Posterior Mitigates Semantic Drift in Class-Incremental Segmentation
por: Yu, Xiao, et al.
Publicado: (2025)
por: Yu, Xiao, et al.
Publicado: (2025)
Scaling Up AI-Generated Image Detection with Generator-Aware Prototypes
por: Qin, Ziheng, et al.
Publicado: (2025)
por: Qin, Ziheng, et al.
Publicado: (2025)
Leveraging Failed Samples: A Few-Shot and Training-Free Framework for Generalized Deepfake Detection
por: Yao, Shibo, et al.
Publicado: (2025)
por: Yao, Shibo, et al.
Publicado: (2025)
SIE3D: Single-Image Expressive 3D Avatar Generation via Semantic Embedding and Perceptual Expression Loss
por: Huang, Zhiqi, et al.
Publicado: (2025)
por: Huang, Zhiqi, et al.
Publicado: (2025)
GS-Marker: Generalizable and Robust Watermarking for 3D Gaussian Splatting
por: Li, Lijiang, et al.
Publicado: (2025)
por: Li, Lijiang, et al.
Publicado: (2025)
PreFM: Online Audio-Visual Event Parsing via Predictive Future Modeling
por: Yu, Xiao, et al.
Publicado: (2025)
por: Yu, Xiao, et al.
Publicado: (2025)
Unsupervised Region-Based Image Editing of Denoising Diffusion Models
por: Li, Zixiang, et al.
Publicado: (2024)
por: Li, Zixiang, et al.
Publicado: (2024)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
por: Li, Xiang, et al.
Publicado: (2023)
por: Li, Xiang, et al.
Publicado: (2023)
Frozen CLIP: A Strong Backbone for Weakly Supervised Semantic Segmentation
por: Zhang, Bingfeng, et al.
Publicado: (2024)
por: Zhang, Bingfeng, et al.
Publicado: (2024)
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
por: Zia, Ali, et al.
Publicado: (2026)
por: Zia, Ali, et al.
Publicado: (2026)
Collaborative Feature-Logits Contrastive Learning for Open-Set Semi-Supervised Object Detection
por: Zhong, Xinhao, et al.
Publicado: (2024)
por: Zhong, Xinhao, et al.
Publicado: (2024)
ACTRESS: Active Retraining for Semi-supervised Visual Grounding
por: Kang, Weitai, et al.
Publicado: (2024)
por: Kang, Weitai, et al.
Publicado: (2024)
CLIP-Flow: A Universal Discriminator for AI-Generated Images Inspired by Anomaly Detection
por: Yuan, Zhipeng, et al.
Publicado: (2025)
por: Yuan, Zhipeng, et al.
Publicado: (2025)
VarAD: Lightweight High-Resolution Image Anomaly Detection via Visual Autoregressive Modeling
por: Cao, Yunkang, et al.
Publicado: (2024)
por: Cao, Yunkang, et al.
Publicado: (2024)
A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis
por: Lin, Dongheng, et al.
Publicado: (2025)
por: Lin, Dongheng, et al.
Publicado: (2025)
DCEdit: Dual-Level Controlled Image Editing via Precisely Localized Semantics
por: Hu, Yihan, et al.
Publicado: (2025)
por: Hu, Yihan, et al.
Publicado: (2025)
DreamLCM: Towards High-Quality Text-to-3D Generation via Latent Consistency Model
por: Zhong, Yiming, et al.
Publicado: (2024)
por: Zhong, Yiming, et al.
Publicado: (2024)
AlignGen: Boosting Personalized Image Generation with Cross-Modality Prior Alignment
por: Lin, Yiheng, et al.
Publicado: (2025)
por: Lin, Yiheng, et al.
Publicado: (2025)
Leveraging Unlabeled Data from Unknown Sources via Dual-Path Guidance for Deepfake Face Detection
por: Yang, Zhiqiang, et al.
Publicado: (2025)
por: Yang, Zhiqiang, et al.
Publicado: (2025)
Diffusion for Natural Image Matting
por: Hu, Yihan, et al.
Publicado: (2023)
por: Hu, Yihan, et al.
Publicado: (2023)
ThinkGen: Generalized Thinking for Visual Generation
por: Jiao, Siyu, et al.
Publicado: (2025)
por: Jiao, Siyu, et al.
Publicado: (2025)
ThinkFake: Reasoning in Multimodal Large Language Models for AI-Generated Image Detection
por: Huang, Tai-Ming, et al.
Publicado: (2025)
por: Huang, Tai-Ming, et al.
Publicado: (2025)
ForgeryVCR: Visual-Centric Reasoning via Efficient Forensic Tools in MLLMs for Image Forgery Detection and Localization
por: Wang, Youqi, et al.
Publicado: (2026)
por: Wang, Youqi, et al.
Publicado: (2026)
Let ViT Speak: Generative Language-Image Pre-training
por: Fang, Yan, et al.
Publicado: (2026)
por: Fang, Yan, et al.
Publicado: (2026)
Towards Visual Discrimination and Reasoning of Real-World Physical Dynamics: Physics-Grounded Anomaly Detection
por: Li, Wenqiao, et al.
Publicado: (2025)
por: Li, Wenqiao, et al.
Publicado: (2025)
PSVMA+: Exploring Multi-granularity Semantic-visual Adaption for Generalized Zero-shot Learning
por: Liu, Man, et al.
Publicado: (2024)
por: Liu, Man, et al.
Publicado: (2024)
Ejemplares similares
-
ForenX: Towards Explainable AI-Generated Image Detection with Multimodal Large Language Models
por: Tan, Chuangchuang, et al.
Publicado: (2025) -
C2P-CLIP: Injecting Category Common Prompt in CLIP to Enhance Generalization in Deepfake Detection
por: Tan, Chuangchuang, et al.
Publicado: (2024) -
Can a Second-View Image Be a Language? Geometric and Semantic Cross-Modal Reasoning for X-ray Prohibited Item Detection
por: Peng, Chuang, et al.
Publicado: (2025) -
ODDN: Addressing Unpaired Data Challenges in Open-World Deepfake Detection on Online Social Networks
por: Tao, Renshuai, et al.
Publicado: (2024) -
Pay Less Attention to Deceptive Artifacts: Robust Detection of Compressed Deepfakes on Online Social Networks
por: Li, Manyi, et al.
Publicado: (2025)