Leveraging Hierarchical Image-Text Misalignment for Universal Fake Image Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Daichi, Zhang, Tong, Bao, Jianmin, Ge, Shiming, Süsstrunk, Sabine |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Frequency Forgery Clues for Diffusion-Generated Image Detection
by: Zhang, Daichi, et al.
Published: (2025)
by: Zhang, Daichi, et al.
Published: (2025)
Divide and Conquer: Static-Dynamic Collaboration for Few-Shot Class-Incremental Learning
by: Bao, Kexin, et al.
Published: (2026)
by: Bao, Kexin, et al.
Published: (2026)
CD^2: Constrained Dataset Distillation for Few-Shot Class-Incremental Learning
by: Bao, Kexin, et al.
Published: (2026)
by: Bao, Kexin, et al.
Published: (2026)
Interpret the Predictions of Deep Networks via Re-Label Distillation
by: Hua, Yingying, et al.
Published: (2024)
by: Hua, Yingying, et al.
Published: (2024)
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
by: Li, Meiling, et al.
Published: (2024)
by: Li, Meiling, et al.
Published: (2024)
Security Risk of Misalignment between Text and Image in Multi-modal Model
by: Wang, Xiaosen, et al.
Published: (2025)
by: Wang, Xiaosen, et al.
Published: (2025)
R$^2$BD: A Reconstruction-Based Method for Generalizable and Efficient Detection of Fake Images
by: Liu, Qingyu, et al.
Published: (2026)
by: Liu, Qingyu, et al.
Published: (2026)
Latent Spatiotemporal Adaptation for Generalized Face Forgery Video Detection
by: Zhang, Daichi, et al.
Published: (2023)
by: Zhang, Daichi, et al.
Published: (2023)
Look Through Masks: Towards Masked Face Recognition with De-Occlusion Distillation
by: Li, Chenyu, et al.
Published: (2024)
by: Li, Chenyu, et al.
Published: (2024)
Zooming into Comics: Region-Aware RL Improves Fine-Grained Comic Understanding in Vision-Language Models
by: Chen, Yule, et al.
Published: (2025)
by: Chen, Yule, et al.
Published: (2025)
Mitigating Object Dependencies: Improving Point Cloud Self-Supervised Learning through Object Exchange
by: Wu, Yanhao, et al.
Published: (2024)
by: Wu, Yanhao, et al.
Published: (2024)
DSI2I: Dense Style for Unpaired Image-to-Image Translation
by: Ozaydin, Baran, et al.
Published: (2022)
by: Ozaydin, Baran, et al.
Published: (2022)
HCVP: Leveraging Hierarchical Contrastive Visual Prompt for Domain Generalization
by: Zhou, Guanglin, et al.
Published: (2024)
by: Zhou, Guanglin, et al.
Published: (2024)
Mesh Neural Cellular Automata
by: Pajouheshgar, Ehsan, et al.
Published: (2023)
by: Pajouheshgar, Ehsan, et al.
Published: (2023)
FDS: Frequency-Aware Denoising Score for Text-Guided Latent Diffusion Image Editing
by: Ren, Yufan, et al.
Published: (2025)
by: Ren, Yufan, et al.
Published: (2025)
FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models
by: Xu, Zhipei, et al.
Published: (2024)
by: Xu, Zhipei, et al.
Published: (2024)
Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection
by: Jiang, Changjiang, et al.
Published: (2026)
by: Jiang, Changjiang, et al.
Published: (2026)
FakeInversion: Learning to Detect Images from Unseen Text-to-Image Models by Inverting Stable Diffusion
by: Cazenavette, George, et al.
Published: (2024)
by: Cazenavette, George, et al.
Published: (2024)
Ivy-Fake: A Unified Explainable Framework and Benchmark for Image and Video AIGC Detection
by: Jiang, Changjiang, et al.
Published: (2025)
by: Jiang, Changjiang, et al.
Published: (2025)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
by: Zhou, Linhan, et al.
Published: (2025)
by: Zhou, Linhan, et al.
Published: (2025)
Few-shot Class-Incremental Learning via Generative Co-Memory Regularization
by: Bao, Kexin, et al.
Published: (2026)
by: Bao, Kexin, et al.
Published: (2026)
FoCLIP: A Feature-Space Misalignment Framework for CLIP-Based Image Manipulation and Detection
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Anomaly Detection by Effectively Leveraging Synthetic Images
by: Kang, Sungho, et al.
Published: (2025)
by: Kang, Sungho, et al.
Published: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
by: Zhao, Yu, et al.
Published: (2024)
by: Zhao, Yu, et al.
Published: (2024)
Redefining Generalization in Visual Domains: A Two-Axis Framework for Fake Image Detection with FusionDetect
by: Amanzadi, Amirtaha, et al.
Published: (2025)
by: Amanzadi, Amirtaha, et al.
Published: (2025)
15M Multimodal Facial Image-Text Dataset
by: Dai, Dawei, et al.
Published: (2024)
by: Dai, Dawei, et al.
Published: (2024)
NoiseNCA: Noisy Seed Improves Spatio-Temporal Continuity of Neural Cellular Automata
by: Pajouheshgar, Ehsan, et al.
Published: (2024)
by: Pajouheshgar, Ehsan, et al.
Published: (2024)
Learning Natural Consistency Representation for Face Forgery Video Detection
by: Zhang, Daichi, et al.
Published: (2024)
by: Zhang, Daichi, et al.
Published: (2024)
Look One and More: Distilling Hybrid Order Relational Knowledge for Cross-Resolution Image Recognition
by: Ge, Shiming, et al.
Published: (2024)
by: Ge, Shiming, et al.
Published: (2024)
Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement
by: Lee, Daeun, et al.
Published: (2024)
by: Lee, Daeun, et al.
Published: (2024)
Comparative Evaluation of Deep Learning Models for Fake Image Detection
by: Pakala, Akhitha, et al.
Published: (2026)
by: Pakala, Akhitha, et al.
Published: (2026)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
by: Peng, Xingkai, et al.
Published: (2025)
by: Peng, Xingkai, et al.
Published: (2025)
Towards Universal Text-driven CT Image Segmentation
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
VIRTUE: Visual-Interactive Text-Image Universal Embedder
by: Wang, Wei-Yao, et al.
Published: (2025)
by: Wang, Wei-Yao, et al.
Published: (2025)
Towards More Accurate Fake Detection on Images Generated from Advanced Generative and Neural Rendering Models
by: Dong, Chengdong, et al.
Published: (2024)
by: Dong, Chengdong, et al.
Published: (2024)
Unveiling the Truth: Exploring Human Gaze Patterns in Fake Images
by: Cartella, Giuseppe, et al.
Published: (2024)
by: Cartella, Giuseppe, et al.
Published: (2024)
Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models
by: Zhang, Huixuan, et al.
Published: (2025)
by: Zhang, Huixuan, et al.
Published: (2025)
SINE: SINgle Image Editing with Text-to-Image Diffusion Models
by: Zhang, Zhixing, et al.
Published: (2022)
by: Zhang, Zhixing, et al.
Published: (2022)
StableAnimator++: Overcoming Pose Misalignment and Face Distortion for Human Image Animation
by: Tu, Shuyuan, et al.
Published: (2025)
by: Tu, Shuyuan, et al.
Published: (2025)
Condition Weaving Meets Expert Modulation: Towards Universal and Controllable Image Generation
by: Zhang, Guoqing, et al.
Published: (2025)
by: Zhang, Guoqing, et al.
Published: (2025)
Similar Items
-
Enhancing Frequency Forgery Clues for Diffusion-Generated Image Detection
by: Zhang, Daichi, et al.
Published: (2025) -
Divide and Conquer: Static-Dynamic Collaboration for Few-Shot Class-Incremental Learning
by: Bao, Kexin, et al.
Published: (2026) -
CD^2: Constrained Dataset Distillation for Few-Shot Class-Incremental Learning
by: Bao, Kexin, et al.
Published: (2026) -
Interpret the Predictions of Deep Networks via Re-Label Distillation
by: Hua, Yingying, et al.
Published: (2024) -
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
by: Li, Meiling, et al.
Published: (2024)