Toward Generalizable Forgery Detection and Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Yueying, Chang, Dongliang, Yu, Bingyao, Qin, Haotian, Diao, Muxi, Chen, Lei, Liang, Kongming, Ma, Zhanyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Conditional Information Bottleneck for Generalizable AI-Generated Image Detection
von: Qin, Haotian, et al.
Veröffentlicht: (2025)
von: Qin, Haotian, et al.
Veröffentlicht: (2025)
Towards Privacy-Preserving Fine-Grained Visual Classification via Hierarchical Learning from Label Proportions
von: Chang, Jinyi, et al.
Veröffentlicht: (2025)
von: Chang, Jinyi, et al.
Veröffentlicht: (2025)
IncreFA: Breaking the Static Wall of Generative Model Attribution
von: Qin, Haotian, et al.
Veröffentlicht: (2026)
von: Qin, Haotian, et al.
Veröffentlicht: (2026)
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
von: Diao, Muxi, et al.
Veröffentlicht: (2025)
von: Diao, Muxi, et al.
Veröffentlicht: (2025)
From Simple to Professional: A Combinatorial Controllable Image Captioning Agent
von: Wang, Xinran, et al.
Veröffentlicht: (2024)
von: Wang, Xinran, et al.
Veröffentlicht: (2024)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
MedReasoner: Reinforcement Learning Drives Reasoning Grounding from Clinical Thought to Pixel-Level Precision
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
Hepato-LLaVA: An Expert MLLM with Sparse Topo-Pack Attention for Hepatocellular Pathology Analysis on Whole Slide Images
von: Yang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Yang, Yuxuan, et al.
Veröffentlicht: (2026)
CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
RO-Bench: Large-scale robustness evaluation of MLLMs with text-driven counterfactual videos
von: Yang, Zixi, et al.
Veröffentlicht: (2025)
von: Yang, Zixi, et al.
Veröffentlicht: (2025)
ForgeLens: Data-Efficient Forgery Focus for Generalizable Forgery Image Detection
von: Chen, Yingjian, et al.
Veröffentlicht: (2024)
von: Chen, Yingjian, et al.
Veröffentlicht: (2024)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation
von: Huang, Qing, et al.
Veröffentlicht: (2026)
von: Huang, Qing, et al.
Veröffentlicht: (2026)
Benchmarking Segmentation Models with Mask-Preserved Attribute Editing
von: Yin, Zijin, et al.
Veröffentlicht: (2024)
von: Yin, Zijin, et al.
Veröffentlicht: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
Poisoned Forgery Face: Towards Backdoor Attacks on Face Forgery Detection
von: Liang, Jiawei, et al.
Veröffentlicht: (2024)
von: Liang, Jiawei, et al.
Veröffentlicht: (2024)
Decoupling Forgery Semantics for Generalizable Deepfake Detection
von: Ye, Wei, et al.
Veröffentlicht: (2024)
von: Ye, Wei, et al.
Veröffentlicht: (2024)
Recolour What Matters: Region-Aware Colour Editing via Token-Level Diffusion
von: Yang, Yuqi, et al.
Veröffentlicht: (2026)
von: Yang, Yuqi, et al.
Veröffentlicht: (2026)
Generalizable Face Forgery Detection via Separable Prompt Learning
von: Yang, Enrui, et al.
Veröffentlicht: (2026)
von: Yang, Enrui, et al.
Veröffentlicht: (2026)
Learning Expressive And Generalizable Motion Features For Face Forgery Detection
von: Zhang, Jingyi, et al.
Veröffentlicht: (2024)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2024)
Exploring Specular Reflection Inconsistency for Generalizable Face Forgery Detection
von: Fei, Hongyan, et al.
Veröffentlicht: (2026)
von: Fei, Hongyan, et al.
Veröffentlicht: (2026)
MFVLR: Multi-domain Fine-grained Vision-Language Reconstruction for Generalizable Diffusion Face Forgery Detection and Localization
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
von: Gao, Jiayi, et al.
Veröffentlicht: (2025)
Loupe: A Generalizable and Adaptive Framework for Image Forgery Detection
von: Jiang, Yuchu, et al.
Veröffentlicht: (2025)
von: Jiang, Yuchu, et al.
Veröffentlicht: (2025)
Efficient Face Super-Resolution via Wavelet-based Feature Enhancement Network
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
Evaluating Attribute Comprehension in Large Vision-Language Models
von: Zhang, Haiwen, et al.
Veröffentlicht: (2024)
von: Zhang, Haiwen, et al.
Veröffentlicht: (2024)
Benchmarking Semantic Segmentation Models via Appearance and Geometry Attribute Editing
von: Yin, Zijin, et al.
Veröffentlicht: (2026)
von: Yin, Zijin, et al.
Veröffentlicht: (2026)
Seeing as Experts Do: A Knowledge-Augmented Agent for Open-Set Fine-Grained Visual Understanding
von: Chen, Junhan, et al.
Veröffentlicht: (2026)
von: Chen, Junhan, et al.
Veröffentlicht: (2026)
FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2025)
PGP-SAM: Prototype-Guided Prompt Learning for Efficient Few-Shot Medical Image Segmentation
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
von: Yan, Zhonghao, et al.
Veröffentlicht: (2025)
Learning Universal Features for Generalizable Image Forgery Localization
von: Zhao, Hengrun, et al.
Veröffentlicht: (2025)
von: Zhao, Hengrun, et al.
Veröffentlicht: (2025)
Transcending Forgery Specificity with Latent Space Augmentation for Generalizable Deepfake Detection
von: Yan, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Yan, Zhiyuan, et al.
Veröffentlicht: (2023)
Low-rank Orthogonal Subspace Intervention for Generalizable Face Forgery Detection
von: Wang, Chi, et al.
Veröffentlicht: (2026)
von: Wang, Chi, et al.
Veröffentlicht: (2026)
Learning to Discover Forgery Cues for Face Forgery Detection
von: Tian, Jiahe, et al.
Veröffentlicht: (2024)
von: Tian, Jiahe, et al.
Veröffentlicht: (2024)
Suppressing Forgery-Specific Shortcuts for Generalizable Deepfake Detection
von: Wang, Yihui, et al.
Veröffentlicht: (2026)
von: Wang, Yihui, et al.
Veröffentlicht: (2026)
OmniFD: A Unified Model for Versatile Face Forgery Detection
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models
von: Tong, Yujun, et al.
Veröffentlicht: (2026)
von: Tong, Yujun, et al.
Veröffentlicht: (2026)
ForgeryVCR: Visual-Centric Reasoning via Efficient Forensic Tools in MLLMs for Image Forgery Detection and Localization
von: Wang, Youqi, et al.
Veröffentlicht: (2026)
von: Wang, Youqi, et al.
Veröffentlicht: (2026)
Inclusion 2024 Global Multimedia Deepfake Detection Challenge: Towards Multi-dimensional Face Forgery Detection
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
Towards General Visual-Linguistic Face Forgery Detection
von: Sun, Ke, et al.
Veröffentlicht: (2023)
von: Sun, Ke, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Multimodal Conditional Information Bottleneck for Generalizable AI-Generated Image Detection
von: Qin, Haotian, et al.
Veröffentlicht: (2025) -
Towards Privacy-Preserving Fine-Grained Visual Classification via Hierarchical Learning from Label Proportions
von: Chang, Jinyi, et al.
Veröffentlicht: (2025) -
IncreFA: Breaking the Static Wall of Generative Model Attribution
von: Qin, Haotian, et al.
Veröffentlicht: (2026) -
DriveRX: A Vision-Language Reasoning Model for Cross-Task Autonomous Driving
von: Diao, Muxi, et al.
Veröffentlicht: (2025) -
From Simple to Professional: A Combinatorial Controllable Image Captioning Agent
von: Wang, Xinran, et al.
Veröffentlicht: (2024)