Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Yikun, Yan, Hong, Lan, Jun, Zhu, Huijia, Wang, Weiqiang, Fan, Qi, Zhang, Liqing, Zhang, Jianfu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
Towards Explainable Fake Image Detection with Multi-Modal Large Language Models
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
von: Ji, Yikun, et al.
Veröffentlicht: (2025)
Robustness in AI-Generated Detection: Enhancing Resistance to Adversarial Attacks
von: Haoxuan, Sun, et al.
Veröffentlicht: (2025)
von: Haoxuan, Sun, et al.
Veröffentlicht: (2025)
GAMMA: Generalizable Alignment via Multi-task and Manipulation-Augmented Training for AI-Generated Image Detection
von: Yan, Haozhen, et al.
Veröffentlicht: (2025)
von: Yan, Haozhen, et al.
Veröffentlicht: (2025)
DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric Finetuning
von: Duan, Yuxuan, et al.
Veröffentlicht: (2024)
von: Duan, Yuxuan, et al.
Veröffentlicht: (2024)
Supervised Contrastive Learning for Snapshot Spectral Imaging Face Anti-Spoofing
von: Song, Chuanbiao, et al.
Veröffentlicht: (2024)
von: Song, Chuanbiao, et al.
Veröffentlicht: (2024)
COCO-Inpaint: A Benchmark for Detecting and Localizing Inpainting-Based Image Manipulations
von: Yan, Haozhen, et al.
Veröffentlicht: (2025)
von: Yan, Haozhen, et al.
Veröffentlicht: (2025)
DS-VTON: An Enhanced Dual-Scale Coarse-to-Fine Framework for Virtual Try-On
von: Sun, Xianbing, et al.
Veröffentlicht: (2025)
von: Sun, Xianbing, et al.
Veröffentlicht: (2025)
EMIT: Enhancing MLLMs for Industrial Anomaly Detection via Difficulty-Aware GRPO
von: Guan, Wei, et al.
Veröffentlicht: (2025)
von: Guan, Wei, et al.
Veröffentlicht: (2025)
DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
WildFake: A Large-scale Challenging Dataset for AI-Generated Images Detection
von: Hong, Yan, et al.
Veröffentlicht: (2024)
von: Hong, Yan, et al.
Veröffentlicht: (2024)
Generalizable and Adaptive Continual Learning Framework for AI-generated Image Detection
von: Wang, Hanyi, et al.
Veröffentlicht: (2026)
von: Wang, Hanyi, et al.
Veröffentlicht: (2026)
Assessing Image Inpainting via Re-Inpainting Self-Consistency Evaluation
von: Chen, Tianyi, et al.
Veröffentlicht: (2024)
von: Chen, Tianyi, et al.
Veröffentlicht: (2024)
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
von: Tan, Hao, et al.
Veröffentlicht: (2026)
von: Tan, Hao, et al.
Veröffentlicht: (2026)
Towards Source-Aware Object Swapping with Initial Noise Perturbation
von: Zhan, Jiahui, et al.
Veröffentlicht: (2026)
von: Zhan, Jiahui, et al.
Veröffentlicht: (2026)
ComFusion: Personalized Subject Generation in Multiple Specific Scenes From Single Image
von: Hong, Yan, et al.
Veröffentlicht: (2024)
von: Hong, Yan, et al.
Veröffentlicht: (2024)
Enhancing Domain Generalization in 3D Human Pose Estimation through Controllable Generative Augmentation
von: Hu, Xinhao, et al.
Veröffentlicht: (2026)
von: Hu, Xinhao, et al.
Veröffentlicht: (2026)
Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning
von: Zhang, Bob, et al.
Veröffentlicht: (2025)
von: Zhang, Bob, et al.
Veröffentlicht: (2025)
Veritas: Generalizable Deepfake Detection via Pattern-Aware Reasoning
von: Tan, Hao, et al.
Veröffentlicht: (2025)
von: Tan, Hao, et al.
Veröffentlicht: (2025)
VTONGuard: Automatic Detection and Authentication of AI-Generated Virtual Try-On Content
von: Wu, Shengyi, et al.
Veröffentlicht: (2026)
von: Wu, Shengyi, et al.
Veröffentlicht: (2026)
DirectTryOn: One-Step Virtual Try-On via Straightened Conditional Transport
von: Sun, Xianbing, et al.
Veröffentlicht: (2026)
von: Sun, Xianbing, et al.
Veröffentlicht: (2026)
User-Friendly Customized Generation with Multi-Modal Prompts
von: Zhong, Linhao, et al.
Veröffentlicht: (2024)
von: Zhong, Linhao, et al.
Veröffentlicht: (2024)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
WeditGAN: Few-Shot Image Generation via Latent Space Relocation
von: Duan, Yuxuan, et al.
Veröffentlicht: (2023)
von: Duan, Yuxuan, et al.
Veröffentlicht: (2023)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
von: Guo, Pinxue, et al.
Veröffentlicht: (2025)
Any3DAvatar: Fast and High-Quality Full-Head 3D Avatar Reconstruction from Single Portrait Image
von: Gao, Yujie, et al.
Veröffentlicht: (2026)
von: Gao, Yujie, et al.
Veröffentlicht: (2026)
Beyond Unimodal Shortcuts: MLLMs as Cross-Modal Reasoners for Grounded Named Entity Recognition
von: Ma, Jinlong, et al.
Veröffentlicht: (2026)
von: Ma, Jinlong, et al.
Veröffentlicht: (2026)
Boosting Audio-visual Zero-shot Learning with Large Language Models
von: Chen, Haoxing, et al.
Veröffentlicht: (2023)
von: Chen, Haoxing, et al.
Veröffentlicht: (2023)
Conditional Prototype Rectification Prompt Learning
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
von: Chen, Haoxing, et al.
Veröffentlicht: (2024)
Adaptive and Balanced Re-initialization for Long-timescale Continual Test-time Domain Adaptation
von: Wang, Yanshuo, et al.
Veröffentlicht: (2026)
von: Wang, Yanshuo, et al.
Veröffentlicht: (2026)
Maintain Plasticity in Long-timescale Continual Test-time Adaptation
von: Wang, Yanshuo, et al.
Veröffentlicht: (2024)
von: Wang, Yanshuo, et al.
Veröffentlicht: (2024)
Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos
von: Tang, Yuqi, et al.
Veröffentlicht: (2026)
von: Tang, Yuqi, et al.
Veröffentlicht: (2026)
High-Quality 3D Head Reconstruction from Any Single Portrait Image
von: Zhang, Jianfu, et al.
Veröffentlicht: (2025)
von: Zhang, Jianfu, et al.
Veröffentlicht: (2025)
PlantMarkerBench: A Multi-Species Benchmark for Evidence-Grounded Plant Marker Reasoning
von: Dip, Sajib Acharjee, et al.
Veröffentlicht: (2026)
von: Dip, Sajib Acharjee, et al.
Veröffentlicht: (2026)
Stochastic Layer-Wise Shuffle for Improving Vision Mamba Training
von: Huang, Zizheng, et al.
Veröffentlicht: (2024)
von: Huang, Zizheng, et al.
Veröffentlicht: (2024)
Skyra: AI-Generated Video Detection via Grounded Artifact Reasoning
von: Li, Yifei, et al.
Veröffentlicht: (2025)
von: Li, Yifei, et al.
Veröffentlicht: (2025)
Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs
von: Huang, Jincai, et al.
Veröffentlicht: (2026)
von: Huang, Jincai, et al.
Veröffentlicht: (2026)
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
von: Wang, Shanshan, et al.
Veröffentlicht: (2026)
von: Wang, Shanshan, et al.
Veröffentlicht: (2026)
CameraBench: Benchmarking Visual Reasoning in MLLMs via Photography
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception
von: Wei, Lai, et al.
Veröffentlicht: (2026)
von: Wei, Lai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images
von: Ji, Yikun, et al.
Veröffentlicht: (2025) -
Towards Explainable Fake Image Detection with Multi-Modal Large Language Models
von: Ji, Yikun, et al.
Veröffentlicht: (2025) -
Robustness in AI-Generated Detection: Enhancing Resistance to Adversarial Attacks
von: Haoxuan, Sun, et al.
Veröffentlicht: (2025) -
GAMMA: Generalizable Alignment via Multi-task and Manipulation-Augmented Training for AI-Generated Image Detection
von: Yan, Haozhen, et al.
Veröffentlicht: (2025) -
DomainGallery: Few-shot Domain-driven Image Generation by Attribute-centric Finetuning
von: Duan, Yuxuan, et al.
Veröffentlicht: (2024)