When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Jiaqi, Zhou, Yuchen, Ng, Dennis Tsang, Shen, Xingyu, Zewde, Kidus, Raj, Ankit, Duong, Tommy, Ren, Simiao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Synthetic Eye Movement Dataset for Script Reading Detection: Real Trajectory Replay on a 3D Simulator
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
di: Zhang, Yan, et al.
Pubblicazione: (2026)
di: Zhang, Yan, et al.
Pubblicazione: (2026)
Can Multi-modal (reasoning) LLMs work as deepfake detectors?
di: Ren, Simiao, et al.
Pubblicazione: (2025)
di: Ren, Simiao, et al.
Pubblicazione: (2025)
Do Deepfake Detectors Work in Reality?
di: Ren, Simiao, et al.
Pubblicazione: (2025)
di: Ren, Simiao, et al.
Pubblicazione: (2025)
Can a Teenager Fool an AI? Evaluating Low-Cost Cosmetic Attacks on Age Estimation Systems
di: Shen, Xingyu, et al.
Pubblicazione: (2026)
di: Shen, Xingyu, et al.
Pubblicazione: (2026)
Conditions of Cognition: Why Meaning Cannot Explain Its Own Origin
di: kondakly, Yousif
Pubblicazione: (2026)
di: kondakly, Yousif
Pubblicazione: (2026)
Forger la citoyenneté juvénile
di: Csupor, Isabelle, et al.
Pubblicazione: (2024)
di: Csupor, Isabelle, et al.
Pubblicazione: (2024)
Can Multi-modal (reasoning) LLMs detect document manipulation?
di: Liang, Zisheng, et al.
Pubblicazione: (2025)
di: Liang, Zisheng, et al.
Pubblicazione: (2025)
To Each Generation Its Own Rabbits
di: Flanagan, Dennis
Pubblicazione: (1974)
di: Flanagan, Dennis
Pubblicazione: (1974)
You Can Use But Cannot Recognize: Preserving Visual Privacy in Deep Neural Networks
di: Li, Qiushi, et al.
Pubblicazione: (2024)
di: Li, Qiushi, et al.
Pubblicazione: (2024)
LLM Evaluators Recognize and Favor Their Own Generations
di: Panickssery, Arjun, et al.
Pubblicazione: (2024)
di: Panickssery, Arjun, et al.
Pubblicazione: (2024)
Can AI Recognize Its Own Reflection? Self-Detection Performance of LLMs in Computing Education
di: Burger, Christopher, et al.
Pubblicazione: (2025)
di: Burger, Christopher, et al.
Pubblicazione: (2025)
Out of the box age estimation through facial imagery: A Comprehensive Benchmark of Vision-Language Models vs. out-of-the-box Traditional Architectures
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
The Lazy Student's Dream: ChatGPT Passing an Engineering Course on Its Own
di: Puthumanaillam, Gokul, et al.
Pubblicazione: (2025)
di: Puthumanaillam, Gokul, et al.
Pubblicazione: (2025)
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation
di: Corbara, Silvia, et al.
Pubblicazione: (2024)
di: Corbara, Silvia, et al.
Pubblicazione: (2024)
AIForge-Doc: A Benchmark for Detecting AI-Forged Tampering in Financial and Form Documents
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
Learning Human–AI Relationships Through Astro Boy — Why the Capability Race Cannot Stop on Its Own v1.2
di: Seo, Y
Pubblicazione: (2026)
di: Seo, Y
Pubblicazione: (2026)
The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement
di: Deng, Lin, et al.
Pubblicazione: (2026)
di: Deng, Lin, et al.
Pubblicazione: (2026)
Motion-Coupled Sensing: When the State Change Powers Its Own Sensing
di: Tahir, Muhammad, et al.
Pubblicazione: (2026)
di: Tahir, Muhammad, et al.
Pubblicazione: (2026)
Each Fake News is Fake in its Own Way: An Attribution Multi-Granularity Benchmark for Multimodal Fake News Detection
di: Guo, Hao, et al.
Pubblicazione: (2024)
di: Guo, Hao, et al.
Pubblicazione: (2024)
LLMs Cannot Reliably Judge (Yet?): A Comprehensive Assessment on the Robustness of LLM-as-a-Judge
di: Li, Songze, et al.
Pubblicazione: (2025)
di: Li, Songze, et al.
Pubblicazione: (2025)
DiCache: Let Diffusion Model Determine Its Own Cache
di: Bu, Jiazi, et al.
Pubblicazione: (2025)
di: Bu, Jiazi, et al.
Pubblicazione: (2025)
Vascular Dementia and Risk Factors in Ethiopia
di: Yared Z Zewde
Pubblicazione: (2025)
di: Yared Z Zewde
Pubblicazione: (2025)
Designing Culturally Appropriate Cognitive Tools in Resource Limited Setting: Lessons from Ethiopia's Multilingual and Culturally Diverse Context
di: Yared Z Zewde
Pubblicazione: (2025)
di: Yared Z Zewde
Pubblicazione: (2025)
LLM-as-a-Judge & Reward Model: What They Can and Cannot Do
di: Son, Guijin, et al.
Pubblicazione: (2024)
di: Son, Guijin, et al.
Pubblicazione: (2024)
SpellForger: Prompting Custom Spell Properties In-Game using BERT supervised-trained model
di: Silva, Emanuel C., et al.
Pubblicazione: (2025)
di: Silva, Emanuel C., et al.
Pubblicazione: (2025)
GPT-4V Cannot Generate Radiology Reports Yet
di: Jiang, Yuyang, et al.
Pubblicazione: (2024)
di: Jiang, Yuyang, et al.
Pubblicazione: (2024)
“Consciousness in Its Own Self Provides Its Own Standard”. Hegel and the Spirit as a Process of Thinking
di: Violetta L. Waibel
Pubblicazione: (2022)
di: Violetta L. Waibel
Pubblicazione: (2022)
Benchmarking is Broken -- Don't Let AI be its Own Judge
di: Cheng, Zerui, et al.
Pubblicazione: (2025)
di: Cheng, Zerui, et al.
Pubblicazione: (2025)
College with Its Own High School.
di: Moed, Martin G., et al.
Pubblicazione: (1982)
di: Moed, Martin G., et al.
Pubblicazione: (1982)
MYC: The Guardian of Its Own Chaos
di: Abdallah Gaballa, et al.
Pubblicazione: (2025)
di: Abdallah Gaballa, et al.
Pubblicazione: (2025)
VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation
di: Kumar, Divake, et al.
Pubblicazione: (2026)
di: Kumar, Divake, et al.
Pubblicazione: (2026)
Detection of ChatGPT Fake Science with the xFakeSci Learning Algorithm
di: Hamed, Ahmed Abdeen, et al.
Pubblicazione: (2023)
di: Hamed, Ahmed Abdeen, et al.
Pubblicazione: (2023)
FakeGPT: Fake News Generation, Explanation and Detection of Large Language Models
di: Huang, Yue, et al.
Pubblicazione: (2023)
di: Huang, Yue, et al.
Pubblicazione: (2023)
When AI Evaluates Its Own Work: Validating Learner-Initiated, AI-Generated Physics Practice Problems
di: Geisler, Tobias, et al.
Pubblicazione: (2025)
di: Geisler, Tobias, et al.
Pubblicazione: (2025)
Beating Backdoor Attack at Its Own Game
di: Liu, Min, et al.
Pubblicazione: (2023)
di: Liu, Min, et al.
Pubblicazione: (2023)
When Abel Kills Cain: What Machine Translation Cannot Capture
di: Bénel, Aurélien, et al.
Pubblicazione: (2024)
di: Bénel, Aurélien, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Synthetic Eye Movement Dataset for Script Reading Detection: Real Trajectory Replay on a 3D Simulator
di: Zewde, Kidus, et al.
Pubblicazione: (2026) -
GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment
di: Zewde, Kidus, et al.
Pubblicazione: (2026) -
How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study
di: Ren, Simiao, et al.
Pubblicazione: (2026) -
Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate
di: Ren, Simiao, et al.
Pubblicazione: (2026) -
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
di: Zhang, Yan, et al.
Pubblicazione: (2026)