How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study
Fuente:
arXiv
Salvato in:
| Autori principali: | Ren, Simiao, Zhou, Yuchen, Shen, Xingyu, Zewde, Kidus, Duong, Tommy, Huang, George, Hatsanai, Tiangratanakul, Tsang, Ng, Wei, En, Xue, Jiayu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Synthetic Eye Movement Dataset for Script Reading Detection: Real Trajectory Replay on a 3D Simulator
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
di: Zewde, Kidus, et al.
Pubblicazione: (2026)
Can Multi-modal (reasoning) LLMs detect document manipulation?
di: Liang, Zisheng, et al.
Pubblicazione: (2025)
di: Liang, Zisheng, et al.
Pubblicazione: (2025)
Can Multi-modal (reasoning) LLMs work as deepfake detectors?
di: Ren, Simiao, et al.
Pubblicazione: (2025)
di: Ren, Simiao, et al.
Pubblicazione: (2025)
Do Deepfake Detectors Work in Reality?
di: Ren, Simiao, et al.
Pubblicazione: (2025)
di: Ren, Simiao, et al.
Pubblicazione: (2025)
Chinese Language Is Not More Efficient Than English in Vibe Coding: A Preliminary Study on Token Cost and Problem-Solving Rate
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
Scheming in the wild: detecting real-world AI scheming incidents with open-source intelligence
di: Shane, Tommy Shaffer, et al.
Pubblicazione: (2026)
di: Shane, Tommy Shaffer, et al.
Pubblicazione: (2026)
Out of the box age estimation through facial imagery: A Comprehensive Benchmark of Vision-Language Models vs. out-of-the-box Traditional Architectures
di: Ren, Simiao, et al.
Pubblicazione: (2026)
di: Ren, Simiao, et al.
Pubblicazione: (2026)
How well do generative models solve inverse problems? A benchmark study
di: Krüger, Patrick, et al.
Pubblicazione: (2026)
di: Krüger, Patrick, et al.
Pubblicazione: (2026)
Can a Teenager Fool an AI? Evaluating Low-Cost Cosmetic Attacks on Age Estimation Systems
di: Shen, Xingyu, et al.
Pubblicazione: (2026)
di: Shen, Xingyu, et al.
Pubblicazione: (2026)
GPT4o-Receipt: A Dataset and Human Study for AI-Generated Document Forensics
di: Zhang, Yan, et al.
Pubblicazione: (2026)
di: Zhang, Yan, et al.
Pubblicazione: (2026)
Vascular Dementia and Risk Factors in Ethiopia
di: Yared Z Zewde
Pubblicazione: (2025)
di: Yared Z Zewde
Pubblicazione: (2025)
Designing Culturally Appropriate Cognitive Tools in Resource Limited Setting: Lessons from Ethiopia's Multilingual and Culturally Diverse Context
di: Yared Z Zewde
Pubblicazione: (2025)
di: Yared Z Zewde
Pubblicazione: (2025)
MedGEN-Bench: Contextually entangled benchmark for open-ended multimodal medical generation
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
Optically detected magnetic resonance with an open source platform
di: Babashah, Hossein, et al.
Pubblicazione: (2022)
di: Babashah, Hossein, et al.
Pubblicazione: (2022)
\textsc{CantoNLU}: A benchmark for Cantonese natural language understanding
di: Min, Junghyun, et al.
Pubblicazione: (2025)
di: Min, Junghyun, et al.
Pubblicazione: (2025)
Hyper-reduction methods for accelerating nonlinear finite element simulations: open source implementation and reproducible benchmarks
di: Larsson, Axel, et al.
Pubblicazione: (2026)
di: Larsson, Axel, et al.
Pubblicazione: (2026)
AIForge-Doc: A Benchmark for Detecting AI-Forged Tampering in Financial and Form Documents
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
di: Wu, Jiaqi, et al.
Pubblicazione: (2026)
Saliency strikes back: How filtering out high frequencies improves white-box explanations
di: Muzellec, Sabine, et al.
Pubblicazione: (2023)
di: Muzellec, Sabine, et al.
Pubblicazione: (2023)
Versatile, open-source program for simulating high-harmonic generation
di: Schröder, Christian A., et al.
Pubblicazione: (2025)
di: Schröder, Christian A., et al.
Pubblicazione: (2025)
Consumer views on safety of over-the-counter drugs, preferred retailers and information sources in Sweden: after re-regulation of the pharmacy market
di: Tommy Westerlund
Pubblicazione: (2017)
di: Tommy Westerlund
Pubblicazione: (2017)
TeleAI-Safety: A comprehensive LLM jailbreaking benchmark towards attacks, defenses, and evaluations
di: Chen, Xiuyuan, et al.
Pubblicazione: (2025)
di: Chen, Xiuyuan, et al.
Pubblicazione: (2025)
Harmonic balance-automatic differentiation method: an out-of-the-box and efficient solver for general nonlinear dynamics simulation
di: Chen, Yi, et al.
Pubblicazione: (2025)
di: Chen, Yi, et al.
Pubblicazione: (2025)
A comprehensive benchmark of an Ising machine on the Max-Cut problem
di: Shaglel, Salwa, et al.
Pubblicazione: (2025)
di: Shaglel, Salwa, et al.
Pubblicazione: (2025)
LLMzSzŁ: a comprehensive LLM benchmark for Polish
di: Jassem, Krzysztof, et al.
Pubblicazione: (2025)
di: Jassem, Krzysztof, et al.
Pubblicazione: (2025)
We've Got You Covered: Rebooting American Health Careby LiranEinav and AmyFinkelstein. Penguin, 2023, 304 pp., $29 (paperback).
di: Naomi Zewde, et al.
Pubblicazione: (2025)
di: Naomi Zewde, et al.
Pubblicazione: (2025)
Continuous Prompt Generation from Linear Combination of Discrete Prompt Embeddings
di: Passigan, Pascal, et al.
Pubblicazione: (2023)
di: Passigan, Pascal, et al.
Pubblicazione: (2023)
Nonlinear Programming Solvers for Unconstrained and Constrained Optimization Problems: a Benchmark Analysis
di: Lavezzi, Giovanni, et al.
Pubblicazione: (2022)
di: Lavezzi, Giovanni, et al.
Pubblicazione: (2022)
A novel open-source ultrasound dataset with deep learning benchmarks for spinal cord injury localization and anatomical segmentation
di: Kumar, Avisha, et al.
Pubblicazione: (2024)
di: Kumar, Avisha, et al.
Pubblicazione: (2024)
Writer to Writer: How To Conference Young Authors. The Bill Harp Professional Teachers Library Series.
di: Thomason, Tommy
Pubblicazione: (1998)
di: Thomason, Tommy
Pubblicazione: (1998)
O desafio analítico da cidadania à brasileira: apontamentos para um campo de estudos cada vez mais necessário
di: Daniel Simião
Pubblicazione: (2025)
di: Daniel Simião
Pubblicazione: (2025)
Sensibilidades jurídicas e respeito às diferenças: cultura, controle e negociação de sentidos em práticas judiciais no Brasil e em Timor-Leste
di: Daniel S. Simião
Pubblicazione: (2014)
di: Daniel S. Simião
Pubblicazione: (2014)
Modeling of the positron sources: an experiment-based benchmarking
di: Alharthi, Fahad, et al.
Pubblicazione: (2025)
di: Alharthi, Fahad, et al.
Pubblicazione: (2025)
Mislabeled examples detection viewed as probing machine learning models: concepts, survey and extensive benchmark
di: George, Thomas, et al.
Pubblicazione: (2024)
di: George, Thomas, et al.
Pubblicazione: (2024)
Test signal generator for the evaluation of detection algorithms of the number of sources.
di: Raydel Ortigueira Ruiz
Pubblicazione: (2018)
di: Raydel Ortigueira Ruiz
Pubblicazione: (2018)
Efficacy of static analysis tools for software defect detection on open-source projects
di: Yeboah, Jones, et al.
Pubblicazione: (2024)
di: Yeboah, Jones, et al.
Pubblicazione: (2024)
pyMSER -- An open-source library for automatic equilibration detection in molecular simulations
di: Oliveira, Felipe Lopes, et al.
Pubblicazione: (2024)
di: Oliveira, Felipe Lopes, et al.
Pubblicazione: (2024)
CAMB: A comprehensive industrial LLM benchmark on civil aviation maintenance
di: Zhang, Feng, et al.
Pubblicazione: (2025)
di: Zhang, Feng, et al.
Pubblicazione: (2025)
A comprehensive survey of oracle character recognition: challenges, benchmarks, and beyond
di: Li, Jing, et al.
Pubblicazione: (2024)
di: Li, Jing, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Synthetic Eye Movement Dataset for Script Reading Detection: Real Trajectory Replay on a 3D Simulator
di: Zewde, Kidus, et al.
Pubblicazione: (2026) -
When the Forger Is the Judge: GPT-Image-2 Cannot Recognize Its Own Faked Documents
di: Wu, Jiaqi, et al.
Pubblicazione: (2026) -
GPT-Image-2 in the Wild: A Twitter Dataset of Self-Reported AI-Generated Images from the First Week of Deployment
di: Zewde, Kidus, et al.
Pubblicazione: (2026) -
Can Multi-modal (reasoning) LLMs detect document manipulation?
di: Liang, Zisheng, et al.
Pubblicazione: (2025) -
Can Multi-modal (reasoning) LLMs work as deepfake detectors?
di: Ren, Simiao, et al.
Pubblicazione: (2025)