Seeing What Matters: Generalizable AI-generated Video Detection with Forensic-Oriented Augmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Corvi, Riccardo, Cozzolino, Davide, Prashnani, Ekta, De Mello, Shalini, Nagano, Koki, Verdoliva, Luisa |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Synthetic Image Verification in the Era of Generative AI: What Works and What Isn't There Yet
by: Tariang, Diangarti, et al.
Published: (2024)
by: Tariang, Diangarti, et al.
Published: (2024)
Raising the Bar of AI-generated Image Detection with CLIP
by: Cozzolino, Davide, et al.
Published: (2023)
by: Cozzolino, Davide, et al.
Published: (2023)
Avatar Fingerprinting for Authorized Use of Synthetic Talking-Head Videos
by: Prashnani, Ekta, et al.
Published: (2023)
by: Prashnani, Ekta, et al.
Published: (2023)
M3Dsynth: A dataset of medical 3D images with AI-generated local manipulations
by: Zingarini, Giada, et al.
Published: (2023)
by: Zingarini, Giada, et al.
Published: (2023)
Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection
by: De Rosa, Vincenzo, et al.
Published: (2024)
by: De Rosa, Vincenzo, et al.
Published: (2024)
Quality-Aware Calibration for AI-Generated Image Detection in the Wild
by: Guillaro, Fabrizio, et al.
Published: (2026)
by: Guillaro, Fabrizio, et al.
Published: (2026)
Zero-Shot Detection of AI-Generated Images
by: Cozzolino, Davide, et al.
Published: (2024)
by: Cozzolino, Davide, et al.
Published: (2024)
A Bias-Free Training Paradigm for More General AI-generated Image Detection
by: Guillaro, Fabrizio, et al.
Published: (2024)
by: Guillaro, Fabrizio, et al.
Published: (2024)
Training-Free Deepfake Voice Recognition by Leveraging Large-Scale Pre-Trained Models
by: Pianese, Alessandro, et al.
Published: (2024)
by: Pianese, Alessandro, et al.
Published: (2024)
AI-GenBench: A New Ongoing Benchmark for AI-Generated Image Detection
by: Pellegrini, Lorenzo, et al.
Published: (2025)
by: Pellegrini, Lorenzo, et al.
Published: (2025)
Unmasking Puppeteers: Leveraging Biometric Leakage to Expose Impersonation in AI-Based Videoconferencing
by: Vahdati, Danial Samadi, et al.
Published: (2025)
by: Vahdati, Danial Samadi, et al.
Published: (2025)
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
by: Trevithick, Alex, et al.
Published: (2024)
by: Trevithick, Alex, et al.
Published: (2024)
Don't Guess, Escalate: Towards Explainable Uncertainty-Calibrated AI Forensic Agents
by: Boato, Giulia, et al.
Published: (2025)
by: Boato, Giulia, et al.
Published: (2025)
Instant Expressive Gaussian Head Avatar via 3D-Aware Expression Distillation
by: Jiang, Kaiwen, et al.
Published: (2025)
by: Jiang, Kaiwen, et al.
Published: (2025)
COSY: Compositional 3DGS Synthesis for Disentangled Human Head Editing
by: Barthel, Florian, et al.
Published: (2026)
by: Barthel, Florian, et al.
Published: (2026)
Coherent3D: Coherent 3D Portrait Video Reconstruction via Triplane Fusion
by: Wang, Shengze, et al.
Published: (2024)
by: Wang, Shengze, et al.
Published: (2024)
Coherent 3D Portrait Video Reconstruction via Triplane Fusion
by: Wang, Shengze, et al.
Published: (2024)
by: Wang, Shengze, et al.
Published: (2024)
A Unified Approach for Text- and Image-guided 4D Scene Generation
by: Zheng, Yufeng, et al.
Published: (2023)
by: Zheng, Yufeng, et al.
Published: (2023)
GAvatar: Animatable 3D Gaussian Avatars with Implicit Mesh Learning
by: Yuan, Ye, et al.
Published: (2023)
by: Yuan, Ye, et al.
Published: (2023)
VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents
by: Mazumdar, Amrita, et al.
Published: (2026)
by: Mazumdar, Amrita, et al.
Published: (2026)
BLADE: Single-view Body Mesh Learning through Accurate Depth Estimation
by: Wang, Shengze, et al.
Published: (2024)
by: Wang, Shengze, et al.
Published: (2024)
See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model
by: Feng, Yixu, et al.
Published: (2026)
by: Feng, Yixu, et al.
Published: (2026)
Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection
by: Zhang, Yuchen, et al.
Published: (2026)
by: Zhang, Yuchen, et al.
Published: (2026)
83‐1: Invited Paper: AI 3D Selfie: Real‐Time Single‐Image 3D Face Reconstruction for Light‐Field Displays
by: Jonghyun Kim, et al.
Published: (2025)
by: Jonghyun Kim, et al.
Published: (2025)
When Detectors Forget Forensics: Blocking Semantic Shortcuts for Generalizable AI-Generated Image Detection
by: Shuai, Chao, et al.
Published: (2026)
by: Shuai, Chao, et al.
Published: (2026)
Do You See What I Say? Generalizable Deepfake Detection based on Visual Speech Recognition
by: Bora, Maheswar, et al.
Published: (2025)
by: Bora, Maheswar, et al.
Published: (2025)
What You See Is What Matters: A Novel Visual and Physics-Based Metric for Evaluating Video Generation Quality
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
FrameOracle: Learning What to See and How Much to See in Videos
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
Forensics Adapter: Unleashing CLIP for Generalizable Face Forgery Detection
by: Cui, Xinjie, et al.
Published: (2024)
by: Cui, Xinjie, et al.
Published: (2024)
Generalizable and Adaptive Continual Learning Framework for AI-generated Image Detection
by: Wang, Hanyi, et al.
Published: (2026)
by: Wang, Hanyi, et al.
Published: (2026)
Deepfake Forensics Adapter: A Dual-Stream Network for Generalizable Deepfake Detection
by: Liao, Jianfeng, et al.
Published: (2026)
by: Liao, Jianfeng, et al.
Published: (2026)
Cutup and Detect: Human Fall Detection on Cutup Untrimmed Videos Using a Large Foundational Video Understanding Model
by: Grutschus, Till, et al.
Published: (2024)
by: Grutschus, Till, et al.
Published: (2024)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
by: Pei, Gensheng, et al.
Published: (2025)
by: Pei, Gensheng, et al.
Published: (2025)
Celeb-DF++: A Large-scale Challenging Video DeepFake Benchmark for Generalizable Forensics
by: Li, Yuezun, et al.
Published: (2025)
by: Li, Yuezun, et al.
Published: (2025)
Seeing What Matters: Visual Preference Policy Optimization for Visual Generation
by: Ni, Ziqi, et al.
Published: (2025)
by: Ni, Ziqi, et al.
Published: (2025)
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
by: De Nadai, Marco, et al.
Published: (2025)
by: De Nadai, Marco, et al.
Published: (2025)
QUEEN: QUantized Efficient ENcoding of Dynamic Gaussians for Streaming Free-viewpoint Videos
by: Girish, Sharath, et al.
Published: (2024)
by: Girish, Sharath, et al.
Published: (2024)
Seeing Before Reasoning: A Unified Framework for Generalizable and Explainable Fake Image Detection
by: Lin, Kaiqing, et al.
Published: (2025)
by: Lin, Kaiqing, et al.
Published: (2025)
Exploring the Robustness of AI-Driven Tools in Digital Forensics: A Preliminary Study
by: Sanna, Silvia Lucia, et al.
Published: (2024)
by: Sanna, Silvia Lucia, et al.
Published: (2024)
Color Matters: Demosaicing-Guided Color Correlation Training for Generalizable AI-Generated Image Detection
by: Zhong, Nan, et al.
Published: (2026)
by: Zhong, Nan, et al.
Published: (2026)
Similar Items
-
Synthetic Image Verification in the Era of Generative AI: What Works and What Isn't There Yet
by: Tariang, Diangarti, et al.
Published: (2024) -
Raising the Bar of AI-generated Image Detection with CLIP
by: Cozzolino, Davide, et al.
Published: (2023) -
Avatar Fingerprinting for Authorized Use of Synthetic Talking-Head Videos
by: Prashnani, Ekta, et al.
Published: (2023) -
M3Dsynth: A dataset of medical 3D images with AI-generated local manipulations
by: Zingarini, Giada, et al.
Published: (2023) -
Exploring the Adversarial Robustness of CLIP for AI-generated Image Detection
by: De Rosa, Vincenzo, et al.
Published: (2024)