Advance Fake Video Detection via Vision Transformers
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Battocchio, Joy, Dell'Anna, Stefano, Montibeller, Andrea, Boato, Giulia |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TrueFake: A Real World Case Dataset of Last Generation Fake Images also Shared on Social Networks
par: Dell'Anna, Stefano, et autres
Publié: (2025)
par: Dell'Anna, Stefano, et autres
Publié: (2025)
Backbone is All You Need: Assessing Vulnerabilities of Frozen Foundation Models in Synthetic Image Forensics
par: Musso, Chiara, et autres
Publié: (2026)
par: Musso, Chiara, et autres
Publié: (2026)
Don't Guess, Escalate: Towards Explainable Uncertainty-Calibrated AI Forensic Agents
par: Boato, Giulia, et autres
Publié: (2025)
par: Boato, Giulia, et autres
Publié: (2025)
Bridging the Gap: A Framework for Real-World Video Deepfake Detection via Social Network Compression Emulation
par: Montibeller, Andrea, et autres
Publié: (2025)
par: Montibeller, Andrea, et autres
Publié: (2025)
WILD: a new in-the-Wild Image Linkage Dataset for synthetic image attribution
par: Bongini, Pietro, et autres
Publié: (2025)
par: Bongini, Pietro, et autres
Publié: (2025)
Official-NV: An LLM-Generated News Video Dataset for Multimodal Fake News Detection
par: Wang, Yihao, et autres
Publié: (2024)
par: Wang, Yihao, et autres
Publié: (2024)
FakeParts: a New Family of AI-Generated DeepFakes
par: Liu, Ziyi, et autres
Publié: (2025)
par: Liu, Ziyi, et autres
Publié: (2025)
Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training
par: Baraldi, Lorenzo, et autres
Publié: (2023)
par: Baraldi, Lorenzo, et autres
Publié: (2023)
Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion
par: Zhang, Yinghui, et autres
Publié: (2025)
par: Zhang, Yinghui, et autres
Publié: (2025)
Parents and Children: Distinguishing Multimodal DeepFakes from Natural Images
par: Amoroso, Roberto, et autres
Publié: (2023)
par: Amoroso, Roberto, et autres
Publié: (2023)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
par: Sar, Ayan, et autres
Publié: (2025)
par: Sar, Ayan, et autres
Publié: (2025)
Unsupervised Transcript-assisted Video Summarization and Highlight Detection
par: Barbakos, Spyros, et autres
Publié: (2025)
par: Barbakos, Spyros, et autres
Publié: (2025)
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
par: Ouyang, Shuyi, et autres
Publié: (2024)
par: Ouyang, Shuyi, et autres
Publié: (2024)
Interactive Video Generation via Domain Adaptation
par: Rawal, Ishaan, et autres
Publié: (2025)
par: Rawal, Ishaan, et autres
Publié: (2025)
Compressed Deepfake Video Detection Based on 3D Spatiotemporal Trajectories
par: Chen, Zongmei, et autres
Publié: (2024)
par: Chen, Zongmei, et autres
Publié: (2024)
MSCrackMamba: Leveraging Vision Mamba for Crack Detection in Fused Multispectral Imagery
par: Zhu, Qinfeng, et autres
Publié: (2024)
par: Zhu, Qinfeng, et autres
Publié: (2024)
DIBS: Enhancing Dense Video Captioning with Unlabeled Videos via Pseudo Boundary Enrichment and Online Refinement
par: Wu, Hao, et autres
Publié: (2024)
par: Wu, Hao, et autres
Publié: (2024)
Composing Concepts from Images and Videos via Concept-prompt Binding
par: Kong, Xianghao, et autres
Publié: (2025)
par: Kong, Xianghao, et autres
Publié: (2025)
Consistency-aware Fake Videos Detection on Short Video Platforms
par: Wang, Junxi, et autres
Publié: (2025)
par: Wang, Junxi, et autres
Publié: (2025)
DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation
par: Cai, Minghong, et autres
Publié: (2024)
par: Cai, Minghong, et autres
Publié: (2024)
TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
par: Zeng, Xiangyu, et autres
Publié: (2024)
par: Zeng, Xiangyu, et autres
Publié: (2024)
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation
par: Huang, Feizhen, et autres
Publié: (2025)
par: Huang, Feizhen, et autres
Publié: (2025)
Video Seal: Open and Efficient Video Watermarking
par: Fernandez, Pierre, et autres
Publié: (2024)
par: Fernandez, Pierre, et autres
Publié: (2024)
Distilling Vision-Language Foundation Models: A Data-Free Approach via Prompt Diversification
par: Xuan, Yunyi, et autres
Publié: (2024)
par: Xuan, Yunyi, et autres
Publié: (2024)
DIRECT: Video Mashup Creation via Hierarchical Multi-Agent Planning and Intent-Guided Editing
par: Li, Ke, et autres
Publié: (2026)
par: Li, Ke, et autres
Publié: (2026)
EvRepSL: Event-Stream Representation via Self-Supervised Learning for Event-Based Vision
par: Qu, Qiang, et autres
Publié: (2024)
par: Qu, Qiang, et autres
Publié: (2024)
FedVideoMAE: Efficient Privacy-Preserving Federated Video Moderation
par: Tao, Ziyuan, et autres
Publié: (2025)
par: Tao, Ziyuan, et autres
Publié: (2025)
Enhancing Fake News Video Detection via LLM-Driven Creative Process Simulation
par: Bu, Yuyan, et autres
Publié: (2025)
par: Bu, Yuyan, et autres
Publié: (2025)
Video-EM: Event-Centric Episodic Memory for Long-Form Video Understanding
par: Wang, Yun, et autres
Publié: (2025)
par: Wang, Yun, et autres
Publié: (2025)
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
par: Mandelli, Sara, et autres
Publié: (2024)
par: Mandelli, Sara, et autres
Publié: (2024)
Moiré Video Authentication: A Physical Signature Against AI Video Generation
par: Qing, Yuan, et autres
Publié: (2026)
par: Qing, Yuan, et autres
Publié: (2026)
VIA: Unified Spatiotemporal Video Adaptation Framework for Global and Local Video Editing
par: Gu, Jing, et autres
Publié: (2024)
par: Gu, Jing, et autres
Publié: (2024)
AI-based System for Transforming text and sound to Educational Videos
par: ElAlami, M. E., et autres
Publié: (2026)
par: ElAlami, M. E., et autres
Publié: (2026)
Open-o3-Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence
par: Meng, Jiahao, et autres
Publié: (2025)
par: Meng, Jiahao, et autres
Publié: (2025)
VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
par: Bian, Yuxuan, et autres
Publié: (2025)
par: Bian, Yuxuan, et autres
Publié: (2025)
Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
par: Liu, Jiajun, et autres
Publié: (2024)
par: Liu, Jiajun, et autres
Publié: (2024)
Lightning Fast Video Anomaly Detection via Adversarial Knowledge Distillation
par: Croitoru, Florinel-Alin, et autres
Publié: (2022)
par: Croitoru, Florinel-Alin, et autres
Publié: (2022)
Recurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval
par: Caffagni, Davide, et autres
Publié: (2025)
par: Caffagni, Davide, et autres
Publié: (2025)
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
par: Yan, Xin, et autres
Publié: (2024)
par: Yan, Xin, et autres
Publié: (2024)
AToken: A Unified Tokenizer for Vision
par: Lu, Jiasen, et autres
Publié: (2025)
par: Lu, Jiasen, et autres
Publié: (2025)
Documents similaires
-
TrueFake: A Real World Case Dataset of Last Generation Fake Images also Shared on Social Networks
par: Dell'Anna, Stefano, et autres
Publié: (2025) -
Backbone is All You Need: Assessing Vulnerabilities of Frozen Foundation Models in Synthetic Image Forensics
par: Musso, Chiara, et autres
Publié: (2026) -
Don't Guess, Escalate: Towards Explainable Uncertainty-Calibrated AI Forensic Agents
par: Boato, Giulia, et autres
Publié: (2025) -
Bridging the Gap: A Framework for Real-World Video Deepfake Detection via Social Network Compression Emulation
par: Montibeller, Andrea, et autres
Publié: (2025) -
WILD: a new in-the-Wild Image Linkage Dataset for synthetic image attribution
par: Bongini, Pietro, et autres
Publié: (2025)