A Timely Survey on Vision Transformer for Deepfake Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhikan, Cheng, Zhongyao, Xiong, Jiajie, Xu, Xun, Li, Tianrui, Veeravalli, Bharadwaj, Yang, Xulei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training
di: Lin, Jinpeng, et al.
Pubblicazione: (2024)
di: Lin, Jinpeng, et al.
Pubblicazione: (2024)
Exploiting Vision Language Model for Training-Free 3D Point Cloud OOD Detection via Graph Score Propagation
di: Chen, Tiankai, et al.
Pubblicazione: (2025)
di: Chen, Tiankai, et al.
Pubblicazione: (2025)
Global-Aware Monocular Semantic Scene Completion with State Space Models
di: Li, Shijie, et al.
Pubblicazione: (2025)
di: Li, Shijie, et al.
Pubblicazione: (2025)
On-the-fly Point Feature Representation for Point Clouds Analysis
di: Wang, Jiangyi, et al.
Pubblicazione: (2024)
di: Wang, Jiangyi, et al.
Pubblicazione: (2024)
Fus-MAE: A cross-attention-based data fusion approach for Masked Autoencoders in remote sensing
di: Chan-To-Hing, Hugo, et al.
Pubblicazione: (2024)
di: Chan-To-Hing, Hugo, et al.
Pubblicazione: (2024)
Integrating SAM Supervision for 3D Weakly Supervised Point Cloud Segmentation
di: You, Lechun, et al.
Pubblicazione: (2025)
di: You, Lechun, et al.
Pubblicazione: (2025)
CAT: Exploiting Inter-Class Dynamics for Domain Adaptive Object Detection
di: Kennerley, Mikhail, et al.
Pubblicazione: (2024)
di: Kennerley, Mikhail, et al.
Pubblicazione: (2024)
Exploring Human-in-the-Loop Test-Time Adaptation by Synergizing Active Learning and Model Selection
di: Li, Yushu, et al.
Pubblicazione: (2024)
di: Li, Yushu, et al.
Pubblicazione: (2024)
Multi-View Industrial Anomaly Detection with Epipolar Constrained Cross-View Fusion
di: Liu, Yifan, et al.
Pubblicazione: (2025)
di: Liu, Yifan, et al.
Pubblicazione: (2025)
LaVida Drive: Vision-Text Interaction VLM for Autonomous Driving with Token Selection, Recovery and Enhancement
di: Jiao, Siwen, et al.
Pubblicazione: (2024)
di: Jiao, Siwen, et al.
Pubblicazione: (2024)
Deepfake Generation and Detection: A Benchmark and Survey
di: Pei, Gan, et al.
Pubblicazione: (2024)
di: Pei, Gan, et al.
Pubblicazione: (2024)
Future-Aware Interaction Network For Motion Forecasting
di: Li, Shijie, et al.
Pubblicazione: (2025)
di: Li, Shijie, et al.
Pubblicazione: (2025)
Utilizing the Mean Teacher with Supcontrast Loss for Wafer Pattern Recognition
di: Wei, Qiyu, et al.
Pubblicazione: (2024)
di: Wei, Qiyu, et al.
Pubblicazione: (2024)
Robust Distribution Alignment for Industrial Anomaly Detection under Distribution Shift
di: Liao, Jingyi, et al.
Pubblicazione: (2025)
di: Liao, Jingyi, et al.
Pubblicazione: (2025)
Joint Architecture-Token-Bitwidth Multi-Axis Optimization of Vision Transformers for Semiconductor IC Packaging
di: Nguyen, Phat, et al.
Pubblicazione: (2026)
di: Nguyen, Phat, et al.
Pubblicazione: (2026)
Exploring Self-Supervised Vision Transformers for Deepfake Detection: A Comparative Analysis
di: Nguyen, Huy H., et al.
Pubblicazione: (2024)
di: Nguyen, Huy H., et al.
Pubblicazione: (2024)
Detecting Lip-Syncing Deepfakes: Vision Temporal Transformer for Analyzing Mouth Inconsistencies
di: Datta, Soumyya Kanti, et al.
Pubblicazione: (2025)
di: Datta, Soumyya Kanti, et al.
Pubblicazione: (2025)
LPViT: Low-Power Semi-structured Pruning for Vision Transformers
di: Xu, Kaixin, et al.
Pubblicazione: (2024)
di: Xu, Kaixin, et al.
Pubblicazione: (2024)
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer
di: Deressa, Deressa Wodajo, et al.
Pubblicazione: (2023)
di: Deressa, Deressa Wodajo, et al.
Pubblicazione: (2023)
AD-FM: Multimodal LLMs for Anomaly Detection via Multi-Stage Reasoning and Fine-Grained Reward Optimization
di: Liao, Jingyi, et al.
Pubblicazione: (2025)
di: Liao, Jingyi, et al.
Pubblicazione: (2025)
An Efficient 3D Convolutional Neural Network with Channel-wise, Spatial-grouped, and Temporal Convolutions
di: Wang, Zhe, et al.
Pubblicazione: (2025)
di: Wang, Zhe, et al.
Pubblicazione: (2025)
CERSA: Cumulative Energy-Retaining Subspace Adaptation for Memory-Efficient Fine-Tuning
di: Ge, Jingze, et al.
Pubblicazione: (2026)
di: Ge, Jingze, et al.
Pubblicazione: (2026)
Deepfake Detection: A Comprehensive Survey from the Reliability Perspective
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
Image Recognition with Online Lightweight Vision Transformer: A Survey
di: Zhang, Zherui, et al.
Pubblicazione: (2025)
di: Zhang, Zherui, et al.
Pubblicazione: (2025)
On the Adversarial Risk of Test Time Adaptation: An Investigation into Realistic Test-Time Data Poisoning
di: Su, Yongyi, et al.
Pubblicazione: (2024)
di: Su, Yongyi, et al.
Pubblicazione: (2024)
Learning Real Facial Concepts for Independent Deepfake Detection
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
Scaling Laws for Deepfake Detection
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
di: Wang, Wenhao, et al.
Pubblicazione: (2025)
Chain-of-Adaptation: Surgical Vision-Language Adaptation with Reinforcement Learning
di: Li, Jiajie, et al.
Pubblicazione: (2026)
di: Li, Jiajie, et al.
Pubblicazione: (2026)
Revisiting Audio-Visual Segmentation with Vision-Centric Transformer
di: Huang, Shaofei, et al.
Pubblicazione: (2025)
di: Huang, Shaofei, et al.
Pubblicazione: (2025)
Suppressing Gradient Conflict for Generalizable Deepfake Detection
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
di: Liu, Ming-Hui, et al.
Pubblicazione: (2025)
Unleashing Vision-Language Semantics for Deepfake Video Detection
di: Zhu, Jiawen, et al.
Pubblicazione: (2026)
di: Zhu, Jiawen, et al.
Pubblicazione: (2026)
VELVET-Med: Vision and Efficient Language Pre-training for Volumetric Imaging Tasks in Medicine
di: Zhang, Ziyang, et al.
Pubblicazione: (2025)
di: Zhang, Ziyang, et al.
Pubblicazione: (2025)
Real-Time Deepfake Detection in the Real-World
di: Cavia, Bar, et al.
Pubblicazione: (2024)
di: Cavia, Bar, et al.
Pubblicazione: (2024)
MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
di: Xu, Wenbo, et al.
Pubblicazione: (2026)
di: Xu, Wenbo, et al.
Pubblicazione: (2026)
Knowledge-Guided Prompt Learning for Deepfake Facial Image Detection
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
Refine-and-Contrast: Adaptive Instance-Aware BEV Representations for Multi-UAV Collaborative Object Detection
di: Li, Zhongyao, et al.
Pubblicazione: (2025)
di: Li, Zhongyao, et al.
Pubblicazione: (2025)
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
di: Fang, Hongwei, et al.
Pubblicazione: (2026)
di: Fang, Hongwei, et al.
Pubblicazione: (2026)
Patch-Discontinuity Mining for Generalized Deepfake Detection
di: Yuan, Huanhuan, et al.
Pubblicazione: (2025)
di: Yuan, Huanhuan, et al.
Pubblicazione: (2025)
TALL: Thumbnail Layout for Deepfake Video Detection
di: Xu, Yuting, et al.
Pubblicazione: (2023)
di: Xu, Yuting, et al.
Pubblicazione: (2023)
Zero-Shot 3D Visual Grounding from Vision-Language Models
di: Li, Rong, et al.
Pubblicazione: (2025)
di: Li, Rong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training
di: Lin, Jinpeng, et al.
Pubblicazione: (2024) -
Exploiting Vision Language Model for Training-Free 3D Point Cloud OOD Detection via Graph Score Propagation
di: Chen, Tiankai, et al.
Pubblicazione: (2025) -
Global-Aware Monocular Semantic Scene Completion with State Space Models
di: Li, Shijie, et al.
Pubblicazione: (2025) -
On-the-fly Point Feature Representation for Point Clouds Analysis
di: Wang, Jiangyi, et al.
Pubblicazione: (2024) -
Fus-MAE: A cross-attention-based data fusion approach for Masked Autoencoders in remote sensing
di: Chan-To-Hing, Hugo, et al.
Pubblicazione: (2024)