DVAR: Adversarial Multi-Agent Debate for Video Authenticity Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Qi, Hongyuan, Shao, Feifei, Li, Ming, Fan, Hehe, Xiao, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deepfake Detection Generalization with Diffusion Noise
by: Qi, Hongyuan, et al.
Published: (2026)
by: Qi, Hongyuan, et al.
Published: (2026)
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
by: Chen, Kerui, et al.
Published: (2026)
by: Chen, Kerui, et al.
Published: (2026)
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
by: Xu, Zhiyu, et al.
Published: (2025)
by: Xu, Zhiyu, et al.
Published: (2025)
MMAD: Multi-label Micro-Action Detection in Videos
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
by: Yang, Xiangpeng, et al.
Published: (2025)
by: Yang, Xiangpeng, et al.
Published: (2025)
EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing
by: Yang, Xiangpeng, et al.
Published: (2024)
by: Yang, Xiangpeng, et al.
Published: (2024)
TV-Dialogue: Crafting Theme-Aware Video Dialogues with Immersive Interaction
by: Wang, Sai, et al.
Published: (2025)
by: Wang, Sai, et al.
Published: (2025)
BVINet: Unlocking Blind Video Inpainting with Zero Annotations
by: Wu, Zhiliang, et al.
Published: (2025)
by: Wu, Zhiliang, et al.
Published: (2025)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
Rendering Multi-Human and Multi-Object with 3D Gaussian Splatting
by: Wang, Weiquan, et al.
Published: (2026)
by: Wang, Weiquan, et al.
Published: (2026)
Structured Universal Adversarial Attacks on Object Detection for Video Sequences
by: Jacob, Sven, et al.
Published: (2025)
by: Jacob, Sven, et al.
Published: (2025)
TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal Estimation
by: Li, Mingwei, et al.
Published: (2026)
by: Li, Mingwei, et al.
Published: (2026)
CoMo: Compositional Motion Customization for Text-to-Video Generation
by: Xu, Youcan, et al.
Published: (2025)
by: Xu, Youcan, et al.
Published: (2025)
InfiniDreamer: Arbitrarily Long Human Motion Generation via Segment Score Distillation
by: Zhuo, Wenjie, et al.
Published: (2024)
by: Zhuo, Wenjie, et al.
Published: (2024)
ClusterStyle: Modeling Intra-Style Diversity with Prototypical Clustering for Stylized Motion Generation
by: Chen, Kerui, et al.
Published: (2025)
by: Chen, Kerui, et al.
Published: (2025)
MICAS: Multi-grained In-Context Adaptive Sampling for 3D Point Cloud Processing
by: Shao, Feifei, et al.
Published: (2024)
by: Shao, Feifei, et al.
Published: (2024)
Incentivizing Generative Zero-Shot Learning via Outcome-Reward Reinforcement Learning with Visual Cues
by: Hou, Wenjin, et al.
Published: (2026)
by: Hou, Wenjin, et al.
Published: (2026)
DeMoGen: Towards Decompositional Human Motion Generation with Energy-Based Diffusion Models
by: Zhang, Jianrong, et al.
Published: (2025)
by: Zhang, Jianrong, et al.
Published: (2025)
EnergyMoGen: Compositional Human Motion Generation with Energy-Based Diffusion Model in Latent Space
by: Zhang, Jianrong, et al.
Published: (2024)
by: Zhang, Jianrong, et al.
Published: (2024)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
by: Xu, Youcan, et al.
Published: (2026)
by: Xu, Youcan, et al.
Published: (2026)
$\text{H}^2$em: Learning Hierarchical Hyperbolic Embeddings for Compositional Zero-Shot Learning
by: Li, Lin, et al.
Published: (2025)
by: Li, Lin, et al.
Published: (2025)
Knowledge-guided Causal Intervention for Weakly-supervised Object Localization
by: Shao, Feifei, et al.
Published: (2023)
by: Shao, Feifei, et al.
Published: (2023)
VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation
by: Zhuo, Wenjie, et al.
Published: (2024)
by: Zhuo, Wenjie, et al.
Published: (2024)
Uncertainty-Aware 4D Gaussian Splatting for Monocular Occluded Human Rendering
by: Wang, Weiquan, et al.
Published: (2026)
by: Wang, Weiquan, et al.
Published: (2026)
Unmasking Deep Fakes: Leveraging Deep Learning for Video Authenticity Detection
by: Hasan, Mahmudul, et al.
Published: (2025)
by: Hasan, Mahmudul, et al.
Published: (2025)
FaVChat: Hierarchical Prompt-Query Guided Facial Video Understanding with Data-Efficient GRPO
by: Zhao, Fufangchen, et al.
Published: (2025)
by: Zhao, Fufangchen, et al.
Published: (2025)
GMFL-Net: A Global Multi-geometric Feature Learning Network for Repetitive Action Counting
by: Li, Jun, et al.
Published: (2024)
by: Li, Jun, et al.
Published: (2024)
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment
by: Li, Wei, et al.
Published: (2024)
by: Li, Wei, et al.
Published: (2024)
Concealed Object Detection
by: Fan, Deng-Ping, et al.
Published: (2021)
by: Fan, Deng-Ping, et al.
Published: (2021)
Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework
by: Liu, Ziyang, et al.
Published: (2026)
by: Liu, Ziyang, et al.
Published: (2026)
Let's Reward Step-by-Step: Step-Aware Contrastive Alignment for Vision-Language Navigation in Continuous Environments
by: Li, Haoyuan, et al.
Published: (2026)
by: Li, Haoyuan, et al.
Published: (2026)
Training-Free and Interpretable Hateful Video Detection via Multi-stage Adversarial Reasoning
by: Yang, Shuonan, et al.
Published: (2026)
by: Yang, Shuonan, et al.
Published: (2026)
Multi-Paradigm Collaborative Adversarial Attack Against Multi-Modal Large Language Models
by: Li, Yuanbo, et al.
Published: (2026)
by: Li, Yuanbo, et al.
Published: (2026)
VideoAgent: A Memory-augmented Multimodal Agent for Video Understanding
by: Fan, Yue, et al.
Published: (2024)
by: Fan, Yue, et al.
Published: (2024)
HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting
by: Zhou, Zhenglin, et al.
Published: (2024)
by: Zhou, Zhenglin, et al.
Published: (2024)
TSGS: Improving Gaussian Splatting for Transparent Surface Reconstruction via Normal and De-lighting Priors
by: Li, Mingwei, et al.
Published: (2025)
by: Li, Mingwei, et al.
Published: (2025)
A Reinforcement Learning-Based Automatic Video Editing Method Using Pre-trained Vision-Language Model
by: Hu, Panwen, et al.
Published: (2024)
by: Hu, Panwen, et al.
Published: (2024)
AdvGPS: Adversarial GPS for Multi-Agent Perception Attack
by: Li, Jinlong, et al.
Published: (2024)
by: Li, Jinlong, et al.
Published: (2024)
A Multi-task Adversarial Attack Against Face Authentication
by: Wang, Hanrui, et al.
Published: (2024)
by: Wang, Hanrui, et al.
Published: (2024)
4DPC$^2$hat: Towards Dynamic Point Cloud Understanding with Failure-Aware Bootstrapping
by: Zhang, Xindan, et al.
Published: (2026)
by: Zhang, Xindan, et al.
Published: (2026)
Similar Items
-
Deepfake Detection Generalization with Diffusion Noise
by: Qi, Hongyuan, et al.
Published: (2026) -
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
by: Chen, Kerui, et al.
Published: (2026) -
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
by: Xu, Zhiyu, et al.
Published: (2025) -
MMAD: Multi-label Micro-Action Detection in Videos
by: Li, Kun, et al.
Published: (2024) -
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
by: Yang, Xiangpeng, et al.
Published: (2025)