DVAR: Adversarial Multi-Agent Debate for Video Authenticity Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Qi, Hongyuan, Shao, Feifei, Li, Ming, Fan, Hehe, Xiao, Jun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deepfake Detection Generalization with Diffusion Noise
por: Qi, Hongyuan, et al.
Publicado: (2026)
por: Qi, Hongyuan, et al.
Publicado: (2026)
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
por: Chen, Kerui, et al.
Publicado: (2026)
por: Chen, Kerui, et al.
Publicado: (2026)
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
por: Xu, Zhiyu, et al.
Publicado: (2025)
por: Xu, Zhiyu, et al.
Publicado: (2025)
MMAD: Multi-label Micro-Action Detection in Videos
por: Li, Kun, et al.
Publicado: (2024)
por: Li, Kun, et al.
Publicado: (2024)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
por: Yang, Xiangpeng, et al.
Publicado: (2025)
por: Yang, Xiangpeng, et al.
Publicado: (2025)
EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing
por: Yang, Xiangpeng, et al.
Publicado: (2024)
por: Yang, Xiangpeng, et al.
Publicado: (2024)
TV-Dialogue: Crafting Theme-Aware Video Dialogues with Immersive Interaction
por: Wang, Sai, et al.
Publicado: (2025)
por: Wang, Sai, et al.
Publicado: (2025)
BVINet: Unlocking Blind Video Inpainting with Zero Annotations
por: Wu, Zhiliang, et al.
Publicado: (2025)
por: Wu, Zhiliang, et al.
Publicado: (2025)
Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion
por: Zhou, Zhenglin, et al.
Publicado: (2025)
por: Zhou, Zhenglin, et al.
Publicado: (2025)
Rendering Multi-Human and Multi-Object with 3D Gaussian Splatting
por: Wang, Weiquan, et al.
Publicado: (2026)
por: Wang, Weiquan, et al.
Publicado: (2026)
Structured Universal Adversarial Attacks on Object Detection for Video Sequences
por: Jacob, Sven, et al.
Publicado: (2025)
por: Jacob, Sven, et al.
Publicado: (2025)
TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal Estimation
por: Li, Mingwei, et al.
Publicado: (2026)
por: Li, Mingwei, et al.
Publicado: (2026)
CoMo: Compositional Motion Customization for Text-to-Video Generation
por: Xu, Youcan, et al.
Publicado: (2025)
por: Xu, Youcan, et al.
Publicado: (2025)
InfiniDreamer: Arbitrarily Long Human Motion Generation via Segment Score Distillation
por: Zhuo, Wenjie, et al.
Publicado: (2024)
por: Zhuo, Wenjie, et al.
Publicado: (2024)
ClusterStyle: Modeling Intra-Style Diversity with Prototypical Clustering for Stylized Motion Generation
por: Chen, Kerui, et al.
Publicado: (2025)
por: Chen, Kerui, et al.
Publicado: (2025)
MICAS: Multi-grained In-Context Adaptive Sampling for 3D Point Cloud Processing
por: Shao, Feifei, et al.
Publicado: (2024)
por: Shao, Feifei, et al.
Publicado: (2024)
Incentivizing Generative Zero-Shot Learning via Outcome-Reward Reinforcement Learning with Visual Cues
por: Hou, Wenjin, et al.
Publicado: (2026)
por: Hou, Wenjin, et al.
Publicado: (2026)
DeMoGen: Towards Decompositional Human Motion Generation with Energy-Based Diffusion Models
por: Zhang, Jianrong, et al.
Publicado: (2025)
por: Zhang, Jianrong, et al.
Publicado: (2025)
EnergyMoGen: Compositional Human Motion Generation with Energy-Based Diffusion Model in Latent Space
por: Zhang, Jianrong, et al.
Publicado: (2024)
por: Zhang, Jianrong, et al.
Publicado: (2024)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
por: Xu, Youcan, et al.
Publicado: (2026)
por: Xu, Youcan, et al.
Publicado: (2026)
$\text{H}^2$em: Learning Hierarchical Hyperbolic Embeddings for Compositional Zero-Shot Learning
por: Li, Lin, et al.
Publicado: (2025)
por: Li, Lin, et al.
Publicado: (2025)
Knowledge-guided Causal Intervention for Weakly-supervised Object Localization
por: Shao, Feifei, et al.
Publicado: (2023)
por: Shao, Feifei, et al.
Publicado: (2023)
VividDreamer: Invariant Score Distillation For Hyper-Realistic Text-to-3D Generation
por: Zhuo, Wenjie, et al.
Publicado: (2024)
por: Zhuo, Wenjie, et al.
Publicado: (2024)
Uncertainty-Aware 4D Gaussian Splatting for Monocular Occluded Human Rendering
por: Wang, Weiquan, et al.
Publicado: (2026)
por: Wang, Weiquan, et al.
Publicado: (2026)
Unmasking Deep Fakes: Leveraging Deep Learning for Video Authenticity Detection
por: Hasan, Mahmudul, et al.
Publicado: (2025)
por: Hasan, Mahmudul, et al.
Publicado: (2025)
FaVChat: Hierarchical Prompt-Query Guided Facial Video Understanding with Data-Efficient GRPO
por: Zhao, Fufangchen, et al.
Publicado: (2025)
por: Zhao, Fufangchen, et al.
Publicado: (2025)
GMFL-Net: A Global Multi-geometric Feature Learning Network for Repetitive Action Counting
por: Li, Jun, et al.
Publicado: (2024)
por: Li, Jun, et al.
Publicado: (2024)
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment
por: Li, Wei, et al.
Publicado: (2024)
por: Li, Wei, et al.
Publicado: (2024)
Concealed Object Detection
por: Fan, Deng-Ping, et al.
Publicado: (2021)
por: Fan, Deng-Ping, et al.
Publicado: (2021)
Towards Universal Physical Adversarial Attacks via a Joint Multi-Objective and Multi-Model Optimization Framework
por: Liu, Ziyang, et al.
Publicado: (2026)
por: Liu, Ziyang, et al.
Publicado: (2026)
Let's Reward Step-by-Step: Step-Aware Contrastive Alignment for Vision-Language Navigation in Continuous Environments
por: Li, Haoyuan, et al.
Publicado: (2026)
por: Li, Haoyuan, et al.
Publicado: (2026)
Training-Free and Interpretable Hateful Video Detection via Multi-stage Adversarial Reasoning
por: Yang, Shuonan, et al.
Publicado: (2026)
por: Yang, Shuonan, et al.
Publicado: (2026)
Multi-Paradigm Collaborative Adversarial Attack Against Multi-Modal Large Language Models
por: Li, Yuanbo, et al.
Publicado: (2026)
por: Li, Yuanbo, et al.
Publicado: (2026)
VideoAgent: A Memory-augmented Multimodal Agent for Video Understanding
por: Fan, Yue, et al.
Publicado: (2024)
por: Fan, Yue, et al.
Publicado: (2024)
HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting
por: Zhou, Zhenglin, et al.
Publicado: (2024)
por: Zhou, Zhenglin, et al.
Publicado: (2024)
TSGS: Improving Gaussian Splatting for Transparent Surface Reconstruction via Normal and De-lighting Priors
por: Li, Mingwei, et al.
Publicado: (2025)
por: Li, Mingwei, et al.
Publicado: (2025)
A Reinforcement Learning-Based Automatic Video Editing Method Using Pre-trained Vision-Language Model
por: Hu, Panwen, et al.
Publicado: (2024)
por: Hu, Panwen, et al.
Publicado: (2024)
AdvGPS: Adversarial GPS for Multi-Agent Perception Attack
por: Li, Jinlong, et al.
Publicado: (2024)
por: Li, Jinlong, et al.
Publicado: (2024)
A Multi-task Adversarial Attack Against Face Authentication
por: Wang, Hanrui, et al.
Publicado: (2024)
por: Wang, Hanrui, et al.
Publicado: (2024)
4DPC$^2$hat: Towards Dynamic Point Cloud Understanding with Failure-Aware Bootstrapping
por: Zhang, Xindan, et al.
Publicado: (2026)
por: Zhang, Xindan, et al.
Publicado: (2026)
Ejemplares similares
-
Deepfake Detection Generalization with Diffusion Noise
por: Qi, Hongyuan, et al.
Publicado: (2026) -
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
por: Chen, Kerui, et al.
Publicado: (2026) -
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
por: Xu, Zhiyu, et al.
Publicado: (2025) -
MMAD: Multi-label Micro-Action Detection in Videos
por: Li, Kun, et al.
Publicado: (2024) -
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
por: Yang, Xiangpeng, et al.
Publicado: (2025)