ActivityForensics: A Comprehensive Benchmark for Localizing Manipulated Activity in Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Bao, Peijun, Luo, Anwei, Pan, Gang, Kot, Alex C., Jiang, Xudong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SimBase: A Simple Baseline for Temporal Video Grounding
by: Bao, Peijun, et al.
Published: (2024)
by: Bao, Peijun, et al.
Published: (2024)
ForensicsSAM: Toward Robust and Unified Image Forgery Detection and Localization Resisting to Adversarial Attack
by: Peng, Rongxuan, et al.
Published: (2025)
by: Peng, Rongxuan, et al.
Published: (2025)
Open-Set Deepfake Detection: A Parameter-Efficient Adaptation Method with Forgery Style Mixture
by: Kong, Chenqi, et al.
Published: (2024)
by: Kong, Chenqi, et al.
Published: (2024)
MoE-FFD: Mixture of Experts for Generalized and Parameter-Efficient Face Forgery Detection
by: Kong, Chenqi, et al.
Published: (2024)
by: Kong, Chenqi, et al.
Published: (2024)
Pixel-Inconsistency Modeling for Image Manipulation Localization
by: Kong, Chenqi, et al.
Published: (2023)
by: Kong, Chenqi, et al.
Published: (2023)
Vid-Morp: Video Moment Retrieval Pretraining from Unlabeled Videos in the Wild
by: Bao, Peijun, et al.
Published: (2024)
by: Bao, Peijun, et al.
Published: (2024)
UAL-Bench: The First Comprehensive Unusual Activity Localization Benchmark
by: Abdullah, Hasnat Md, et al.
Published: (2024)
by: Abdullah, Hasnat Md, et al.
Published: (2024)
Generalized Face Forgery Detection via Adaptive Learning for Pre-trained Vision Transformer
by: Luo, Anwei, et al.
Published: (2023)
by: Luo, Anwei, et al.
Published: (2023)
Explainable Forensics of Manipulated Segments in Untrimmed Long Videos
by: Feng, Yue, et al.
Published: (2026)
by: Feng, Yue, et al.
Published: (2026)
IMDL-BenCo: A Comprehensive Benchmark and Codebase for Image Manipulation Detection & Localization
by: Ma, Xiaochen, et al.
Published: (2024)
by: Ma, Xiaochen, et al.
Published: (2024)
Forensics-Bench: A Comprehensive Forgery Detection Benchmark Suite for Large Vision Language Models
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
MMFusion: Combining Image Forensic Filters for Visual Manipulation Detection and Localization
by: Triaridis, Kostas, et al.
Published: (2023)
by: Triaridis, Kostas, et al.
Published: (2023)
GREx: Generalized Referring Expression Segmentation, Comprehension, and Generation
by: Ding, Henghui, et al.
Published: (2026)
by: Ding, Henghui, et al.
Published: (2026)
GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization
by: Chen, Yirui, et al.
Published: (2024)
by: Chen, Yirui, et al.
Published: (2024)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
by: Li, Zhaoxu, et al.
Published: (2026)
by: Li, Zhaoxu, et al.
Published: (2026)
Open-set Anomaly Segmentation in Complex Scenarios
by: Xia, Song, et al.
Published: (2025)
by: Xia, Song, et al.
Published: (2025)
Propose and Rectify: A Forensics-Driven MLLM Framework for Image Manipulation Localization
by: Zhang, Keyang, et al.
Published: (2025)
by: Zhang, Keyang, et al.
Published: (2025)
SynthForensics: Benchmarking and Evaluating People-Centric Synthetic Video Deepfakes
by: Leotta, Roberto, et al.
Published: (2026)
by: Leotta, Roberto, et al.
Published: (2026)
MVFNet: Multipurpose Video Forensics Network using Multiple Forms of Forensic Evidence
by: Nguyen, Tai D., et al.
Published: (2025)
by: Nguyen, Tai D., et al.
Published: (2025)
Active Adversarial Noise Suppression for Image Forgery Localization
by: Peng, Rongxuan, et al.
Published: (2025)
by: Peng, Rongxuan, et al.
Published: (2025)
From Pretrain to Pain: Adversarial Vulnerability of Video Foundation Models Without Task Knowledge
by: Lu, Hui, et al.
Published: (2025)
by: Lu, Hui, et al.
Published: (2025)
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
by: Ma, Xiaochen, et al.
Published: (2023)
by: Ma, Xiaochen, et al.
Published: (2023)
MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
by: Li, Kunchang, et al.
Published: (2023)
by: Li, Kunchang, et al.
Published: (2023)
Celeb-DF++: A Large-scale Challenging Video DeepFake Benchmark for Generalizable Forensics
by: Li, Yuezun, et al.
Published: (2025)
by: Li, Yuezun, et al.
Published: (2025)
ForensicHub: A Unified Benchmark & Codebase for All-Domain Fake Image Detection and Localization
by: Du, Bo, et al.
Published: (2025)
by: Du, Bo, et al.
Published: (2025)
Semi-Supervised Pipe Video Temporal Defect Interval Localization
by: Huang, Zhu, et al.
Published: (2024)
by: Huang, Zhu, et al.
Published: (2024)
Evolving Storytelling: Benchmarks and Methods for New Character Customization with Diffusion Models
by: Wang, Xiyu, et al.
Published: (2024)
by: Wang, Xiyu, et al.
Published: (2024)
Benchmarking Joint Face Spoofing and Forgery Detection with Visual and Physiological Cues
by: Yu, Zitong, et al.
Published: (2022)
by: Yu, Zitong, et al.
Published: (2022)
Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection
by: Zhang, Yuchen, et al.
Published: (2026)
by: Zhang, Yuchen, et al.
Published: (2026)
Color Space Learning for Cross-Color Person Re-Identification
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
Single-Image Shadow Removal Using Deep Learning: A Comprehensive Survey
by: Guo, Laniqng, et al.
Published: (2024)
by: Guo, Laniqng, et al.
Published: (2024)
MeViS: A Multi-Modal Dataset for Referring Motion Expression Video Segmentation
by: Ding, Henghui, et al.
Published: (2025)
by: Ding, Henghui, et al.
Published: (2025)
Hiding Local Manipulations on SAR Images: a Counter-Forensic Attack
by: Mandelli, Sara, et al.
Published: (2024)
by: Mandelli, Sara, et al.
Published: (2024)
Uncertainty-boosted Robust Video Activity Anticipation
by: Qi, Zhaobo, et al.
Published: (2024)
by: Qi, Zhaobo, et al.
Published: (2024)
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
by: Zhang, Yifei, et al.
Published: (2023)
by: Zhang, Yifei, et al.
Published: (2023)
GigaHands: A Massive Annotated Dataset of Bimanual Hand Activities
by: Fu, Rao, et al.
Published: (2024)
by: Fu, Rao, et al.
Published: (2024)
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
by: Li, Zhaoxu, et al.
Published: (2025)
by: Li, Zhaoxu, et al.
Published: (2025)
MMRel: Benchmarking Relation Understanding in Multi-Modal Large Language Models
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
Fast Low-parameter Video Activity Localization in Collaborative Learning Environments
by: Jatla, Venkatesh, et al.
Published: (2024)
by: Jatla, Venkatesh, et al.
Published: (2024)
Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
Similar Items
-
SimBase: A Simple Baseline for Temporal Video Grounding
by: Bao, Peijun, et al.
Published: (2024) -
ForensicsSAM: Toward Robust and Unified Image Forgery Detection and Localization Resisting to Adversarial Attack
by: Peng, Rongxuan, et al.
Published: (2025) -
Open-Set Deepfake Detection: A Parameter-Efficient Adaptation Method with Forgery Style Mixture
by: Kong, Chenqi, et al.
Published: (2024) -
MoE-FFD: Mixture of Experts for Generalized and Parameter-Efficient Face Forgery Detection
by: Kong, Chenqi, et al.
Published: (2024) -
Pixel-Inconsistency Modeling for Image Manipulation Localization
by: Kong, Chenqi, et al.
Published: (2023)