A Survey of AI-Generated Video Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xiao, Xiang, Xinhao, Li, Zizhong, Wang, Yongheng, Li, Zhuoheng, Liu, Zhuosheng, Zhang, Weidi, Ye, Weiqi, Zhang, Jiawei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AIGVE-Tool: AI-Generated Video Evaluation Toolkit with Multifaceted Benchmark
by: Xiang, Xinhao, et al.
Published: (2025)
by: Xiang, Xinhao, et al.
Published: (2025)
Are AI-Generated Driving Videos Ready for Autonomous Driving? A Diagnostic Evaluation Framework
by: Xiang, Xinhao, et al.
Published: (2025)
by: Xiang, Xinhao, et al.
Published: (2025)
MMViR: A Multi-Modal and Multi-Granularity Representation for Long-range Video Understanding
by: Li, Zizhong, et al.
Published: (2026)
by: Li, Zizhong, et al.
Published: (2026)
AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation
by: Liu, Xiao, et al.
Published: (2025)
by: Liu, Xiao, et al.
Published: (2025)
Are Video Generation Models Geographically Fair? An Attraction-Centric Evaluation of Global Visual Knowledge
by: Liu, Xiao, et al.
Published: (2026)
by: Liu, Xiao, et al.
Published: (2026)
TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction
by: Zheng, Zhijie, et al.
Published: (2026)
by: Zheng, Zhijie, et al.
Published: (2026)
EffiPerception: an Efficient Framework for Various Perception Tasks
by: Xiang, Xinhao, et al.
Published: (2024)
by: Xiang, Xinhao, et al.
Published: (2024)
VQ-Insight: Teaching VLMs for AI-Generated Video Quality Understanding via Progressive Visual Reinforcement Learning
by: Zhang, Xuanyu, et al.
Published: (2025)
by: Zhang, Xuanyu, et al.
Published: (2025)
VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model
by: Li, Xinhao, et al.
Published: (2024)
by: Li, Xinhao, et al.
Published: (2024)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
OccFace: Unified Occlusion-Aware Facial Landmark Detection with Per-Point Visibility
by: Xiang, Xinhao, et al.
Published: (2026)
by: Xiang, Xinhao, et al.
Published: (2026)
OmniDrag: Enabling Motion Control for Omnidirectional Image-to-Video Generation
by: Li, Weiqi, et al.
Published: (2024)
by: Li, Weiqi, et al.
Published: (2024)
A Survey of Interactive Generative Video
by: Yu, Jiwen, et al.
Published: (2025)
by: Yu, Jiwen, et al.
Published: (2025)
A Survey: Spatiotemporal Consistency in Video Generation
by: Yin, Zhiyu, et al.
Published: (2025)
by: Yin, Zhiyu, et al.
Published: (2025)
G$^2$TR: Generation-Guided Visual Token Reduction for Separate-Encoder Unified Multimodal Models
by: Li, Junxian, et al.
Published: (2026)
by: Li, Junxian, et al.
Published: (2026)
ACD: Direct Conditional Control for Video Diffusion Models via Attention Supervision
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
Multi-Sentence Grounding for Long-term Instructional Video
by: Li, Zeqian, et al.
Published: (2023)
by: Li, Zeqian, et al.
Published: (2023)
Controllable Video Generation: A Survey
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
by: Hu, Kairui, et al.
Published: (2025)
by: Hu, Kairui, et al.
Published: (2025)
Universal Video Temporal Grounding with Generative Multi-modal Large Language Models
by: Li, Zeqian, et al.
Published: (2025)
by: Li, Zeqian, et al.
Published: (2025)
Human Motion Video Generation: A Survey
by: Xue, Haiwei, et al.
Published: (2025)
by: Xue, Haiwei, et al.
Published: (2025)
Grounded Question-Answering in Long Egocentric Videos
by: Di, Shangzhe, et al.
Published: (2023)
by: Di, Shangzhe, et al.
Published: (2023)
DARK: Denoising, Amplification, Restoration Kit
by: Li, Zhuoheng, et al.
Published: (2024)
by: Li, Zhuoheng, et al.
Published: (2024)
Shadow Generation for Composite Image Using Diffusion model
by: Liu, Qingyang, et al.
Published: (2024)
by: Liu, Qingyang, et al.
Published: (2024)
ZeroI2V: Zero-Cost Adaptation of Pre-trained Transformers from Image to Video
by: Li, Xinhao, et al.
Published: (2023)
by: Li, Xinhao, et al.
Published: (2023)
Generalizable Detection of AI Generated Images with Large Models and Fuzzy Decision Tree
by: Wu, Fei, et al.
Published: (2026)
by: Wu, Fei, et al.
Published: (2026)
Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
Improved Adversarial Diffusion Compression for Real-World Video Super-Resolution
by: Chen, Bin, et al.
Published: (2026)
by: Chen, Bin, et al.
Published: (2026)
Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos
by: Tang, Yuqi, et al.
Published: (2026)
by: Tang, Yuqi, et al.
Published: (2026)
DenseStep2M: A Scalable, Training-Free Pipeline for Dense Instructional Video Annotation
by: Ge, Mingji, et al.
Published: (2026)
by: Ge, Mingji, et al.
Published: (2026)
UVE: Are MLLMs Unified Evaluators for AI-Generated Videos?
by: Liu, Yuanxin, et al.
Published: (2025)
by: Liu, Yuanxin, et al.
Published: (2025)
Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos
by: Chen, Qirui, et al.
Published: (2024)
by: Chen, Qirui, et al.
Published: (2024)
HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
Thinking with Frames: Generative Video Distortion Evaluation via Frame Reward Model
by: Wang, Yuan, et al.
Published: (2026)
by: Wang, Yuan, et al.
Published: (2026)
A Comprehensive Survey on World Models for Embodied AI
by: Li, Xinqing, et al.
Published: (2025)
by: Li, Xinqing, et al.
Published: (2025)
Deep Learning For Point Cloud Denoising: A Survey
by: Zhang, Chengwei, et al.
Published: (2025)
by: Zhang, Chengwei, et al.
Published: (2025)
VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking
by: Meng, Desen, et al.
Published: (2025)
by: Meng, Desen, et al.
Published: (2025)
Extended to Reality: Prompt Injection in 3D Environments
by: Li, Zhuoheng, et al.
Published: (2026)
by: Li, Zhuoheng, et al.
Published: (2026)
Survey on AI-Generated Media Detection: From Non-MLLM to MLLM
by: Zou, Yueying, et al.
Published: (2025)
by: Zou, Yueying, et al.
Published: (2025)
VideoMamba: State Space Model for Efficient Video Understanding
by: Li, Kunchang, et al.
Published: (2024)
by: Li, Kunchang, et al.
Published: (2024)
Similar Items
-
AIGVE-Tool: AI-Generated Video Evaluation Toolkit with Multifaceted Benchmark
by: Xiang, Xinhao, et al.
Published: (2025) -
Are AI-Generated Driving Videos Ready for Autonomous Driving? A Diagnostic Evaluation Framework
by: Xiang, Xinhao, et al.
Published: (2025) -
MMViR: A Multi-Modal and Multi-Granularity Representation for Long-range Video Understanding
by: Li, Zizhong, et al.
Published: (2026) -
AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation
by: Liu, Xiao, et al.
Published: (2025) -
Are Video Generation Models Geographically Fair? An Attraction-Centric Evaluation of Global Visual Knowledge
by: Liu, Xiao, et al.
Published: (2026)