A Physical Coherence Benchmark for Evaluating Video Generation Models via Optical Flow-guided Frame Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yongfan, Zhu, Xiuwen, Li, Tianyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
Optical Flow Representation Alignment Mamba Diffusion Model for Medical Video Generation
von: Wang, Zhenbin, et al.
Veröffentlicht: (2024)
von: Wang, Zhenbin, et al.
Veröffentlicht: (2024)
Landmark-guided Diffusion Model for High-fidelity and Temporally Coherent Talking Head Generation
von: Tan, Jintao, et al.
Veröffentlicht: (2024)
von: Tan, Jintao, et al.
Veröffentlicht: (2024)
Detecting AI-Generated Video via Frame Consistency
von: Ma, Long, et al.
Veröffentlicht: (2024)
von: Ma, Long, et al.
Veröffentlicht: (2024)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models
von: Li, Chenglin, et al.
Veröffentlicht: (2024)
von: Li, Chenglin, et al.
Veröffentlicht: (2024)
Benefits of Feature Extraction and Temporal Sequence Analysis for Video Frame Prediction: An Evaluation of Hybrid Deep Learning Models
von: Velázquez, Jose M. Sánchez, et al.
Veröffentlicht: (2025)
von: Velázquez, Jose M. Sánchez, et al.
Veröffentlicht: (2025)
Optical-Flow Guided Prompt Optimization for Coherent Video Generation
von: Nam, Hyelin, et al.
Veröffentlicht: (2024)
von: Nam, Hyelin, et al.
Veröffentlicht: (2024)
AEGIS: Authenticity Evaluation Benchmark for AI-Generated Video Sequences
von: Li, Jieyu, et al.
Veröffentlicht: (2025)
von: Li, Jieyu, et al.
Veröffentlicht: (2025)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
Video-Bench: Human-Aligned Video Generation Benchmark
von: Han, Hui, et al.
Veröffentlicht: (2025)
von: Han, Hui, et al.
Veröffentlicht: (2025)
KFS-Bench: Comprehensive Evaluation of Key Frame Sampling in Long Video Understanding
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
SCBench: A Sports Commentary Benchmark for Video LLMs
von: Ge, Kuangzhi, et al.
Veröffentlicht: (2024)
von: Ge, Kuangzhi, et al.
Veröffentlicht: (2024)
CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
von: Deng, Yufan, et al.
Veröffentlicht: (2025)
Playing with Transformer at 30+ FPS via Next-Frame Diffusion
von: Cheng, Xinle, et al.
Veröffentlicht: (2025)
von: Cheng, Xinle, et al.
Veröffentlicht: (2025)
VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation
von: Jiang, Longteng, et al.
Veröffentlicht: (2026)
von: Jiang, Longteng, et al.
Veröffentlicht: (2026)
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
von: Jang, Sangwon, et al.
Veröffentlicht: (2025)
CoAgent: Collaborative Planning and Consistency Agent for Coherent Video Generation
von: Zeng, Qinglin, et al.
Veröffentlicht: (2025)
von: Zeng, Qinglin, et al.
Veröffentlicht: (2025)
MF-LPR$^2$: Multi-Frame License Plate Image Restoration and Recognition using Optical Flow
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
Next Block Prediction: Video Generation via Semi-Autoregressive Modeling
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
von: Ren, Shuhuai, et al.
Veröffentlicht: (2025)
SLVMEval: Synthetic Meta Evaluation Benchmark for Text-to-Long Video Generation
von: Matsuda, Ryosuke, et al.
Veröffentlicht: (2026)
von: Matsuda, Ryosuke, et al.
Veröffentlicht: (2026)
VFIMamba: Video Frame Interpolation with State Space Models
von: Zhang, Guozhen, et al.
Veröffentlicht: (2024)
von: Zhang, Guozhen, et al.
Veröffentlicht: (2024)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
von: Chen, Zhifei, et al.
Veröffentlicht: (2025)
von: Chen, Zhifei, et al.
Veröffentlicht: (2025)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
PipeFlow: Pipelined Processing and Motion-Aware Frame Selection for Long-Form Video Editing
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
Intriguing Differences Between Zero-Shot and Systematic Evaluations of Vision-Language Transformer Models
von: Salman, Shaeke, et al.
Veröffentlicht: (2024)
von: Salman, Shaeke, et al.
Veröffentlicht: (2024)
PhysicsMind: Sim and Real Mechanics Benchmarking for Physical Reasoning and Prediction in Foundational VLMs and World Models
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
von: Feng, Huidong, et al.
Veröffentlicht: (2026)
von: Feng, Huidong, et al.
Veröffentlicht: (2026)
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
DySink: Dynamic Frame Sinks for Autoregressive Long Video Generation
von: Ye, Bo, et al.
Veröffentlicht: (2026)
von: Ye, Bo, et al.
Veröffentlicht: (2026)
MEMFOF: High-Resolution Training for Memory-Efficient Multi-Frame Optical Flow Estimation
von: Bargatin, Vladislav, et al.
Veröffentlicht: (2025)
von: Bargatin, Vladislav, et al.
Veröffentlicht: (2025)
OSCBench: Benchmarking Object State Change in Text-to-Video Generation
von: Han, Xianjing, et al.
Veröffentlicht: (2026)
von: Han, Xianjing, et al.
Veröffentlicht: (2026)
Fostering Video Reasoning via Next-Event Prediction
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
von: Wang, Haonan, et al.
Veröffentlicht: (2025)
Culture In a Frame: C$^3$B as a Comic-Based Benchmark for Multimodal Culturally Awareness
von: Song, Yuchen, et al.
Veröffentlicht: (2025)
von: Song, Yuchen, et al.
Veröffentlicht: (2025)
Infinite-World: Scaling Interactive World Models to 1000-Frame Horizons via Pose-Free Hierarchical Memory
von: Wu, Ruiqi, et al.
Veröffentlicht: (2026)
von: Wu, Ruiqi, et al.
Veröffentlicht: (2026)
ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images
von: Kong, Xianghao, et al.
Veröffentlicht: (2025)
von: Kong, Xianghao, et al.
Veröffentlicht: (2025)
NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
von: Li, Hongyu, et al.
Veröffentlicht: (2025)
MambaFlow: A Novel and Flow-guided State Space Model for Scene Flow Estimation
von: Luo, Jiehao, et al.
Veröffentlicht: (2025)
von: Luo, Jiehao, et al.
Veröffentlicht: (2025)
VideoLLM Benchmarks and Evaluation: A Survey
von: Kumar, Yogesh
Veröffentlicht: (2025)
von: Kumar, Yogesh
Veröffentlicht: (2025)
Ähnliche Einträge
-
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
von: Ji, Longbin, et al.
Veröffentlicht: (2026) -
Optical Flow Representation Alignment Mamba Diffusion Model for Medical Video Generation
von: Wang, Zhenbin, et al.
Veröffentlicht: (2024) -
Landmark-guided Diffusion Model for High-fidelity and Temporally Coherent Talking Head Generation
von: Tan, Jintao, et al.
Veröffentlicht: (2024) -
Detecting AI-Generated Video via Frame Consistency
von: Ma, Long, et al.
Veröffentlicht: (2024) -
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)