Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Ge Ya, Favero, Gian Mario, Luo, Zhi Hao, Jolicoeur-Martineau, Alexia, Pal, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion
by: Luo, Ge Ya, et al.
Published: (2024)
by: Luo, Ge Ya, et al.
Published: (2024)
Ctrl-Crash: Controllable Diffusion for Realistic Car Crashes
by: Gosselin, Anthony, et al.
Published: (2025)
by: Gosselin, Anthony, et al.
Published: (2025)
One Pass Is Not Enough: Recursive Latent Refinement for Generative Models
by: Esmaeilzadeh, Mehdi, et al.
Published: (2026)
by: Esmaeilzadeh, Mehdi, et al.
Published: (2026)
Multi-Agent Game Generation and Evaluation via Audio-Visual Recordings
by: Jolicoeur-Martineau, Alexia
Published: (2025)
by: Jolicoeur-Martineau, Alexia
Published: (2025)
Less is More: Recursive Reasoning with Tiny Networks
by: Jolicoeur-Martineau, Alexia
Published: (2025)
by: Jolicoeur-Martineau, Alexia
Published: (2025)
PopulAtion Parameter Averaging (PAPA)
by: Jolicoeur-Martineau, Alexia, et al.
Published: (2023)
by: Jolicoeur-Martineau, Alexia, et al.
Published: (2023)
Spatio-Temporal Conditional Diffusion Models for Forecasting Future Multiple Sclerosis Lesion Masks Conditioned on Treatments
by: Favero, Gian Mario, et al.
Published: (2025)
by: Favero, Gian Mario, et al.
Published: (2025)
AIGCBench: Comprehensive Evaluation of Image-to-Video Content Generated by AI
by: Fan, Fanda, et al.
Published: (2024)
by: Fan, Fanda, et al.
Published: (2024)
Video-As-Prompt: Unified Semantic Control for Video Generation
by: Bian, Yuxuan, et al.
Published: (2025)
by: Bian, Yuxuan, et al.
Published: (2025)
Cross-Modal Transferable Image-to-Video Attack on Video Quality Metrics
by: Gotin, Georgii, et al.
Published: (2025)
by: Gotin, Georgii, et al.
Published: (2025)
Evaluating Design Video Generation: Metrics for Compositional Fidelity
by: Deganutti, Adrienne, et al.
Published: (2026)
by: Deganutti, Adrienne, et al.
Published: (2026)
STREAM: Spatio-TempoRal Evaluation and Analysis Metric for Video Generative Models
by: Kim, Pum Jun, et al.
Published: (2024)
by: Kim, Pum Jun, et al.
Published: (2024)
VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
by: He, Xuan, et al.
Published: (2024)
by: He, Xuan, et al.
Published: (2024)
A Survey on Quality Metrics for Text-to-Image Generation
by: Hartwig, Sebastian, et al.
Published: (2024)
by: Hartwig, Sebastian, et al.
Published: (2024)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
by: Zhang, Zhihong, et al.
Published: (2025)
by: Zhang, Zhihong, et al.
Published: (2025)
VISTA: Mitigating Semantic Inertia in Video-LLMs via Training-Free Dynamic Chain-of-Thought Routing
by: Jin, Hongbo, et al.
Published: (2025)
by: Jin, Hongbo, et al.
Published: (2025)
What You See Is What Matters: A Novel Visual and Physics-Based Metric for Evaluating Video Generation Quality
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Attribute Based Interpretable Evaluation Metrics for Generative Models
by: Kim, Dongkyun, et al.
Published: (2023)
by: Kim, Dongkyun, et al.
Published: (2023)
Progressive Feature Fusion Network for Enhancing Image Quality Assessment
by: Wu, Kaiqun, et al.
Published: (2024)
by: Wu, Kaiqun, et al.
Published: (2024)
Learning to Evaluate the Artness of AI-generated Images
by: Chen, Junyu, et al.
Published: (2023)
by: Chen, Junyu, et al.
Published: (2023)
HiTVideo: Hierarchical Tokenizers for Enhancing Text-to-Video Generation with Autoregressive Large Language Models
by: Zhou, Ziqin, et al.
Published: (2025)
by: Zhou, Ziqin, et al.
Published: (2025)
SSG-Dit: A Spatial Signal Guided Framework for Controllable Video Generation
by: Hu, Peng, et al.
Published: (2025)
by: Hu, Peng, et al.
Published: (2025)
Seeing Beyond 8bits: Subjective and Objective Quality Assessment of HDR-UGC Videos
by: Saini, Shreshth, et al.
Published: (2026)
by: Saini, Shreshth, et al.
Published: (2026)
VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation
by: Jiang, Longteng, et al.
Published: (2026)
by: Jiang, Longteng, et al.
Published: (2026)
SpatialMem: Metric-Aligned Long-Horizon Video Memory for Language Grounding and QA
by: Zheng, Xinyi, et al.
Published: (2026)
by: Zheng, Xinyi, et al.
Published: (2026)
Preacher: Paper-to-Video Agentic System
by: Liu, Jingwei, et al.
Published: (2025)
by: Liu, Jingwei, et al.
Published: (2025)
Generation of Heterogeneous PET Images from Uniform Organ Activity Maps Using a Pretrained Domain-Adapted Diffusion Model
by: Li, Suya, et al.
Published: (2026)
by: Li, Suya, et al.
Published: (2026)
Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models
by: Curl, Emily, et al.
Published: (2026)
by: Curl, Emily, et al.
Published: (2026)
VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance
by: Taesiri, Mohammad Reza, et al.
Published: (2025)
by: Taesiri, Mohammad Reza, et al.
Published: (2025)
Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding
by: Luo, Bingjun, et al.
Published: (2026)
by: Luo, Bingjun, et al.
Published: (2026)
StableAnimator: High-Quality Identity-Preserving Human Image Animation
by: Tu, Shuyuan, et al.
Published: (2024)
by: Tu, Shuyuan, et al.
Published: (2024)
QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering
by: Jung, Woojun, et al.
Published: (2026)
by: Jung, Woojun, et al.
Published: (2026)
Consistent Video Editing as Flow-Driven Image-to-Video Generation
by: Wang, Ge, et al.
Published: (2025)
by: Wang, Ge, et al.
Published: (2025)
Beyond the Frame: Generating 360 Panoramic Videos from Perspective Videos
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
S2DM: Sector-Shaped Diffusion Models for Video Generation
by: Lang, Haoran, et al.
Published: (2024)
by: Lang, Haoran, et al.
Published: (2024)
BodyMetric: Evaluating the Realism of Human Bodies in Text-to-Image Generation
by: Andreou, Nefeli, et al.
Published: (2024)
by: Andreou, Nefeli, et al.
Published: (2024)
VimTS: A Unified Video and Image Text Spotter for Enhancing the Cross-domain Generalization
by: Liu, Yuliang, et al.
Published: (2024)
by: Liu, Yuliang, et al.
Published: (2024)
CHUG: Crowdsourced User-Generated HDR Video Quality Dataset
by: Saini, Shreshth, et al.
Published: (2025)
by: Saini, Shreshth, et al.
Published: (2025)
VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models
by: Li, Chenglin, et al.
Published: (2024)
by: Li, Chenglin, et al.
Published: (2024)
MemCam: Memory-Augmented Camera Control for Consistent Video Generation
by: Gao, Xinhang, et al.
Published: (2026)
by: Gao, Xinhang, et al.
Published: (2026)
Similar Items
-
Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion
by: Luo, Ge Ya, et al.
Published: (2024) -
Ctrl-Crash: Controllable Diffusion for Realistic Car Crashes
by: Gosselin, Anthony, et al.
Published: (2025) -
One Pass Is Not Enough: Recursive Latent Refinement for Generative Models
by: Esmaeilzadeh, Mehdi, et al.
Published: (2026) -
Multi-Agent Game Generation and Evaluation via Audio-Visual Recordings
by: Jolicoeur-Martineau, Alexia
Published: (2025) -
Less is More: Recursive Reasoning with Tiny Networks
by: Jolicoeur-Martineau, Alexia
Published: (2025)