DynamicEval: Rethinking Evaluation for Dynamic Text-to-Video Synthesis
Fuente:
arXiv
Salvato in:
| Autori principali: | Babu, Nithin C., Mahapatra, Aniruddha, Rangwani, Harsh, Soundararajan, Rajiv, Kulkarni, Kuldeep |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Composing Parts for Expressive Object Generation
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
Learning from Limited and Imperfect Data
di: Rangwani, Harsh
Pubblicazione: (2024)
di: Rangwani, Harsh
Pubblicazione: (2024)
Learning from Limited and Imperfect Data
di: Rangwani, Harsh
Pubblicazione: (2025)
di: Rangwani, Harsh
Pubblicazione: (2025)
Factorized Motion Fields for Fast Sparse Input Dynamic View Synthesis
di: Somraj, Nagabhushan, et al.
Pubblicazione: (2024)
di: Somraj, Nagabhushan, et al.
Pubblicazione: (2024)
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)
DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models
di: Wang, Juntong, et al.
Pubblicazione: (2026)
di: Wang, Juntong, et al.
Pubblicazione: (2026)
Selective Mixup Fine-Tuning for Optimizing Non-Decomposable Objectives
di: Ramasubramanian, Shrinivas, et al.
Pubblicazione: (2024)
di: Ramasubramanian, Shrinivas, et al.
Pubblicazione: (2024)
Simple-RF: Regularizing Sparse Input Radiance Fields with Simpler Solutions
di: Somraj, Nagabhushan, et al.
Pubblicazione: (2024)
di: Somraj, Nagabhushan, et al.
Pubblicazione: (2024)
AD-GS: Alternating Densification for Sparse-Input 3D Gaussian Splatting
di: Patle, Gurutva, et al.
Pubblicazione: (2025)
di: Patle, Gurutva, et al.
Pubblicazione: (2025)
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
di: Liu, Zhixuan, et al.
Pubblicazione: (2026)
di: Liu, Zhixuan, et al.
Pubblicazione: (2026)
Text2Place: Affordance-aware Text Guided Human Placement
di: Parihar, Rishubh, et al.
Pubblicazione: (2024)
di: Parihar, Rishubh, et al.
Pubblicazione: (2024)
UnDIVE: Generalized Underwater Video Enhancement Using Generative Priors
di: Srinath, Suhas, et al.
Pubblicazione: (2024)
di: Srinath, Suhas, et al.
Pubblicazione: (2024)
Object-WIPER : Training-Free Object and Associated Effect Removal in Videos
di: Kushwaha, Saksham Singh, et al.
Pubblicazione: (2026)
di: Kushwaha, Saksham Singh, et al.
Pubblicazione: (2026)
Evaluation of Text-to-Video Generation Models: A Dynamics Perspective
di: Liao, Mingxiang, et al.
Pubblicazione: (2024)
di: Liao, Mingxiang, et al.
Pubblicazione: (2024)
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
di: Phung, Quynh, et al.
Pubblicazione: (2026)
di: Phung, Quynh, et al.
Pubblicazione: (2026)
ImPoster: Text and Frequency Guidance for Subject Driven Action Personalization using Diffusion Models
di: Kothandaraman, Divya, et al.
Pubblicazione: (2024)
di: Kothandaraman, Divya, et al.
Pubblicazione: (2024)
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models
di: S, Sridhar, et al.
Pubblicazione: (2025)
di: S, Sridhar, et al.
Pubblicazione: (2025)
Decoupling Dynamic Monocular Videos for Dynamic View Synthesis
di: You, Meng, et al.
Pubblicazione: (2023)
di: You, Meng, et al.
Pubblicazione: (2023)
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation
di: Ma, Wentao, et al.
Pubblicazione: (2025)
di: Ma, Wentao, et al.
Pubblicazione: (2025)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
LIME-Eval: Rethinking Low-light Image Enhancement Evaluation via Object Detection
di: Li, Mingjia, et al.
Pubblicazione: (2024)
di: Li, Mingjia, et al.
Pubblicazione: (2024)
Rethinking Human Evaluation Protocol for Text-to-Video Models: Enhancing Reliability,Reproducibility, and Practicality
di: Zhang, Tianle, et al.
Pubblicazione: (2024)
di: Zhang, Tianle, et al.
Pubblicazione: (2024)
SceneEval: Evaluating Semantic Coherence in Text-Conditioned 3D Indoor Scene Synthesis
di: Tam, Hou In Ivan, et al.
Pubblicazione: (2025)
di: Tam, Hou In Ivan, et al.
Pubblicazione: (2025)
Video-Oasis: Rethinking Evaluation of Video Understanding
di: Lim, Geuntaek, et al.
Pubblicazione: (2026)
di: Lim, Geuntaek, et al.
Pubblicazione: (2026)
Eval3D: Interpretable and Fine-grained Evaluation for 3D Generation
di: Duggal, Shivam, et al.
Pubblicazione: (2025)
di: Duggal, Shivam, et al.
Pubblicazione: (2025)
On the Content Bias in Fréchet Video Distance
di: Ge, Songwei, et al.
Pubblicazione: (2024)
di: Ge, Songwei, et al.
Pubblicazione: (2024)
Evaluating Text-to-Image and Text-to-Video Synthesis with a Conditional Fréchet Distance
di: Koo, Jaywon, et al.
Pubblicazione: (2025)
di: Koo, Jaywon, et al.
Pubblicazione: (2025)
EvalCrafter: Benchmarking and Evaluating Large Video Generation Models
di: Liu, Yaofang, et al.
Pubblicazione: (2023)
di: Liu, Yaofang, et al.
Pubblicazione: (2023)
Enhanced Rooftop Solar Panel Detection by Efficiently Aggregating Local Features
di: Kurte, Kuldeep, et al.
Pubblicazione: (2025)
di: Kurte, Kuldeep, et al.
Pubblicazione: (2025)
MotionCanvas: Cinematic Shot Design with Controllable Image-to-Video Generation
di: Xing, Jinbo, et al.
Pubblicazione: (2025)
di: Xing, Jinbo, et al.
Pubblicazione: (2025)
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
di: Mahapatra, Aniruddha, et al.
Pubblicazione: (2026)
di: Mahapatra, Aniruddha, et al.
Pubblicazione: (2026)
VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model
di: Li, Xinhao, et al.
Pubblicazione: (2024)
di: Li, Xinhao, et al.
Pubblicazione: (2024)
Dynamic Reflections: Probing Video Representations with Text Alignment
di: Zhu, Tyler, et al.
Pubblicazione: (2025)
di: Zhu, Tyler, et al.
Pubblicazione: (2025)
Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification
di: Sharan, S P, et al.
Pubblicazione: (2024)
di: Sharan, S P, et al.
Pubblicazione: (2024)
Progressive Growing of Video Tokenizers for Temporally Compact Latent Spaces
di: Mahapatra, Aniruddha, et al.
Pubblicazione: (2025)
di: Mahapatra, Aniruddha, et al.
Pubblicazione: (2025)
Physion-Eval: Evaluating Physical Realism in Generated Video via Human Reasoning
di: Zhang, Qin, et al.
Pubblicazione: (2026)
di: Zhang, Qin, et al.
Pubblicazione: (2026)
T2I-FineEval: Fine-Grained Compositional Metric for Text-to-Image Evaluation
di: Hosseini, Seyed Mohammad Hadi, et al.
Pubblicazione: (2025)
di: Hosseini, Seyed Mohammad Hadi, et al.
Pubblicazione: (2025)
FlashEval: Towards Fast and Accurate Evaluation of Text-to-image Diffusion Generative Models
di: Zhao, Lin, et al.
Pubblicazione: (2024)
di: Zhao, Lin, et al.
Pubblicazione: (2024)
Rethinking Video-Text Understanding: Retrieval from Counterfactually Augmented Data
di: Ma, Wufei, et al.
Pubblicazione: (2024)
di: Ma, Wufei, et al.
Pubblicazione: (2024)
NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Composing Parts for Expressive Object Generation
di: Rangwani, Harsh, et al.
Pubblicazione: (2024) -
Learning from Limited and Imperfect Data
di: Rangwani, Harsh
Pubblicazione: (2024) -
Learning from Limited and Imperfect Data
di: Rangwani, Harsh
Pubblicazione: (2025) -
Factorized Motion Fields for Fast Sparse Input Dynamic View Synthesis
di: Somraj, Nagabhushan, et al.
Pubblicazione: (2024) -
DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasets
di: Rangwani, Harsh, et al.
Pubblicazione: (2024)