FAVOR-Bench: A Comprehensive Benchmark for Fine-Grained Video Motion Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tu, Chongjun, Zhang, Lin, Chen, Pengtao, Ye, Peng, Zeng, Xianfang, Cheng, Wei, Yu, Gang, Chen, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
MotionAgent: Fine-grained Controllable Video Generation via Motion Field Agent
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
MikuDance: Animating Character Art with Mixed Motion Dynamics
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
KPM-Bench: A Kinematic Parsing Motion Benchmark for Fine-grained Motion-centric Video Understanding
von: Lin, Boda, et al.
Veröffentlicht: (2026)
von: Lin, Boda, et al.
Veröffentlicht: (2026)
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
von: Zhuang, Cailin, et al.
Veröffentlicht: (2025)
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
OneIG-Bench: Omni-dimensional Nuanced Evaluation for Image Generation
von: Chang, Jingjing, et al.
Veröffentlicht: (2025)
von: Chang, Jingjing, et al.
Veröffentlicht: (2025)
$Δ$-DiT: A Training-Free Acceleration Method Tailored for Diffusion Transformers
von: Chen, Pengtao, et al.
Veröffentlicht: (2024)
von: Chen, Pengtao, et al.
Veröffentlicht: (2024)
FineState-Bench: A Comprehensive Benchmark for Fine-Grained State Control in GUI Agents
von: Ji, Fengxian, et al.
Veröffentlicht: (2025)
von: Ji, Fengxian, et al.
Veröffentlicht: (2025)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
InstructionBench: An Instructional Video Understanding Benchmark
von: Wei, Haiwan, et al.
Veröffentlicht: (2025)
von: Wei, Haiwan, et al.
Veröffentlicht: (2025)
GEditBench v2: A Human-Aligned Benchmark for General Image Editing
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2026)
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2026)
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation
von: Yu, Hong-Tao, et al.
Veröffentlicht: (2025)
von: Yu, Hong-Tao, et al.
Veröffentlicht: (2025)
KRIS-Bench: Benchmarking Next-Level Intelligent Image Editing Models
von: Wu, Yongliang, et al.
Veröffentlicht: (2025)
von: Wu, Yongliang, et al.
Veröffentlicht: (2025)
VERIFIED: A Video Corpus Moment Retrieval Benchmark for Fine-Grained Video Understanding
von: Chen, Houlun, et al.
Veröffentlicht: (2024)
von: Chen, Houlun, et al.
Veröffentlicht: (2024)
Q-Bench-Video: Benchmarking the Video Quality Understanding of LMMs
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
YourSkatingCoach: A Figure Skating Video Benchmark for Fine-Grained Element Analysis
von: Chen, Wei-Yi, et al.
Veröffentlicht: (2024)
von: Chen, Wei-Yi, et al.
Veröffentlicht: (2024)
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning
von: Zhu, Minghao, et al.
Veröffentlicht: (2023)
von: Zhu, Minghao, et al.
Veröffentlicht: (2023)
RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
von: Zhang, Jun, et al.
Veröffentlicht: (2025)
ActionArt: Advancing Multimodal Large Models for Fine-Grained Human-Centric Video Understanding
von: Peng, Yi-Xing, et al.
Veröffentlicht: (2025)
von: Peng, Yi-Xing, et al.
Veröffentlicht: (2025)
BrokenVideos: A Benchmark Dataset for Fine-Grained Artifact Localization in AI-Generated Videos
von: Lin, Jiahao, et al.
Veröffentlicht: (2025)
von: Lin, Jiahao, et al.
Veröffentlicht: (2025)
In-Context Learning with Unpaired Clips for Instruction-based Video Editing
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
von: Liao, Xinyao, et al.
Veröffentlicht: (2025)
LongInsightBench: A Comprehensive Benchmark for Evaluating Omni-Modal Models on Human-Centric Long-Video Understanding
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
PresentBench: A Fine-Grained Rubric-Based Benchmark for Slide Generation
von: Chen, Xin-Sheng, et al.
Veröffentlicht: (2026)
von: Chen, Xin-Sheng, et al.
Veröffentlicht: (2026)
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation
von: Li, Shengze, et al.
Veröffentlicht: (2024)
von: Li, Shengze, et al.
Veröffentlicht: (2024)
FineVAU: A Novel Human-Aligned Benchmark for Fine-Grained Video Anomaly Understanding
von: Pereira, João, et al.
Veröffentlicht: (2026)
von: Pereira, João, et al.
Veröffentlicht: (2026)
TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation
von: Tong, Haibo, et al.
Veröffentlicht: (2025)
von: Tong, Haibo, et al.
Veröffentlicht: (2025)
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
MotionSight: Boosting Fine-Grained Motion Understanding in Multimodal LLMs
von: Du, Yipeng, et al.
Veröffentlicht: (2025)
von: Du, Yipeng, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Human Motion Video Captioning
von: Song, Guorui, et al.
Veröffentlicht: (2025)
von: Song, Guorui, et al.
Veröffentlicht: (2025)
TokenCarve: Information-Preserving Visual Token Compression in Multimodal Large Language Models
von: Tan, Xudong, et al.
Veröffentlicht: (2025)
von: Tan, Xudong, et al.
Veröffentlicht: (2025)
M-ErasureBench: A Comprehensive Multimodal Evaluation Benchmark for Concept Erasure in Diffusion Models
von: Weng, Ju-Hsuan, et al.
Veröffentlicht: (2025)
von: Weng, Ju-Hsuan, et al.
Veröffentlicht: (2025)
CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
von: Xu, Yifan, et al.
Veröffentlicht: (2024)
VEU-Bench: Towards Comprehensive Understanding of Video Editing
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
LiViBench: An Omnimodal Benchmark for Interactive Livestream Video Understanding
von: Wang, Xiaodong, et al.
Veröffentlicht: (2026)
von: Wang, Xiaodong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
von: Chen, Pengtao, et al.
Veröffentlicht: (2025) -
MotionAgent: Fine-grained Controllable Video Generation via Motion Field Agent
von: Liao, Xinyao, et al.
Veröffentlicht: (2025) -
MikuDance: Animating Character Art with Mixed Motion Dynamics
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024) -
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
von: Chen, Pengtao, et al.
Veröffentlicht: (2025) -
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)