VEU-Bench: Towards Comprehensive Understanding of Video Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Bozheng, Wu, Yongliang, Lu, Yi, Yu, Jiashuo, Tang, Licheng, Cao, Jiawang, Zhu, Wenqing, Sun, Yuyang, Wu, Jay, Zhu, Wenbo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought
von: Lu, Yi, et al.
Veröffentlicht: (2025)
von: Lu, Yi, et al.
Veröffentlicht: (2025)
Zero-Shot Long-Form Video Understanding through Screenplay
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
OpusAnimation: Code-Based Dynamic Chart Generation
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
Reframe Anything: LLM Agent for Open World Video Reframing
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)
Adaptive Dense Evidence Refinement for Video Relational Reasoning for VRR-QA Challenge
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
KRIS-Bench: Benchmarking Next-Level Intelligent Image Editing Models
von: Wu, Yongliang, et al.
Veröffentlicht: (2025)
von: Wu, Yongliang, et al.
Veröffentlicht: (2025)
Number it: Temporal Grounding Videos like Flipping Manga
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
von: Wu, Yongliang, et al.
Veröffentlicht: (2024)
Temporal Evidence Routing with Structured Visual Evidence for TimeLogicQA
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
GEditBench v2: A Human-Aligned Benchmark for General Image Editing
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2026)
von: Jiang, Zhangqi, et al.
Veröffentlicht: (2026)
SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding
von: Fu, Chaoyou, et al.
Veröffentlicht: (2026)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2026)
U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
Omni-WorldBench: Towards a Comprehensive Interaction-Centric Evaluation for World Models
von: Wu, Meiqi, et al.
Veröffentlicht: (2026)
von: Wu, Meiqi, et al.
Veröffentlicht: (2026)
E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding
von: Liu, Ye, et al.
Veröffentlicht: (2024)
von: Liu, Ye, et al.
Veröffentlicht: (2024)
VidText: Towards Comprehensive Evaluation for Video Text Understanding
von: Yang, Zhoufaran, et al.
Veröffentlicht: (2025)
von: Yang, Zhoufaran, et al.
Veröffentlicht: (2025)
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model
von: Xu, Lu, et al.
Veröffentlicht: (2024)
von: Xu, Lu, et al.
Veröffentlicht: (2024)
Q-Bench-Video: Benchmarking the Video Quality Understanding of LMMs
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zicheng, et al.
Veröffentlicht: (2024)
NarrLV: Towards a Comprehensive Narrative-Centric Evaluation for Long Video Generation
von: Feng, X., et al.
Veröffentlicht: (2025)
von: Feng, X., et al.
Veröffentlicht: (2025)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihong, et al.
Veröffentlicht: (2025)
FAVOR-Bench: A Comprehensive Benchmark for Fine-Grained Video Motion Understanding
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
von: Tu, Chongjun, et al.
Veröffentlicht: (2025)
VTEdit-Bench: A Comprehensive Benchmark for Multi-Reference Image Editing Models in Virtual Try-On
von: Liang, Xiaoye, et al.
Veröffentlicht: (2026)
von: Liang, Xiaoye, et al.
Veröffentlicht: (2026)
Envisioning Class Entity Reasoning by Large Language Models for Few-shot Learning
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
ExpVid: A Benchmark for Experiment Video Understanding & Reasoning
von: Xu, Yicheng, et al.
Veröffentlicht: (2025)
von: Xu, Yicheng, et al.
Veröffentlicht: (2025)
MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence
von: Lin, Jingli, et al.
Veröffentlicht: (2025)
von: Lin, Jingli, et al.
Veröffentlicht: (2025)
ICE-Bench: A Unified and Comprehensive Benchmark for Image Creating and Editing
von: Pan, Yulin, et al.
Veröffentlicht: (2025)
von: Pan, Yulin, et al.
Veröffentlicht: (2025)
Vidi: Large Multimodal Models for Video Understanding and Editing
von: Vidi Team, et al.
Veröffentlicht: (2025)
von: Vidi Team, et al.
Veröffentlicht: (2025)
MetaphorVU: Towards Metaphorical Video Understanding
von: Li, Zhuoqun, et al.
Veröffentlicht: (2026)
von: Li, Zhuoqun, et al.
Veröffentlicht: (2026)
Fully Fine-tuned CLIP Models are Efficient Few-Shot Learners
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
von: Liu, Mushui, et al.
Veröffentlicht: (2024)
Edit-Your-Interest: Efficient Video Editing via Feature Most-Similar Propagation
von: Zuo, Yi, et al.
Veröffentlicht: (2025)
von: Zuo, Yi, et al.
Veröffentlicht: (2025)
BackdoorBench: A Comprehensive Benchmark and Analysis of Backdoor Learning
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Baoyuan, et al.
Veröffentlicht: (2024)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
CamDirector: Towards Long-Term Coherent Video Trajectory Editing
von: Shi, Zhihao, et al.
Veröffentlicht: (2026)
von: Shi, Zhihao, et al.
Veröffentlicht: (2026)
FATE: Full-head Gaussian Avatar with Textural Editing from Monocular Video
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
KFS-Bench: Comprehensive Evaluation of Key Frame Sampling in Long Video Understanding
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
FlexSelect: Flexible Token Selection for Efficient Long Video Understanding
von: Zhang, Yunzhu, et al.
Veröffentlicht: (2025)
von: Zhang, Yunzhu, et al.
Veröffentlicht: (2025)
MA-Bench: Towards Fine-grained Micro-Action Understanding
von: Li, Kun, et al.
Veröffentlicht: (2026)
von: Li, Kun, et al.
Veröffentlicht: (2026)
LvBench: A Benchmark for Long-form Video Understanding with Versatile Multi-modal Question Answering
von: Zhang, Hongjie, et al.
Veröffentlicht: (2023)
von: Zhang, Hongjie, et al.
Veröffentlicht: (2023)
UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RSVP: Reasoning Segmentation via Visual Prompting and Multi-modal Chain-of-Thought
von: Lu, Yi, et al.
Veröffentlicht: (2025) -
Zero-Shot Long-Form Video Understanding through Screenplay
von: Wu, Yongliang, et al.
Veröffentlicht: (2024) -
OpusAnimation: Code-Based Dynamic Chart Generation
von: Li, Bozheng, et al.
Veröffentlicht: (2025) -
Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
von: Wu, Yongliang, et al.
Veröffentlicht: (2024) -
Reframe Anything: LLM Agent for Open World Video Reframing
von: Cao, Jiawang, et al.
Veröffentlicht: (2024)