LiViBench: An Omnimodal Benchmark for Interactive Livestream Video Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Xiaodong, Huang, Langling, Wu, Zhirong, Zhao, Xu, Xu, Teng, Xia, Xuhong, Peng, Peixi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs
di: Tao, Keda, et al.
Pubblicazione: (2026)
di: Tao, Keda, et al.
Pubblicazione: (2026)
ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
di: Wang, Xiaodong, et al.
Pubblicazione: (2025)
Active Perception Agent for Omnimodal Audio-Video Understanding
di: Tao, Keda, et al.
Pubblicazione: (2025)
di: Tao, Keda, et al.
Pubblicazione: (2025)
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
di: Zhuang, Cailin, et al.
Pubblicazione: (2025)
di: Zhuang, Cailin, et al.
Pubblicazione: (2025)
TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
di: Chen, Harold Haodong, et al.
Pubblicazione: (2025)
Q-Bench-Video: Benchmarking the Video Quality Understanding of LMMs
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
di: Chen, Guo, et al.
Pubblicazione: (2024)
di: Chen, Guo, et al.
Pubblicazione: (2024)
ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation
di: Liang, Yongyuan, et al.
Pubblicazione: (2025)
di: Liang, Yongyuan, et al.
Pubblicazione: (2025)
ViMU: Benchmarking Video Metaphorical Understanding
di: Li, Qi, et al.
Pubblicazione: (2026)
di: Li, Qi, et al.
Pubblicazione: (2026)
EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs
di: He, Yuping, et al.
Pubblicazione: (2025)
di: He, Yuping, et al.
Pubblicazione: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
InstructionBench: An Instructional Video Understanding Benchmark
di: Wei, Haiwan, et al.
Pubblicazione: (2025)
di: Wei, Haiwan, et al.
Pubblicazione: (2025)
Benchmarking Scientific Understanding and Reasoning for Video Generation using VideoScience-Bench
di: Hu, Lanxiang, et al.
Pubblicazione: (2025)
di: Hu, Lanxiang, et al.
Pubblicazione: (2025)
SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning
di: Kong, Fanqi, et al.
Pubblicazione: (2025)
di: Kong, Fanqi, et al.
Pubblicazione: (2025)
Prototype Embedding Optimization for Human-Object Interaction Detection in Livestreaming
di: Zhang, Menghui, et al.
Pubblicazione: (2025)
di: Zhang, Menghui, et al.
Pubblicazione: (2025)
EgoSocial: Benchmarking Proactive Intervention Ability of Omnimodal LLMs via Egocentric Social Interaction Perception
di: Wang, Xijun, et al.
Pubblicazione: (2025)
di: Wang, Xijun, et al.
Pubblicazione: (2025)
FreeGen: Feed-Forward Reconstruction-Generation Co-Training for Free-Viewpoint Driving Scene Synthesis
di: Chen, Shijie, et al.
Pubblicazione: (2025)
di: Chen, Shijie, et al.
Pubblicazione: (2025)
CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
di: Wang, Xinran, et al.
Pubblicazione: (2025)
di: Wang, Xinran, et al.
Pubblicazione: (2025)
ConViS-Bench: Estimating Video Similarity Through Semantic Concepts
di: Liberatori, Benedetta, et al.
Pubblicazione: (2025)
di: Liberatori, Benedetta, et al.
Pubblicazione: (2025)
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
di: Hong, Wenyi, et al.
Pubblicazione: (2025)
di: Hong, Wenyi, et al.
Pubblicazione: (2025)
FAVOR-Bench: A Comprehensive Benchmark for Fine-Grained Video Motion Understanding
di: Tu, Chongjun, et al.
Pubblicazione: (2025)
di: Tu, Chongjun, et al.
Pubblicazione: (2025)
TextVidBench: A Benchmark for Long Video Scene Text Understanding
di: Zhong, Yangyang, et al.
Pubblicazione: (2025)
di: Zhong, Yangyang, et al.
Pubblicazione: (2025)
X-LeBench: A Benchmark for Extremely Long Egocentric Video Understanding
di: Zhou, Wenqi, et al.
Pubblicazione: (2025)
di: Zhou, Wenqi, et al.
Pubblicazione: (2025)
VideoRewardBench: Comprehensive Evaluation of Multimodal Reward Models for Video Understanding
di: Zhang, Zhihong, et al.
Pubblicazione: (2025)
di: Zhang, Zhihong, et al.
Pubblicazione: (2025)
LvBench: A Benchmark for Long-form Video Understanding with Versatile Multi-modal Question Answering
di: Zhang, Hongjie, et al.
Pubblicazione: (2023)
di: Zhang, Hongjie, et al.
Pubblicazione: (2023)
WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs
di: Hong, Jack, et al.
Pubblicazione: (2025)
di: Hong, Jack, et al.
Pubblicazione: (2025)
Video-SafetyBench: A Benchmark for Safety Evaluation of Video LVLMs
di: Liu, Xuannan, et al.
Pubblicazione: (2025)
di: Liu, Xuannan, et al.
Pubblicazione: (2025)
ViLCo-Bench: VIdeo Language COntinual learning Benchmark
di: Tang, Tianqi, et al.
Pubblicazione: (2024)
di: Tang, Tianqi, et al.
Pubblicazione: (2024)
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation
di: Zhang, Shi-Xue, et al.
Pubblicazione: (2025)
di: Zhang, Shi-Xue, et al.
Pubblicazione: (2025)
ViC-Bench: Benchmarking Visual-Interleaved Chain-of-Thought Capability in MLLMs with Free-Style Intermediate State Representations
di: Wu, Xuecheng, et al.
Pubblicazione: (2025)
di: Wu, Xuecheng, et al.
Pubblicazione: (2025)
LongViTU: Instruction Tuning for Long-Form Video Understanding
di: Wu, Rujie, et al.
Pubblicazione: (2025)
di: Wu, Rujie, et al.
Pubblicazione: (2025)
MT-Video-Bench: A Holistic Video Understanding Benchmark for Evaluating Multimodal LLMs in Multi-Turn Dialogues
di: Pan, Yaning, et al.
Pubblicazione: (2025)
di: Pan, Yaning, et al.
Pubblicazione: (2025)
ExpVid: A Benchmark for Experiment Video Understanding & Reasoning
di: Xu, Yicheng, et al.
Pubblicazione: (2025)
di: Xu, Yicheng, et al.
Pubblicazione: (2025)
U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking
di: Tang, Fenghe, et al.
Pubblicazione: (2025)
di: Tang, Fenghe, et al.
Pubblicazione: (2025)
LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding
di: Wu, Haoning, et al.
Pubblicazione: (2024)
di: Wu, Haoning, et al.
Pubblicazione: (2024)
Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration
di: Zhong, Hao, et al.
Pubblicazione: (2025)
di: Zhong, Hao, et al.
Pubblicazione: (2025)
PhoStream: Benchmarking Real-World Streaming for Omnimodal Assistants in Mobile Scenarios
di: Lu, Xudong, et al.
Pubblicazione: (2026)
di: Lu, Xudong, et al.
Pubblicazione: (2026)
Scaling Language-Centric Omnimodal Representation Learning
di: Xiao, Chenghao, et al.
Pubblicazione: (2025)
di: Xiao, Chenghao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model
di: Wang, Xiaodong, et al.
Pubblicazione: (2025) -
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs
di: Tao, Keda, et al.
Pubblicazione: (2026) -
ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos
di: Wang, Xiaodong, et al.
Pubblicazione: (2025) -
LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs
di: Wang, Xiaodong, et al.
Pubblicazione: (2025) -
Active Perception Agent for Omnimodal Audio-Video Understanding
di: Tao, Keda, et al.
Pubblicazione: (2025)