TennisExpert: Towards Expert-Level Analytical Sports Video Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Liu, Zhaoyu, Weng, Xi, Hu, Lianyu, Hou, Zhe, Jiang, Kan, Dong, Jin Song, Liu, Yang |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
F$^3$Set: Towards Analyzing Fast, Frequent, and Fine-grained Events from Videos
par: Liu, Zhaoyu, et autres
Publié: (2025)
par: Liu, Zhaoyu, et autres
Publié: (2025)
Few-Shot Precise Event Spotting via Unified Multi-Entity Graph and Distillation
par: Liu, Zhaoyu, et autres
Publié: (2025)
par: Liu, Zhaoyu, et autres
Publié: (2025)
Enhancing Sports Strategy with Video Analytics and Data Mining: Automated Video-Based Analytics Framework for Tennis Doubles
par: Chen, Jia Wei
Publié: (2025)
par: Chen, Jia Wei
Publié: (2025)
MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
par: Zhao, Yilun, et autres
Publié: (2025)
par: Zhao, Yilun, et autres
Publié: (2025)
AesRM: Improving Video Aesthetics with Expert-Level Feedback
par: Han, Yujin, et autres
Publié: (2026)
par: Han, Yujin, et autres
Publié: (2026)
ShotBench: Expert-Level Cinematic Understanding in Vision-Language Models
par: Liu, Hongbo, et autres
Publié: (2025)
par: Liu, Hongbo, et autres
Publié: (2025)
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
par: Yang, Zuhao, et autres
Publié: (2025)
par: Yang, Zuhao, et autres
Publié: (2025)
VideoExpert: Augmented LLM for Temporal-Sensitive Video Understanding
par: Zhao, Henghao, et autres
Publié: (2025)
par: Zhao, Henghao, et autres
Publié: (2025)
PhyDAE: Physics-Guided Degradation-Adaptive Experts for All-in-One Remote Sensing Image Restoration
par: Dong, Zhe, et autres
Publié: (2025)
par: Dong, Zhe, et autres
Publié: (2025)
ExpertAF: Expert Actionable Feedback from Video
par: Ashutosh, Kumar, et autres
Publié: (2024)
par: Ashutosh, Kumar, et autres
Publié: (2024)
TennisTV: Do Multimodal Large Language Models Understand Tennis Rallies?
par: Bao, Zhongyuan, et autres
Publié: (2025)
par: Bao, Zhongyuan, et autres
Publié: (2025)
SARES-DEIM: Sparse Mixture-of-Experts Meets DETR for Robust SAR Ship Detection
par: Song, Fenghao, et autres
Publié: (2026)
par: Song, Fenghao, et autres
Publié: (2026)
TinyLLaVA-Video: Towards Smaller LMMs for Video Understanding with Group Resampler
par: Zhang, Xingjian, et autres
Publié: (2025)
par: Zhang, Xingjian, et autres
Publié: (2025)
Long-Tailed Distribution-Aware Router For Mixture-of-Experts in Large Vision-Language Model
par: Cai, Chaoxiang, et autres
Publié: (2025)
par: Cai, Chaoxiang, et autres
Publié: (2025)
Unified Multimodal Visual Tracking with Dual Mixture-of-Experts
par: Hong, Lingyi, et autres
Publié: (2026)
par: Hong, Lingyi, et autres
Publié: (2026)
HBridge: H-Shape Bridging of Heterogeneous Experts for Unified Multimodal Understanding and Generation
par: Wang, Xiang, et autres
Publié: (2025)
par: Wang, Xiang, et autres
Publié: (2025)
Watch and Learn: Leveraging Expert Knowledge and Language for Surgical Video Understanding
par: Gastager, David, et autres
Publié: (2025)
par: Gastager, David, et autres
Publié: (2025)
PathMMU: A Massive Multimodal Expert-Level Benchmark for Understanding and Reasoning in Pathology
par: Sun, Yuxuan, et autres
Publié: (2024)
par: Sun, Yuxuan, et autres
Publié: (2024)
ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding
par: Cao, Shuo, et autres
Publié: (2025)
par: Cao, Shuo, et autres
Publié: (2025)
Exposing and Defending the Achilles' Heel of Video Mixture-of-Experts
par: Wang, Songping, et autres
Publié: (2026)
par: Wang, Songping, et autres
Publié: (2026)
E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding
par: Liu, Ye, et autres
Publié: (2024)
par: Liu, Ye, et autres
Publié: (2024)
Gamma: Toward Generic Image Assessment with Mixture of Assessment Experts
par: Zhou, Hantao, et autres
Publié: (2025)
par: Zhou, Hantao, et autres
Publié: (2025)
CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
par: Yang, Zhuoyi, et autres
Publié: (2024)
par: Yang, Zhuoyi, et autres
Publié: (2024)
Towards Video Thinking Test: A Holistic Benchmark for Advanced Video Reasoning and Understanding
par: Zhang, Yuanhan, et autres
Publié: (2025)
par: Zhang, Yuanhan, et autres
Publié: (2025)
ExpertEdit: Learning Skill-Aware Motion Editing from Expert Videos
par: Somayazulu, Arjun, et autres
Publié: (2026)
par: Somayazulu, Arjun, et autres
Publié: (2026)
Pose-Guided Fine-Grained Sign Language Video Generation
par: Shi, Tongkai, et autres
Publié: (2024)
par: Shi, Tongkai, et autres
Publié: (2024)
Understanding Annotation Error Propagation and Learning an Adaptive Policy for Expert Intervention in Barrett's Video Segmentation
par: Rasanjalee, Lokesha, et autres
Publié: (2026)
par: Rasanjalee, Lokesha, et autres
Publié: (2026)
MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning
par: Gan, Ziliang, et autres
Publié: (2024)
par: Gan, Ziliang, et autres
Publié: (2024)
Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection
par: Lu, Yehao, et autres
Publié: (2025)
par: Lu, Yehao, et autres
Publié: (2025)
Towards Temporal Compositional Reasoning in Long-Form Sports Videos
par: Cao, Siyu, et autres
Publié: (2026)
par: Cao, Siyu, et autres
Publié: (2026)
HAMoBE: Hierarchical and Adaptive Mixture of Biometric Experts for Video-based Person ReID
par: Su, Yiyang, et autres
Publié: (2025)
par: Su, Yiyang, et autres
Publié: (2025)
Decoupled Video Generation with Chain of Training-free Diffusion Model Experts
par: Li, Wenhao, et autres
Publié: (2024)
par: Li, Wenhao, et autres
Publié: (2024)
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
par: Long, Chen, et autres
Publié: (2026)
par: Long, Chen, et autres
Publié: (2026)
Solving Token Gradient Conflict in Mixture-of-Experts for Large Vision-Language Model
par: Yang, Longrong, et autres
Publié: (2024)
par: Yang, Longrong, et autres
Publié: (2024)
Multi-Task Dense Prediction via Mixture of Low-Rank Experts
par: Yang, Yuqi, et autres
Publié: (2024)
par: Yang, Yuqi, et autres
Publié: (2024)
Dynamic Spatial-Temporal Aggregation for Skeleton-Aware Sign Language Recognition
par: Hu, Lianyu, et autres
Publié: (2024)
par: Hu, Lianyu, et autres
Publié: (2024)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
par: Zong, Zhuofan, et autres
Publié: (2024)
par: Zong, Zhuofan, et autres
Publié: (2024)
ExAct: A Video-Language Benchmark for Expert Action Analysis
par: Yi, Han, et autres
Publié: (2025)
par: Yi, Han, et autres
Publié: (2025)
Towards Adversarial Robustness of Model-Level Mixture-of-Experts Architectures for Semantic Segmentation
par: Pavlitska, Svetlana, et autres
Publié: (2024)
par: Pavlitska, Svetlana, et autres
Publié: (2024)
Dual-Expert Consistency Model for Efficient and High-Quality Video Generation
par: Lv, Zhengyao, et autres
Publié: (2025)
par: Lv, Zhengyao, et autres
Publié: (2025)
Documents similaires
-
F$^3$Set: Towards Analyzing Fast, Frequent, and Fine-grained Events from Videos
par: Liu, Zhaoyu, et autres
Publié: (2025) -
Few-Shot Precise Event Spotting via Unified Multi-Entity Graph and Distillation
par: Liu, Zhaoyu, et autres
Publié: (2025) -
Enhancing Sports Strategy with Video Analytics and Data Mining: Automated Video-Based Analytics Framework for Tennis Doubles
par: Chen, Jia Wei
Publié: (2025) -
MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
par: Zhao, Yilun, et autres
Publié: (2025) -
AesRM: Improving Video Aesthetics with Expert-Level Feedback
par: Han, Yujin, et autres
Publié: (2026)