ViSTec: Video Modeling for Sports Technique Recognition and Tactical Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Yuchen, Yuan, Zeqing, Wu, Yihong, Cheng, Liqi, Deng, Dazhen, Wu, Yingcai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
von: He, Yuchen, et al.
Veröffentlicht: (2025)
von: He, Yuchen, et al.
Veröffentlicht: (2025)
CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation
von: Deng, Dazhen, et al.
Veröffentlicht: (2025)
von: Deng, Dazhen, et al.
Veröffentlicht: (2025)
FACTS: Fine-Grained Action Classification for Tactical Sports
von: Lai, Christopher, et al.
Veröffentlicht: (2024)
von: Lai, Christopher, et al.
Veröffentlicht: (2024)
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
ViBe: Ultra-High-Resolution Video Synthesis Born from Pure Images
von: Wu, Yunfeng, et al.
Veröffentlicht: (2026)
von: Wu, Yunfeng, et al.
Veröffentlicht: (2026)
Tracking-Aware Deformation Field Estimation for Non-rigid 3D Reconstruction in Robotic Surgeries
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
SportSkills: Physical Skill Learning from Sports Instructional Videos
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2026)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2026)
ViDiC: Video Difference Captioning
von: Wu, Jiangtao, et al.
Veröffentlicht: (2025)
von: Wu, Jiangtao, et al.
Veröffentlicht: (2025)
PRevivor: Reviving Ancient Chinese Paintings using Prior-Guided Color Transformers
von: Tang, Tan, et al.
Veröffentlicht: (2025)
von: Tang, Tan, et al.
Veröffentlicht: (2025)
Shot2Tactic-Caption: Multi-Scale Captioning of Badminton Videos for Tactical Understanding
von: Ding, Ning, et al.
Veröffentlicht: (2025)
von: Ding, Ning, et al.
Veröffentlicht: (2025)
A General Framework for Jersey Number Recognition in Sports Video
von: Koshkina, Maria, et al.
Veröffentlicht: (2024)
von: Koshkina, Maria, et al.
Veröffentlicht: (2024)
TiViBench: Benchmarking Think-in-Video Reasoning for Video Generative Models
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
ViSS-R1: Self-Supervised Reinforcement Video Reasoning
von: Fang, Bo, et al.
Veröffentlicht: (2025)
von: Fang, Bo, et al.
Veröffentlicht: (2025)
ViViD: Video Virtual Try-on using Diffusion Models
von: Fang, Zixun, et al.
Veröffentlicht: (2024)
von: Fang, Zixun, et al.
Veröffentlicht: (2024)
SeViCES: Unifying Semantic-Visual Evidence Consensus for Long Video Understanding
von: Sheng, Yuan, et al.
Veröffentlicht: (2025)
von: Sheng, Yuan, et al.
Veröffentlicht: (2025)
ViLA: Efficient Video-Language Alignment for Video Question Answering
von: Wang, Xijun, et al.
Veröffentlicht: (2023)
von: Wang, Xijun, et al.
Veröffentlicht: (2023)
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
The 1st Winner for 5th PVUW MeViS-Text Challenge: Strong MLLMs Meet SAM3 for Referring Video Object Segmentation
von: He, Xusheng, et al.
Veröffentlicht: (2026)
von: He, Xusheng, et al.
Veröffentlicht: (2026)
Sports-QA: A Large-Scale Video Question Answering Benchmark for Complex and Professional Sports
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
SGA-INTERACT: A 3D Skeleton-based Benchmark for Group Activity Understanding in Modern Basketball Tactic
von: Yang, Yuchen, et al.
Veröffentlicht: (2025)
von: Yang, Yuchen, et al.
Veröffentlicht: (2025)
AdaViP: Aligning Multi-modal LLMs via Adaptive Vision-enhanced Preference Optimization
von: Lu, Jinda, et al.
Veröffentlicht: (2025)
von: Lu, Jinda, et al.
Veröffentlicht: (2025)
LoViT: Long Video Transformer for Surgical Phase Recognition
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
BST: Badminton Stroke-type Transformer for Skeleton-based Action Recognition in Racket Sports
von: Chang, Jing-Yuan
Veröffentlicht: (2025)
von: Chang, Jing-Yuan
Veröffentlicht: (2025)
LiViBench: An Omnimodal Benchmark for Interactive Livestream Video Understanding
von: Wang, Xiaodong, et al.
Veröffentlicht: (2026)
von: Wang, Xiaodong, et al.
Veröffentlicht: (2026)
LongDPM: Overlap-Aware 4D Reconstruction from Long Monocular Videos
von: Xu, Chenyi, et al.
Veröffentlicht: (2026)
von: Xu, Chenyi, et al.
Veröffentlicht: (2026)
LoViC: Efficient Long Video Generation with Context Compression
von: Jiang, Jiaxiu, et al.
Veröffentlicht: (2025)
von: Jiang, Jiaxiu, et al.
Veröffentlicht: (2025)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
ViSpeak: Visual Instruction Feedback in Streaming Videos
von: Fu, Shenghao, et al.
Veröffentlicht: (2025)
von: Fu, Shenghao, et al.
Veröffentlicht: (2025)
SV3.3B: A Sports Video Understanding Model for Action Recognition
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
ViLLa: Video Reasoning Segmentation with Large Language Model
von: Zheng, Rongkun, et al.
Veröffentlicht: (2024)
von: Zheng, Rongkun, et al.
Veröffentlicht: (2024)
Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding
von: Jin, Peng, et al.
Veröffentlicht: (2023)
von: Jin, Peng, et al.
Veröffentlicht: (2023)
GenRec: Unifying Video Generation and Recognition with Diffusion Models
von: Weng, Zejia, et al.
Veröffentlicht: (2024)
von: Weng, Zejia, et al.
Veröffentlicht: (2024)
EmbodiedPlace: Learning Mixture-of-Features with Embodied Constraints for Visual Place Recognition
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
Round Outcome Prediction in VALORANT Using Tactical Features from Video Analysis
von: Hayakawa, Nirai, et al.
Veröffentlicht: (2025)
von: Hayakawa, Nirai, et al.
Veröffentlicht: (2025)
AutoEval-Video: An Automatic Benchmark for Assessing Large Vision Language Models in Open-Ended Video Question Answering
von: Chen, Xiuyuan, et al.
Veröffentlicht: (2023)
von: Chen, Xiuyuan, et al.
Veröffentlicht: (2023)
Poze: Sports Technique Feedback under Data Constraints
von: Singh, Agamdeep, et al.
Veröffentlicht: (2024)
von: Singh, Agamdeep, et al.
Veröffentlicht: (2024)
StegaVAR: Privacy-Preserving Video Action Recognition via Steganographic Domain Analysis
von: Chen, Lixin, et al.
Veröffentlicht: (2025)
von: Chen, Lixin, et al.
Veröffentlicht: (2025)
VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
von: Wang, Zeqing, et al.
Veröffentlicht: (2025)
HuViDPO:Enhancing Video Generation through Direct Preference Optimization for Human-Centric Alignment
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
von: Jiang, Lifan, et al.
Veröffentlicht: (2025)
Beyond Boundary Frames: Context-Centric Video Interpolation with Audio-Visual Semantics
von: Deng, Yuchen, et al.
Veröffentlicht: (2025)
von: Deng, Yuchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
von: He, Yuchen, et al.
Veröffentlicht: (2025) -
CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generation
von: Deng, Dazhen, et al.
Veröffentlicht: (2025) -
FACTS: Fine-Grained Action Classification for Tactical Sports
von: Lai, Christopher, et al.
Veröffentlicht: (2024) -
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
von: Wu, Tao, et al.
Veröffentlicht: (2024) -
ViBe: Ultra-High-Resolution Video Synthesis Born from Pure Images
von: Wu, Yunfeng, et al.
Veröffentlicht: (2026)