Beyond Pixels: Leveraging the Language of Soccer to Improve Spatio-Temporal Action Detection in Broadcast Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Ochin, Jeremie, Chekroun, Raphael, Stanciulescu, Bogdan, Manitsaris, Sotiris |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FOOTPASS: A Multi-Modal Multi-Agent Tactical Context Dataset for Play-by-Play Action Spotting in Soccer Broadcast Videos
by: Ochin, Jeremie, et al.
Published: (2025)
by: Ochin, Jeremie, et al.
Published: (2025)
Game State and Spatio-temporal Action Detection in Soccer using Graph Neural Networks and 3D Convolutional Networks
by: Ochin, Jeremie, et al.
Published: (2025)
by: Ochin, Jeremie, et al.
Published: (2025)
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023)
by: Mkhallati, Hassan, et al.
Published: (2023)
Action Anticipation from SoccerNet Football Video Broadcasts
by: Dalal, Mohamad, et al.
Published: (2025)
by: Dalal, Mohamad, et al.
Published: (2025)
BroadTrack: Broadcast Camera Tracking for Soccer
by: Magera, Floriane, et al.
Published: (2024)
by: Magera, Floriane, et al.
Published: (2024)
Open-Vocabulary Spatio-Temporal Action Detection
by: Wu, Tao, et al.
Published: (2024)
by: Wu, Tao, et al.
Published: (2024)
ASTRA: An Action Spotting TRAnsformer for Soccer Videos
by: Xarles, Artur, et al.
Published: (2024)
by: Xarles, Artur, et al.
Published: (2024)
STF: Spatio-Temporal Fusion Module for Improving Video Object Detection
by: Anwar, Noreen, et al.
Published: (2024)
by: Anwar, Noreen, et al.
Published: (2024)
Leveraging Temporal Contextualization for Video Action Recognition
by: Kim, Minji, et al.
Published: (2024)
by: Kim, Minji, et al.
Published: (2024)
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
Survey of Action Recognition, Spotting and Spatio-Temporal Localization in Soccer -- Current Trends and Research Perspectives
by: Seweryn, Karolina, et al.
Published: (2023)
by: Seweryn, Karolina, et al.
Published: (2023)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
by: Elsharkawi, Ismael, et al.
Published: (2026)
by: Elsharkawi, Ismael, et al.
Published: (2026)
Efficient Spatio-Temporal Vegetation Pixel Classification with Vision Transformers
by: Gomes, Alan, et al.
Published: (2026)
by: Gomes, Alan, et al.
Published: (2026)
Beyond Spatial Frequency: Pixel-wise Temporal Frequency-based Deepfake Video Detection
by: Kim, Taehoon, et al.
Published: (2025)
by: Kim, Taehoon, et al.
Published: (2025)
Improving Weakly-supervised Video Instance Segmentation by Leveraging Spatio-temporal Consistency
by: Arefi, Farnoosh, et al.
Published: (2024)
by: Arefi, Farnoosh, et al.
Published: (2024)
Patch Spatio-Temporal Relation Prediction for Video Anomaly Detection
by: Shen, Hao, et al.
Published: (2024)
by: Shen, Hao, et al.
Published: (2024)
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues
by: Girmaji, Rohit, et al.
Published: (2025)
by: Girmaji, Rohit, et al.
Published: (2025)
Video-Language Alignment via Spatio-Temporal Graph Transformer
by: Zhang, Shi-Xue, et al.
Published: (2024)
by: Zhang, Shi-Xue, et al.
Published: (2024)
PiTe: Pixel-Temporal Alignment for Large Video-Language Model
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Vulnerability-Aware Spatio-Temporal Learning for Generalizable Deepfake Video Detection
by: Nguyen, Dat, et al.
Published: (2025)
by: Nguyen, Dat, et al.
Published: (2025)
One-Stage Open-Vocabulary Temporal Action Detection Leveraging Temporal Multi-scale and Action Label Features
by: Nguyen, Trung Thanh, et al.
Published: (2024)
by: Nguyen, Trung Thanh, et al.
Published: (2024)
SoccerSynth-Detection: A Synthetic Dataset for Soccer Player Detection
by: Qin, Haobin, et al.
Published: (2025)
by: Qin, Haobin, et al.
Published: (2025)
Spatio-Temporal Context Prompting for Zero-Shot Action Detection
by: Huang, Wei-Jhe, et al.
Published: (2024)
by: Huang, Wei-Jhe, et al.
Published: (2024)
Cluster-Wise Spatio-Temporal Masking for Efficient Video-Language Pretraining
by: Zhuang, Weijun, et al.
Published: (2026)
by: Zhuang, Weijun, et al.
Published: (2026)
Semi-Supervised Training to Improve Player and Ball Detection in Soccer
by: Vandeghen, Renaud, et al.
Published: (2022)
by: Vandeghen, Renaud, et al.
Published: (2022)
Modelling Spatio-Temporal Interactions For Compositional Action Recognition
by: Rajendiran, Ramanathan, et al.
Published: (2023)
by: Rajendiran, Ramanathan, et al.
Published: (2023)
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
by: Kim, Yehna, et al.
Published: (2025)
by: Kim, Yehna, et al.
Published: (2025)
PixelRefer: A Unified Framework for Spatio-Temporal Object Referring with Arbitrary Granularity
by: Yuan, Yuqian, et al.
Published: (2025)
by: Yuan, Yuqian, et al.
Published: (2025)
Leveraging Consistent Spatio-Temporal Correspondence for Robust Visual Odometry
by: Zhang, Zhaoxing, et al.
Published: (2024)
by: Zhang, Zhaoxing, et al.
Published: (2024)
DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition
by: Ullah, Hayat, et al.
Published: (2025)
by: Ullah, Hayat, et al.
Published: (2025)
Improving Viewpoint-Invariance and Temporal Consistency for Action Detection
by: Porto, Yannick, et al.
Published: (2026)
by: Porto, Yannick, et al.
Published: (2026)
SoccerNet-Tracking: Multiple Object Tracking Dataset and Benchmark in Soccer Videos
by: Cioppa, Anthony, et al.
Published: (2022)
by: Cioppa, Anthony, et al.
Published: (2022)
Context-Guided Spatio-Temporal Video Grounding
by: Gu, Xin, et al.
Published: (2024)
by: Gu, Xin, et al.
Published: (2024)
Know-Show: Benchmarking Video-Language Models on Spatio-Temporal Grounded Reasoning
by: Sugandhika, Chinthani, et al.
Published: (2025)
by: Sugandhika, Chinthani, et al.
Published: (2025)
Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding
by: Gao, Shida, et al.
Published: (2025)
by: Gao, Shida, et al.
Published: (2025)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
SparseLaneSTP: Leveraging Spatio-Temporal Priors with Sparse Transformers for 3D Lane Detection
by: Pittner, Maximilian, et al.
Published: (2026)
by: Pittner, Maximilian, et al.
Published: (2026)
Towards Universal Soccer Video Understanding
by: Rao, Jiayuan, et al.
Published: (2024)
by: Rao, Jiayuan, et al.
Published: (2024)
Deep Understanding of Soccer Match Videos
by: Xu, Shikun, et al.
Published: (2024)
by: Xu, Shikun, et al.
Published: (2024)
Similar Items
-
FOOTPASS: A Multi-Modal Multi-Agent Tactical Context Dataset for Play-by-Play Action Spotting in Soccer Broadcast Videos
by: Ochin, Jeremie, et al.
Published: (2025) -
Game State and Spatio-temporal Action Detection in Soccer using Graph Neural Networks and 3D Convolutional Networks
by: Ochin, Jeremie, et al.
Published: (2025) -
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023) -
Action Anticipation from SoccerNet Football Video Broadcasts
by: Dalal, Mohamad, et al.
Published: (2025) -
BroadTrack: Broadcast Camera Tracking for Soccer
by: Magera, Floriane, et al.
Published: (2024)