Pixels or Positions? Benchmarking Modalities in Group Activity Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Karki, Drishya, Ramazanova, Merey, Cioppa, Anthony, Giancola, Silvio, Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation
by: Alhuwaider, Shyma, et al.
Published: (2026)
by: Alhuwaider, Shyma, et al.
Published: (2026)
Exploring Missing Modality in Multimodal Egocentric Datasets
by: Ramazanova, Merey, et al.
Published: (2024)
by: Ramazanova, Merey, et al.
Published: (2024)
Test-Time Adaptation for Combating Missing Modalities in Egocentric Videos
by: Ramazanova, Merey, et al.
Published: (2024)
by: Ramazanova, Merey, et al.
Published: (2024)
Deep learning for action spotting in association football videos
by: Giancola, Silvio, et al.
Published: (2024)
by: Giancola, Silvio, et al.
Published: (2024)
X-VARS: Introducing Explainability in Football Refereeing with Multi-Modal Large Language Model
by: Held, Jan, et al.
Published: (2024)
by: Held, Jan, et al.
Published: (2024)
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023)
by: Mkhallati, Hassan, et al.
Published: (2023)
Investigating Event-Based Cameras for Video Frame Interpolation in Sports
by: Deckyvere, Antoine, et al.
Published: (2024)
by: Deckyvere, Antoine, et al.
Published: (2024)
ADVMEM: Adversarial Memory Initialization for Realistic Test-Time Adaptation via Tracklet-Based Benchmarking
by: Alhuwaider, Shyma, et al.
Published: (2025)
by: Alhuwaider, Shyma, et al.
Published: (2025)
OpenTAD: A Unified Framework and Comprehensive Study of Temporal Action Detection
by: Liu, Shuming, et al.
Published: (2025)
by: Liu, Shuming, et al.
Published: (2025)
OSL-ActionSpotting: A Unified Library for Action Spotting in Sports Videos
by: Benzakour, Yassine, et al.
Published: (2024)
by: Benzakour, Yassine, et al.
Published: (2024)
VARS: Video Assistant Referee System for Automated Soccer Decision Making from Multiple Views
by: Held, Jan, et al.
Published: (2023)
by: Held, Jan, et al.
Published: (2023)
SoccerNet-Tracking: Multiple Object Tracking Dataset and Benchmark in Soccer Videos
by: Cioppa, Anthony, et al.
Published: (2022)
by: Cioppa, Anthony, et al.
Published: (2022)
Towards AI-Powered Video Assistant Referee System (VARS) for Association Football
by: Held, Jan, et al.
Published: (2024)
by: Held, Jan, et al.
Published: (2024)
Towards Active Learning for Action Spotting in Association Football Videos
by: Giancola, Silvio, et al.
Published: (2023)
by: Giancola, Silvio, et al.
Published: (2023)
Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders
by: Eymaël, Alexandre, et al.
Published: (2024)
by: Eymaël, Alexandre, et al.
Published: (2024)
Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2026)
by: Eldesokey, Abdelrahman, et al.
Published: (2026)
3D Convex Splatting: Radiance Field Rendering with 3D Smooth Convexes
by: Held, Jan, et al.
Published: (2024)
by: Held, Jan, et al.
Published: (2024)
Hybrid Structure-from-Motion and Camera Relocalization for Enhanced Egocentric Localization
by: Mai, Jinjie, et al.
Published: (2024)
by: Mai, Jinjie, et al.
Published: (2024)
Triangle Splatting for Real-Time Radiance Field Rendering
by: Held, Jan, et al.
Published: (2025)
by: Held, Jan, et al.
Published: (2025)
Learning Semantic Segmentation with Query Points Supervision on Aerial Images
by: Rivier, Santiago, et al.
Published: (2023)
by: Rivier, Santiago, et al.
Published: (2023)
MVTN: Learning Multi-View Transformations for 3D Understanding
by: Hamdi, Abdullah, et al.
Published: (2022)
by: Hamdi, Abdullah, et al.
Published: (2022)
Action Anticipation from SoccerNet Football Video Broadcasts
by: Dalal, Mohamad, et al.
Published: (2025)
by: Dalal, Mohamad, et al.
Published: (2025)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
by: Elsharkawi, Ismael, et al.
Published: (2026)
by: Elsharkawi, Ismael, et al.
Published: (2026)
Evaluation of Test-Time Adaptation Under Computational Time Constraints
by: Alfarra, Motasem, et al.
Published: (2023)
by: Alfarra, Motasem, et al.
Published: (2023)
TrackNeRF: Bundle Adjusting NeRF from Sparse and Noisy Views via Feature Tracks
by: Mai, Jinjie, et al.
Published: (2024)
by: Mai, Jinjie, et al.
Published: (2024)
Towards Athlete Fatigue Assessment from Association Football Videos
by: Bou, Xavier, et al.
Published: (2026)
by: Bou, Xavier, et al.
Published: (2026)
Online Distillation with Continual Learning for Cyclic Domain Shifts
by: Houyon, Joachim, et al.
Published: (2023)
by: Houyon, Joachim, et al.
Published: (2023)
Semi-Supervised Training to Improve Player and Ball Detection in Soccer
by: Vandeghen, Renaud, et al.
Published: (2022)
by: Vandeghen, Renaud, et al.
Published: (2022)
ActivityCLIP: Enhancing Group Activity Recognition by Mining Complementary Information from Text to Supplement Image Modality
by: Xu, Guoliang, et al.
Published: (2024)
by: Xu, Guoliang, et al.
Published: (2024)
SoccerNet Game State Reconstruction: End-to-End Athlete Tracking and Identification on a Minimap
by: Somers, Vladimir, et al.
Published: (2024)
by: Somers, Vladimir, et al.
Published: (2024)
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
LinDeps: A Fine-tuning Free Post-Pruning Method to Remove Layer-Wise Linear Dependencies with Guaranteed Performance Preservation
by: Henry, Maxim, et al.
Published: (2025)
by: Henry, Maxim, et al.
Published: (2025)
Video Self-Stitching Graph Network for Temporal Action Localization
by: Zhao, Chen, et al.
Published: (2020)
by: Zhao, Chen, et al.
Published: (2020)
Behind the Magic, MERLIM: Multi-modal Evaluation Benchmark for Large Image-Language Models
by: Villa, Andrés, et al.
Published: (2023)
by: Villa, Andrés, et al.
Published: (2023)
Unified Framework with Consistency across Modalities for Human Activity Recognition
by: Tran, Tuyen, et al.
Published: (2024)
by: Tran, Tuyen, et al.
Published: (2024)
PromptGAR: Flexible Promptive Group Activity Recognition
by: Jin, Zhangyu, et al.
Published: (2025)
by: Jin, Zhangyu, et al.
Published: (2025)
Multi-Stream Cellular Test-Time Adaptation of Real-Time Models Evolving in Dynamic Environments
by: Gérin, Benoît, et al.
Published: (2024)
by: Gérin, Benoît, et al.
Published: (2024)
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
BOLT: Boost Large Vision-Language Model Without Training for Long-form Video Understanding
by: Liu, Shuming, et al.
Published: (2025)
by: Liu, Shuming, et al.
Published: (2025)
$β$-CLIP: Text-Conditioned Contrastive Learning for Multi-Granular Vision-Language Alignment
by: Zohra, Fatimah, et al.
Published: (2025)
by: Zohra, Fatimah, et al.
Published: (2025)
Similar Items
-
GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation
by: Alhuwaider, Shyma, et al.
Published: (2026) -
Exploring Missing Modality in Multimodal Egocentric Datasets
by: Ramazanova, Merey, et al.
Published: (2024) -
Test-Time Adaptation for Combating Missing Modalities in Egocentric Videos
by: Ramazanova, Merey, et al.
Published: (2024) -
Deep learning for action spotting in association football videos
by: Giancola, Silvio, et al.
Published: (2024) -
X-VARS: Introducing Explainability in Football Refereeing with Multi-Modal Large Language Model
by: Held, Jan, et al.
Published: (2024)