BoxingVI: A Multi-Modal Benchmark for Boxing Action Recognition and Localization
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Rahul, Baghel, Vipul, Singh, Sudhanshu, Badatya, Bikash Kumar, Yadav, Shivam, Srinivasan, Babji, Hegde, Ravi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UTAL-GNN: Unsupervised Temporal Action Localization using Graph Neural Networks
by: Badatya, Bikash Kumar, et al.
Published: (2025)
by: Badatya, Bikash Kumar, et al.
Published: (2025)
Biomechanical-phase based Temporal Segmentation in Sports Videos: a Demonstration on Javelin-Throw
by: Badatya, Bikash Kumar, et al.
Published: (2025)
by: Badatya, Bikash Kumar, et al.
Published: (2025)
BoxMAC -- A Boxing Dataset for Multi-label Action Classification
by: Sahoo, Shashikanta
Published: (2024)
by: Sahoo, Shashikanta
Published: (2024)
BoxCell: Leveraging SAM for Cell Segmentation with Box Supervision
by: Tyagi, Aayush Kumar, et al.
Published: (2023)
by: Tyagi, Aayush Kumar, et al.
Published: (2023)
ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes
by: Kumar, Shivam
Published: (2026)
by: Kumar, Shivam
Published: (2026)
BoxComm: Benchmarking Category-Aware Commentary Generation and Narration Rhythm in Boxing
by: Wang, Kaiwen, et al.
Published: (2026)
by: Wang, Kaiwen, et al.
Published: (2026)
A Visually Attentive Splice Localization Network with Multi-Domain Feature Extractor and Multi-Receptive Field Upsampler
by: Yadav, Ankit, et al.
Published: (2024)
by: Yadav, Ankit, et al.
Published: (2024)
Local Features Meet Stochastic Anonymization: Revolutionizing Privacy-Preserving Face Recognition for Black-Box Models
by: Liu, Yuanwei, et al.
Published: (2024)
by: Liu, Yuanwei, et al.
Published: (2024)
TexTAR : Textual Attribute Recognition in Multi-domain and Multi-lingual Document Images
by: Kumar, Rohan, et al.
Published: (2025)
by: Kumar, Rohan, et al.
Published: (2025)
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
Task2Box: Box Embeddings for Modeling Asymmetric Task Relationships
by: Daroya, Rangel, et al.
Published: (2024)
by: Daroya, Rangel, et al.
Published: (2024)
BoxDreamer: Dreaming Box Corners for Generalizable Object Pose Estimation
by: Yu, Yuanhong, et al.
Published: (2025)
by: Yu, Yuanhong, et al.
Published: (2025)
ChatterBox: Multi-round Multimodal Referring and Grounding
by: Tian, Yunjie, et al.
Published: (2024)
by: Tian, Yunjie, et al.
Published: (2024)
Multi-Modality Co-Learning for Efficient Skeleton-based Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Box-QAymo: Box-Referring VQA Dataset for Autonomous Driving
by: Etchegaray, Djamahl, et al.
Published: (2025)
by: Etchegaray, Djamahl, et al.
Published: (2025)
Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion
by: Singh, Shivam, et al.
Published: (2026)
by: Singh, Shivam, et al.
Published: (2026)
Box for Mask and Mask for Box: weak losses for multi-task partially supervised learning
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
BoxSeg: Quality-Aware and Peer-Assisted Learning for Box-supervised Instance Segmentation
by: Lai, Jinxiang, et al.
Published: (2025)
by: Lai, Jinxiang, et al.
Published: (2025)
MonoBox: Tightness-free Box-supervised Polyp Segmentation using Monotonicity Constraint
by: Hu, Qiang, et al.
Published: (2024)
by: Hu, Qiang, et al.
Published: (2024)
BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion
by: Lan, Yuqing, et al.
Published: (2025)
by: Lan, Yuqing, et al.
Published: (2025)
QuadBox: Accelerating 3D Gaussian Splatting with Geometry-Aware Boxes
by: Li, Xinze, et al.
Published: (2026)
by: Li, Xinze, et al.
Published: (2026)
Box-Free Model Watermarks Are Prone to Black-Box Removal Attacks
by: An, Haonan, et al.
Published: (2024)
by: An, Haonan, et al.
Published: (2024)
Explore Human Parsing Modality for Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Box2Flow: Instance-based Action Flow Graphs from Videos
by: Li, Jiatong, et al.
Published: (2024)
by: Li, Jiatong, et al.
Published: (2024)
Deconvolution with a Box
by: Felzenszwalb, Pedro
Published: (2024)
by: Felzenszwalb, Pedro
Published: (2024)
Data-free Defense of Black Box Models Against Adversarial Attacks
by: Nayak, Gaurav Kumar, et al.
Published: (2022)
by: Nayak, Gaurav Kumar, et al.
Published: (2022)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
BoxSplitGen: A Generative Model for 3D Part Bounding Boxes in Varying Granularity
by: Koo, Juil, et al.
Published: (2026)
by: Koo, Juil, et al.
Published: (2026)
Multi-Grid Redundant Bounding Box Annotation for Accurate Object Detection
by: Tesema, Solomon Negussie, et al.
Published: (2022)
by: Tesema, Solomon Negussie, et al.
Published: (2022)
Scaling White-Box Transformers for Vision
by: Yang, Jinrui, et al.
Published: (2024)
by: Yang, Jinrui, et al.
Published: (2024)
BB-Patch: BlackBox Adversarial Patch-Attack using Zeroth-Order Optimization
by: Kumar, Satyadwyoom, et al.
Published: (2024)
by: Kumar, Satyadwyoom, et al.
Published: (2024)
Theoretical Analysis of Power-law Transformation on Images for Text Polarity Detection
by: Yadav, Narendra Singh, et al.
Published: (2025)
by: Yadav, Narendra Singh, et al.
Published: (2025)
Multi-Modal Character Localization and Extraction for Chinese Text Recognition
by: Li, Qilong, et al.
Published: (2026)
by: Li, Qilong, et al.
Published: (2026)
Threats to Arabic Handwriting Recognition: Investigating Black-Box Adversarial Attacks on embedded ConvNet models
by: Khayati, Mohsine EL, et al.
Published: (2026)
by: Khayati, Mohsine EL, et al.
Published: (2026)
SPACT18: Spiking Human Action Recognition Benchmark Dataset with Complementary RGB and Thermal Modalities
by: Ashraf, Yasser, et al.
Published: (2025)
by: Ashraf, Yasser, et al.
Published: (2025)
IBoxCLA: Towards Robust Box-supervised Segmentation of Polyp via Improved Box-dice and Contrastive Latent-anchors
by: Wang, Zhiwei, et al.
Published: (2023)
by: Wang, Zhiwei, et al.
Published: (2023)
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
by: Jeong, Uyoung, et al.
Published: (2023)
by: Jeong, Uyoung, et al.
Published: (2023)
MSTAR: Box-free Multi-query Scene Text Retrieval with Attention Recycling
by: Yin, Liang, et al.
Published: (2025)
by: Yin, Liang, et al.
Published: (2025)
Transformer-Driven Multimodal Fusion for Explainable Suspiciousness Estimation in Visual Surveillance
by: Yadav, Kuldeep Singh, et al.
Published: (2025)
by: Yadav, Kuldeep Singh, et al.
Published: (2025)
Similar Items
-
UTAL-GNN: Unsupervised Temporal Action Localization using Graph Neural Networks
by: Badatya, Bikash Kumar, et al.
Published: (2025) -
Biomechanical-phase based Temporal Segmentation in Sports Videos: a Demonstration on Javelin-Throw
by: Badatya, Bikash Kumar, et al.
Published: (2025) -
BoxMAC -- A Boxing Dataset for Multi-label Action Classification
by: Sahoo, Shashikanta
Published: (2024) -
BoxCell: Leveraging SAM for Cell Segmentation with Box Supervision
by: Tyagi, Aayush Kumar, et al.
Published: (2023) -
ShapeCodeBench: A Renewable Benchmark for Perception-to-Program Reconstruction of Synthetic Shape Scenes
by: Kumar, Shivam
Published: (2026)