Annotation Techniques for Judo Combat Phase Classification from Tournament Footage
Fuente:
arXiv
Salvato in:
| Autori principali: | Miyaguchi, Anthony, Moutahir, Jed, Sutar, Tanmay |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
di: Wu, Xun, et al.
Pubblicazione: (2024)
di: Wu, Xun, et al.
Pubblicazione: (2024)
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
di: Kontostathis, Ioannis, et al.
Pubblicazione: (2024)
di: Kontostathis, Ioannis, et al.
Pubblicazione: (2024)
MagicAnime: A Hierarchically Annotated, Multimodal and Multitasking Dataset with Benchmarks for Cartoon Animation Generation
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
di: Xu, Shuolin, et al.
Pubblicazione: (2025)
A Hierarchical Compression Technique for 3D Gaussian Splatting Compression
di: Huang, He, et al.
Pubblicazione: (2024)
di: Huang, He, et al.
Pubblicazione: (2024)
LMM-Regularized CLIP Embeddings for Image Classification
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)
Local Neighborhood Features for 3D Classification
di: Sheshappanavar, Shivanand Venkanna, et al.
Pubblicazione: (2022)
di: Sheshappanavar, Shivanand Venkanna, et al.
Pubblicazione: (2022)
Deep Learning Classification of Photoplethysmogram Signal for Hypertension Levels
di: Nasir, Nida, et al.
Pubblicazione: (2024)
di: Nasir, Nida, et al.
Pubblicazione: (2024)
Deep Compositional Phase Diffusion for Long Motion Sequence Generation
di: Au, Ho Yin, et al.
Pubblicazione: (2025)
di: Au, Ho Yin, et al.
Pubblicazione: (2025)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
di: Zhang, Zhihao, et al.
Pubblicazione: (2023)
di: Zhang, Zhihao, et al.
Pubblicazione: (2023)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
di: Baia, Alina Elena, et al.
Pubblicazione: (2025)
di: Baia, Alina Elena, et al.
Pubblicazione: (2025)
TALDS-Net: Task-Aware Adaptive Local Descriptors Selection for Few-shot Image Classification
di: Qiao, Qian, et al.
Pubblicazione: (2023)
di: Qiao, Qian, et al.
Pubblicazione: (2023)
UAU-Net: Uncertainty-aware Representation Learning and Evidential Classification for Facial Action Unit Detection
di: Li, Yuze, et al.
Pubblicazione: (2026)
di: Li, Yuze, et al.
Pubblicazione: (2026)
AsyReC: A Multimodal Graph-based Framework for Spatio-Temporal Asymmetric Dyadic Relationship Classification
di: Tang, Wang, et al.
Pubblicazione: (2025)
di: Tang, Wang, et al.
Pubblicazione: (2025)
LookupForensics: A Large-Scale Multi-Task Dataset for Multi-Phase Image-Based Fact Verification
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
di: Cui, Shuhan, et al.
Pubblicazione: (2024)
Integrating Multi-Modal Sensors: A Review of Fusion Techniques for Intelligent Vehicles
di: Wei, Chuheng, et al.
Pubblicazione: (2025)
di: Wei, Chuheng, et al.
Pubblicazione: (2025)
Video Watermarking: Safeguarding Your Video from (Unauthorized) Annotations by Video-based LLMs
di: Li, Jinmin, et al.
Pubblicazione: (2024)
di: Li, Jinmin, et al.
Pubblicazione: (2024)
Quizzard@INOVA Challenge 2025 -- Track A: Plug-and-Play Technique in Interleaved Multi-Image Model
di: Cuong, Dinh Viet, et al.
Pubblicazione: (2025)
di: Cuong, Dinh Viet, et al.
Pubblicazione: (2025)
CoreMark: Toward Robust and Universal Text Watermarking Technique
di: Meng, Jiale, et al.
Pubblicazione: (2025)
di: Meng, Jiale, et al.
Pubblicazione: (2025)
Multimodal Engagement Analysis from Facial Videos in the Classroom
di: Sümer, Ömer, et al.
Pubblicazione: (2021)
di: Sümer, Ömer, et al.
Pubblicazione: (2021)
Feature CAM: Interpretable AI in Image Classification
di: Clement, Frincy, et al.
Pubblicazione: (2024)
di: Clement, Frincy, et al.
Pubblicazione: (2024)
Harnessing Self-Supervised Features for Art Classification
di: Melis, Federico, et al.
Pubblicazione: (2026)
di: Melis, Federico, et al.
Pubblicazione: (2026)
Generating Attribute-Aware Human Motions from Textual Prompt
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
Deepfake Detection: A Comprehensive Survey from the Reliability Perspective
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
di: Wang, Tianyi, et al.
Pubblicazione: (2022)
Latent Reconstruction from Generated Data for Multimodal Misinformation Detection
di: Papadopoulos, Stefanos-Iordanis, et al.
Pubblicazione: (2025)
di: Papadopoulos, Stefanos-Iordanis, et al.
Pubblicazione: (2025)
Learning from Silence and Noise for Visual Sound Source Localization
di: Juanola, Xavier, et al.
Pubblicazione: (2025)
di: Juanola, Xavier, et al.
Pubblicazione: (2025)
MultiColor: Image Colorization by Learning from Multiple Color Spaces
di: Du, Xiangcheng, et al.
Pubblicazione: (2024)
di: Du, Xiangcheng, et al.
Pubblicazione: (2024)
Gorgeous: Create Your Desired Character Facial Makeup from Any Ideas
di: Sii, Jia Wei, et al.
Pubblicazione: (2024)
di: Sii, Jia Wei, et al.
Pubblicazione: (2024)
CoPRS: Learning Positional Prior from Chain-of-Thought for Reasoning Segmentation
di: Lu, Zhenyu, et al.
Pubblicazione: (2025)
di: Lu, Zhenyu, et al.
Pubblicazione: (2025)
Visual question answering: from early developments to recent advances -- a survey
di: Huynh, Ngoc Dung, et al.
Pubblicazione: (2025)
di: Huynh, Ngoc Dung, et al.
Pubblicazione: (2025)
MSRS: Training Multimodal Speech Recognition Models from Scratch with Sparse Mask Optimization
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
Hand1000: Generating Realistic Hands from Text with Only 1,000 Images
di: Zhang, Haozhuo, et al.
Pubblicazione: (2024)
di: Zhang, Haozhuo, et al.
Pubblicazione: (2024)
ContextBLIP: Doubly Contextual Alignment for Contrastive Image Retrieval from Linguistically Complex Descriptions
di: Lin, Honglin, et al.
Pubblicazione: (2024)
di: Lin, Honglin, et al.
Pubblicazione: (2024)
MVPbev: Multi-view Perspective Image Generation from BEV with Test-time Controllability and Generalizability
di: Liu, Buyu, et al.
Pubblicazione: (2024)
di: Liu, Buyu, et al.
Pubblicazione: (2024)
M2ORT: Many-To-One Regression Transformer for Spatial Transcriptomics Prediction from Histopathology Images
di: Wang, Hongyi, et al.
Pubblicazione: (2024)
di: Wang, Hongyi, et al.
Pubblicazione: (2024)
MTFusion: Reconstructing Any 3D Object from Single Image Using Multi-word Textual Inversion
di: Liu, Yu, et al.
Pubblicazione: (2024)
di: Liu, Yu, et al.
Pubblicazione: (2024)
SynopGround: A Large-Scale Dataset for Multi-Paragraph Video Grounding from TV Dramas and Synopses
di: Tan, Chaolei, et al.
Pubblicazione: (2024)
di: Tan, Chaolei, et al.
Pubblicazione: (2024)
Robust Latent Representation Tuning for Image-text Classification
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
TimeNeRF: Building Generalizable Neural Radiance Fields across Time from Few-Shot Input Views
di: Hung, Hsiang-Hui, et al.
Pubblicazione: (2025)
di: Hung, Hsiang-Hui, et al.
Pubblicazione: (2025)
ByteNet: Rethinking Multimedia File Fragment Classification through Visual Perspectives
di: Liu, Wenyang, et al.
Pubblicazione: (2024)
di: Liu, Wenyang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
di: Wu, Xun, et al.
Pubblicazione: (2024) -
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
di: Kontostathis, Ioannis, et al.
Pubblicazione: (2024) -
MagicAnime: A Hierarchically Annotated, Multimodal and Multitasking Dataset with Benchmarks for Cartoon Animation Generation
di: Xu, Shuolin, et al.
Pubblicazione: (2025) -
A Hierarchical Compression Technique for 3D Gaussian Splatting Compression
di: Huang, He, et al.
Pubblicazione: (2024) -
LMM-Regularized CLIP Embeddings for Image Classification
di: Tzelepi, Maria, et al.
Pubblicazione: (2024)