Poze: Sports Technique Feedback under Data Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Singh, Agamdeep, PB, Sujit, Vatsa, Mayank |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimizing Skin Lesion Classification via Multimodal Data and Auxiliary Task Integration
by: Khurshid, Mahapara, et al.
Published: (2024)
by: Khurshid, Mahapara, et al.
Published: (2024)
Right Looks, Wrong Reasons: Compositional Fidelity in Text-to-Image Generation
by: Vatsa, Mayank, et al.
Published: (2025)
by: Vatsa, Mayank, et al.
Published: (2025)
Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
by: Singh, Jaisidh, et al.
Published: (2024)
by: Singh, Jaisidh, et al.
Published: (2024)
NutriScreener: Retrieval-Augmented Multi-Pose Graph Attention Network for Malnourishment Screening
by: Khan, Misaal, et al.
Published: (2025)
by: Khan, Misaal, et al.
Published: (2025)
HyperSpaceX: Radial and Angular Exploration of HyperSpherical Dimensions
by: Chiranjeev, Chiranjeev, et al.
Published: (2024)
by: Chiranjeev, Chiranjeev, et al.
Published: (2024)
Unbiased Model Prediction Without Using Protected Attribute Information
by: Majumdar, Puspita, et al.
Published: (2026)
by: Majumdar, Puspita, et al.
Published: (2026)
Harmonizing Geometry and Uncertainty: Diffusion with Hyperspheres
by: Dosi, Muskan, et al.
Published: (2025)
by: Dosi, Muskan, et al.
Published: (2025)
Continual Unlearning for Foundational Text-to-Image Models without Generalization Erosion
by: Thakral, Kartik, et al.
Published: (2025)
by: Thakral, Kartik, et al.
Published: (2025)
Fine-Grained Erasure in Text-to-Image Diffusion-based Foundation Models
by: Thakral, Kartik, et al.
Published: (2025)
by: Thakral, Kartik, et al.
Published: (2025)
LitMAS: A Lightweight and Generalized Multi-Modal Anti-Spoofing Framework for Biometric Security
by: Gorthi, Nidheesh, et al.
Published: (2025)
by: Gorthi, Nidheesh, et al.
Published: (2025)
TAIGen: Training-Free Adversarial Image Generation via Diffusion Models
by: Roy, Susim, et al.
Published: (2025)
by: Roy, Susim, et al.
Published: (2025)
AnyTraverse: An off-road traversability framework with VLM and human operator in the loop
by: Sahu, Sattwik, et al.
Published: (2025)
by: Sahu, Sattwik, et al.
Published: (2025)
Off-Road LiDAR Intensity Based Semantic Segmentation
by: Viswanath, Kasi, et al.
Published: (2024)
by: Viswanath, Kasi, et al.
Published: (2024)
Discerning the Chaos: Detecting Adversarial Perturbations while Disentangling Intentional from Unintentional Noises
by: Jain, Anubhooti, et al.
Published: (2024)
by: Jain, Anubhooti, et al.
Published: (2024)
Navigating Text-to-Image Generative Bias across Indic Languages
by: Mittal, Surbhi, et al.
Published: (2024)
by: Mittal, Surbhi, et al.
Published: (2024)
Low-Resolution Chest X-ray Classification via Knowledge Distillation and Multi-task Learning
by: Akhter, Yasmeena, et al.
Published: (2024)
by: Akhter, Yasmeena, et al.
Published: (2024)
On Responsible Machine Learning Datasets with Fairness, Privacy, and Regulatory Norms
by: Mittal, Surbhi, et al.
Published: (2023)
by: Mittal, Surbhi, et al.
Published: (2023)
ViSTec: Video Modeling for Sports Technique Recognition and Tactical Analysis
by: He, Yuchen, et al.
Published: (2024)
by: He, Yuchen, et al.
Published: (2024)
Generalizing Sports Feedback Generation by Watching Competitions and Reading Books: A Rock Climbing Case Study
by: Rai, Arushi, et al.
Published: (2026)
by: Rai, Arushi, et al.
Published: (2026)
RetailKLIP : Finetuning OpenCLIP backbone using metric learning on a single GPU for Zero-shot retail product image classification
by: Srivastava, Muktabh Mayank
Published: (2023)
by: Srivastava, Muktabh Mayank
Published: (2023)
SportSkills: Physical Skill Learning from Sports Instructional Videos
by: Ashutosh, Kumar, et al.
Published: (2026)
by: Ashutosh, Kumar, et al.
Published: (2026)
Efficient Human Pose Estimation: Leveraging Advanced Techniques with MediaPipe
by: Sengar, Sandeep Singh, et al.
Published: (2024)
by: Sengar, Sandeep Singh, et al.
Published: (2024)
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
by: Wu, Tao, et al.
Published: (2024)
by: Wu, Tao, et al.
Published: (2024)
SportR: A Benchmark for Multimodal Large Language Model Reasoning in Sports
by: Xia, Haotian, et al.
Published: (2025)
by: Xia, Haotian, et al.
Published: (2025)
Sports-Traj: A Unified Trajectory Generation Model for Multi-Agent Movement in Sports
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Sports Re-ID: Improving Re-Identification Of Players In Broadcast Videos Of Team Sports
by: Comandur, Bharath
Published: (2022)
by: Comandur, Bharath
Published: (2022)
Women Sport Actions Dataset for Visual Classification Using Small Scale Training Data
by: Ray, Palash, et al.
Published: (2025)
by: Ray, Palash, et al.
Published: (2025)
NaViL: Rethinking Scaling Properties of Native Multimodal Large Language Models under Data Constraints
by: Tian, Changyao, et al.
Published: (2025)
by: Tian, Changyao, et al.
Published: (2025)
RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints
by: Pham, Tan-Hanh, et al.
Published: (2025)
by: Pham, Tan-Hanh, et al.
Published: (2025)
Confident Learning for Object Detection under Model Constraints
by: Yu, Yingda, et al.
Published: (2026)
by: Yu, Yingda, et al.
Published: (2026)
Sports-QA: A Large-Scale Video Question Answering Benchmark for Complex and Professional Sports
by: Li, Haopeng, et al.
Published: (2024)
by: Li, Haopeng, et al.
Published: (2024)
AdaSports-Traj: Role- and Domain-Aware Adaptation for Multi-Agent Trajectory Modeling in Sports
by: Xu, Yi, et al.
Published: (2025)
by: Xu, Yi, et al.
Published: (2025)
OFFSEG: A Semantic Segmentation Framework For Off-Road Driving
by: Viswanath, Kasi, et al.
Published: (2021)
by: Viswanath, Kasi, et al.
Published: (2021)
SportMamba: Adaptive Non-Linear Multi-Object Tracking with State Space Models for Team Sports
by: Khanna, Dheeraj, et al.
Published: (2025)
by: Khanna, Dheeraj, et al.
Published: (2025)
Scale-Aware Recognition in Satellite Images under Resource Constraints
by: Revankar, Shreelekha, et al.
Published: (2024)
by: Revankar, Shreelekha, et al.
Published: (2024)
Data Processing Techniques for Modern Multimodal Models
by: Li, Yinheng, et al.
Published: (2024)
by: Li, Yinheng, et al.
Published: (2024)
Action Valuation in Sports: A Survey
by: Xarles, Artur, et al.
Published: (2025)
by: Xarles, Artur, et al.
Published: (2025)
Landslide Hazard Mapping with Geospatial Foundation Models: Geographical Generalizability, Data Scarcity, and Band Adaptability
by: Li, Wenwen, et al.
Published: (2025)
by: Li, Wenwen, et al.
Published: (2025)
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation
by: Shah, Arya, et al.
Published: (2026)
by: Shah, Arya, et al.
Published: (2026)
Multi-Modal Sensor Fusion using Hybrid Attention for Autonomous Driving
by: Mayank, Mayank, et al.
Published: (2026)
by: Mayank, Mayank, et al.
Published: (2026)
Similar Items
-
Optimizing Skin Lesion Classification via Multimodal Data and Auxiliary Task Integration
by: Khurshid, Mahapara, et al.
Published: (2024) -
Right Looks, Wrong Reasons: Compositional Fidelity in Text-to-Image Generation
by: Vatsa, Mayank, et al.
Published: (2025) -
Learn "No" to Say "Yes" Better: Improving Vision-Language Models via Negations
by: Singh, Jaisidh, et al.
Published: (2024) -
NutriScreener: Retrieval-Augmented Multi-Pose Graph Attention Network for Malnourishment Screening
by: Khan, Misaal, et al.
Published: (2025) -
HyperSpaceX: Radial and Angular Exploration of HyperSpherical Dimensions
by: Chiranjeev, Chiranjeev, et al.
Published: (2024)