Generalizing Sports Feedback Generation by Watching Competitions and Reading Books: A Rock Climbing Case Study
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rai, Arushi, Kovashka, Adriana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Consistent Temporal Grounding between Related Tasks in Sports Coaching
von: Rai, Arushi, et al.
Veröffentlicht: (2026)
von: Rai, Arushi, et al.
Veröffentlicht: (2026)
VEIL: Vetting Extracted Image Labels from In-the-Wild Captions for Weakly-Supervised Object Detection
von: Rai, Arushi, et al.
Veröffentlicht: (2023)
von: Rai, Arushi, et al.
Veröffentlicht: (2023)
The Face of Persuasion: Analyzing Bias and Generating Culture-Aware Ads
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025)
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025)
CAP: Evaluation of Persuasive and Creative Image Generation
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2024)
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2024)
Towards Generalization of Tactile Image Generation: Reference-Free Evaluation in a Leakage-Free Setting
von: Gungor, Cagri, et al.
Veröffentlicht: (2025)
von: Gungor, Cagri, et al.
Veröffentlicht: (2025)
Role Bias in Diffusion Models: Diagnosing and Mitigating through Intermediate Decomposition
von: Malakouti, Sina, et al.
Veröffentlicht: (2025)
von: Malakouti, Sina, et al.
Veröffentlicht: (2025)
Enhancing Weakly-Supervised Object Detection on Static Images through (Hallucinated) Motion
von: Gungor, Cagri, et al.
Veröffentlicht: (2024)
von: Gungor, Cagri, et al.
Veröffentlicht: (2024)
Leveraging Large Models to Evaluate Novel Content: A Case Study on Advertisement Creativity
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
von: Hou, Zhaoyi Joey, et al.
Veröffentlicht: (2025)
Quantifying the Gaps Between Translation and Native Perception in Training for Multimodal, Multilingual Retrieval
von: Buettner, Kyle, et al.
Veröffentlicht: (2024)
von: Buettner, Kyle, et al.
Veröffentlicht: (2024)
ClimbingCap: Multi-Modal Dataset and Method for Rock Climbing in World Coordinate
von: Yan, Ming, et al.
Veröffentlicht: (2025)
von: Yan, Ming, et al.
Veröffentlicht: (2025)
Culture in Action: Evaluating Text-to-Image Models through Social Activities
von: Malakouti, Sina, et al.
Veröffentlicht: (2025)
von: Malakouti, Sina, et al.
Veröffentlicht: (2025)
Integrating Audio Narrations to Strengthen Domain Generalization in Multimodal First-Person Action Recognition
von: Gungor, Cagri, et al.
Veröffentlicht: (2024)
von: Gungor, Cagri, et al.
Veröffentlicht: (2024)
The Way Up: A Dataset for Hold Usage Detection in Sport Climbing
von: Maschek, Anna, et al.
Veröffentlicht: (2025)
von: Maschek, Anna, et al.
Veröffentlicht: (2025)
A Multimodal Recaptioning Framework to Account for Perceptual Diversity Across Languages in Vision-Language Modeling
von: Buettner, Kyle, et al.
Veröffentlicht: (2025)
von: Buettner, Kyle, et al.
Veröffentlicht: (2025)
Benchmarking VLMs' Reasoning About Persuasive Atypical Images
von: Malakouti, Sina, et al.
Veröffentlicht: (2024)
von: Malakouti, Sina, et al.
Veröffentlicht: (2024)
Read, Watch and Scream! Sound Generation from Text and Video
von: Jeong, Yujin, et al.
Veröffentlicht: (2024)
von: Jeong, Yujin, et al.
Veröffentlicht: (2024)
Sports-Traj: A Unified Trajectory Generation Model for Multi-Agent Movement in Sports
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Poze: Sports Technique Feedback under Data Constraints
von: Singh, Agamdeep, et al.
Veröffentlicht: (2024)
von: Singh, Agamdeep, et al.
Veröffentlicht: (2024)
OrienText: Surface Oriented Textual Image Generation
von: Paliwal, Shubham Singh, et al.
Veröffentlicht: (2025)
von: Paliwal, Shubham Singh, et al.
Veröffentlicht: (2025)
A General Framework for Jersey Number Recognition in Sports Video
von: Koshkina, Maria, et al.
Veröffentlicht: (2024)
von: Koshkina, Maria, et al.
Veröffentlicht: (2024)
Towards Understanding Ambiguity Resolution in Multimodal Inference of Meaning
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition
von: Buettner, Kyle, et al.
Veröffentlicht: (2024)
von: Buettner, Kyle, et al.
Veröffentlicht: (2024)
How Well Can General Vision-Language Models Learn Medicine By Watching Public Educational Videos?
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
No Train Yet Gain: Towards Generic Multi-Object Tracking in Sports and Beyond
von: Stanczyk, Tomasz, et al.
Veröffentlicht: (2025)
von: Stanczyk, Tomasz, et al.
Veröffentlicht: (2025)
CustomText: Customized Textual Image Generation using Diffusion Models
von: Paliwal, Shubham, et al.
Veröffentlicht: (2024)
von: Paliwal, Shubham, et al.
Veröffentlicht: (2024)
Combining OCR Models for Reading Early Modern Printed Books
von: Seuret, Mathias, et al.
Veröffentlicht: (2023)
von: Seuret, Mathias, et al.
Veröffentlicht: (2023)
Aligning Anime Video Generation with Human Feedback
von: Zhu, Bingwen, et al.
Veröffentlicht: (2025)
von: Zhu, Bingwen, et al.
Veröffentlicht: (2025)
Rich Human Feedback for Text-to-Image Generation
von: Liang, Youwei, et al.
Veröffentlicht: (2023)
von: Liang, Youwei, et al.
Veröffentlicht: (2023)
Visually Interpretable Subtask Reasoning for Visual Question Answering
von: Cheng, Yu, et al.
Veröffentlicht: (2025)
von: Cheng, Yu, et al.
Veröffentlicht: (2025)
AGFSync: Leveraging AI-Generated Feedback for Preference Optimization in Text-to-Image Generation
von: An, Jingkun, et al.
Veröffentlicht: (2024)
von: An, Jingkun, et al.
Veröffentlicht: (2024)
Platypus: A Generalized Specialist Model for Reading Text in Various Forms
von: Wang, Peng, et al.
Veröffentlicht: (2024)
von: Wang, Peng, et al.
Veröffentlicht: (2024)
Adversarial Reconstruction Feedback for Robust Fine-grained Generalization
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
von: Wang, Shijie, et al.
Veröffentlicht: (2025)
Improving Personalized Image Generation through Social Context Feedback
von: Gupta, Parul, et al.
Veröffentlicht: (2025)
von: Gupta, Parul, et al.
Veröffentlicht: (2025)
Towards Reliable Advertising Image Generation Using Human Feedback
von: Du, Zhenbang, et al.
Veröffentlicht: (2024)
von: Du, Zhenbang, et al.
Veröffentlicht: (2024)
SportR: A Benchmark for Multimodal Large Language Model Reasoning in Sports
von: Xia, Haotian, et al.
Veröffentlicht: (2025)
von: Xia, Haotian, et al.
Veröffentlicht: (2025)
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
Real-Time Polygonal Semantic Mapping for Humanoid Robot Stair Climbing
von: Bin, Teng, et al.
Veröffentlicht: (2024)
von: Bin, Teng, et al.
Veröffentlicht: (2024)
FeTrIL++: Feature Translation for Exemplar-Free Class-Incremental Learning with Hill-Climbing
von: Hogea, Eduard, et al.
Veröffentlicht: (2024)
von: Hogea, Eduard, et al.
Veröffentlicht: (2024)
Steering Generative Models for Accessibility: EasyRead Image Generation
von: Dickenmann, Nicolas, et al.
Veröffentlicht: (2026)
von: Dickenmann, Nicolas, et al.
Veröffentlicht: (2026)
WAT: Online Video Understanding Needs Watching Before Thinking
von: Han, Zifan, et al.
Veröffentlicht: (2026)
von: Han, Zifan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Learning Consistent Temporal Grounding between Related Tasks in Sports Coaching
von: Rai, Arushi, et al.
Veröffentlicht: (2026) -
VEIL: Vetting Extracted Image Labels from In-the-Wild Captions for Weakly-Supervised Object Detection
von: Rai, Arushi, et al.
Veröffentlicht: (2023) -
The Face of Persuasion: Analyzing Bias and Generating Culture-Aware Ads
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025) -
CAP: Evaluation of Persuasive and Creative Image Generation
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2024) -
Towards Generalization of Tactile Image Generation: Reference-Free Evaluation in a Leakage-Free Setting
von: Gungor, Cagri, et al.
Veröffentlicht: (2025)