Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nagy, Rajmund, Voss, Hendric, Hoang-Minh, Thanh, Tsakov, Mihail, Nikolov, Teodor, Zhang, Zeyi, Ao, Tenglong, Yang, Sicheng, Huang, Shaoli, Cheng, Yongkang, Mughal, M. Hamza, Dabral, Rishabh, Chhatre, Kiran, Theobalt, Christian, Liu, Libin, Kopp, Stefan, McDonnell, Rachel, Neff, Michael, Kucherenko, Taras, Yoon, Youngwoo, Henter, Gustav Eje |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating gesture generation in a large-scale open challenge: The GENEA Challenge 2022
von: Kucherenko, Taras, et al.
Veröffentlicht: (2023)
von: Kucherenko, Taras, et al.
Veröffentlicht: (2023)
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
Matcha-TTS: A fast TTS architecture with conditional flow matching
von: Mehta, Shivam, et al.
Veröffentlicht: (2023)
von: Mehta, Shivam, et al.
Veröffentlicht: (2023)
Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
von: Mehta, Shivam, et al.
Veröffentlicht: (2024)
von: Mehta, Shivam, et al.
Veröffentlicht: (2024)
Unified speech and gesture synthesis using flow matching
von: Mehta, Shivam, et al.
Veröffentlicht: (2023)
von: Mehta, Shivam, et al.
Veröffentlicht: (2023)
Fake it to make it: Using synthetic data to remedy the data shortage in joint multimodal speech-and-gesture synthesis
von: Mehta, Shivam, et al.
Veröffentlicht: (2024)
von: Mehta, Shivam, et al.
Veröffentlicht: (2024)
3HANDS Dataset: Learning from Humans for Generating Naturalistic Handovers with Supernumerary Robotic Limbs
von: Abadian, Artin Saberpour, et al.
Veröffentlicht: (2025)
von: Abadian, Artin Saberpour, et al.
Veröffentlicht: (2025)
THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion
von: Ioan, Calin Teodor
Veröffentlicht: (2025)
von: Ioan, Calin Teodor
Veröffentlicht: (2025)
Hybrid-AIRL: Enhancing Inverse Reinforcement Learning with Supervised Expert Guidance
von: Silue, Bram, et al.
Veröffentlicht: (2025)
von: Silue, Bram, et al.
Veröffentlicht: (2025)
Transformers Meet Relational Databases
von: Peleška, Jakub, et al.
Veröffentlicht: (2024)
von: Peleška, Jakub, et al.
Veröffentlicht: (2024)
Exploring Internal Numeracy in Language Models: A Case Study on ALBERT
von: Wennberg, Ulme, et al.
Veröffentlicht: (2024)
von: Wennberg, Ulme, et al.
Veröffentlicht: (2024)
Composition of rocks and minerals from the Mid-Atlantic Ridge 5-7°N
von: Pushcharovsky, Yury M, et al.
Veröffentlicht: (2004)
von: Pushcharovsky, Yury M, et al.
Veröffentlicht: (2004)
Chemical composition of basalts and basaltic glasses from the Sierra Leone Fracture Zone region
von: Skolotnev, Sergey G, et al.
Veröffentlicht: (2003)
von: Skolotnev, Sergey G, et al.
Veröffentlicht: (2003)
Physical oceanography during L' Atalante cruise Almofront-1
von: Prieur, Louis Marie
Veröffentlicht: (2012)
von: Prieur, Louis Marie
Veröffentlicht: (2012)
OnlineSplatter: Pose-Free Online 3D Reconstruction for Free-Moving Objects
von: Huang, Mark He, et al.
Veröffentlicht: (2025)
von: Huang, Mark He, et al.
Veröffentlicht: (2025)
Chemical and isotopic compositions of basalts and glasses from lavas of the Sierra Leone fault site
von: Sharkov, E V, et al.
Veröffentlicht: (2008)
von: Sharkov, E V, et al.
Veröffentlicht: (2008)
(Table 1) Rock types dredged at stations of R/V Akademik Nikolaj Strakhov (Cruise 22) and R/V Akademik Ioffe (Cruise 10), Mid-Atlantic Ridge 5-7°N
von: Pushcharovsky, Yury M, et al.
Veröffentlicht: (2004)
von: Pushcharovsky, Yury M, et al.
Veröffentlicht: (2004)
Nutrients measured on water bottle samples during L'Atalante cruise Almofront-1
von: Prieur, Louis Marie
Veröffentlicht: (2012)
von: Prieur, Louis Marie
Veröffentlicht: (2012)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
FLD+: Data-efficient Evaluation Metric for Generative Models
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Normalizing Flow-Based Metric for Image Generation
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
Pigments measured on water bottle samples during cruise ALMOFRONT-1
von: Prieur, Louis Marie, et al.
Veröffentlicht: (2012)
von: Prieur, Louis Marie, et al.
Veröffentlicht: (2012)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
IMUVIE: Pickup Timeline Action Localization via Motion Movies
von: Clapham, John, et al.
Veröffentlicht: (2024)
von: Clapham, John, et al.
Veröffentlicht: (2024)
Trapped in texture bias? A large scale comparison of deep instance segmentation
von: Theodoridis, Johannes, et al.
Veröffentlicht: (2024)
von: Theodoridis, Johannes, et al.
Veröffentlicht: (2024)
Evaluation Metric for Quality Control and Generative Models in Histopathology Images
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
WaveMix: A Resource-efficient Neural Network for Image Analysis
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Few-Shot Learning of a Graph-Based Neural Network Model Without Backpropagation
von: Lapin, Mykyta, et al.
Veröffentlicht: (2025)
von: Lapin, Mykyta, et al.
Veröffentlicht: (2025)
Low-Cost Tree Crown Dieback Estimation Using Deep Learning-Based Segmentation
von: Allen, M. J., et al.
Veröffentlicht: (2024)
von: Allen, M. J., et al.
Veröffentlicht: (2024)
(Table 13) Chemical composition of Fe-Mn crusts between the Bogdanov Fault and the Sierra Leone Tectonic Zone
von: Bazilevskaya, Elena S
Veröffentlicht: (2007)
von: Bazilevskaya, Elena S
Veröffentlicht: (2007)
The Company You Keep: How LLMs Respond to Dark Triad Traits
von: Lu, Zeyi, et al.
Veröffentlicht: (2026)
von: Lu, Zeyi, et al.
Veröffentlicht: (2026)
Physical oceanography during Jean Charcot cruise MEDATLANTE-I
von: Raunet, J, et al.
Veröffentlicht: (2012)
von: Raunet, J, et al.
Veröffentlicht: (2012)
FEDTAIL: Federated Long-Tailed Domain Generalization with Sharpness-Guided Gradient Matching
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
PhysHand: A Hand Simulation Model with Physiological Geometry, Physical Deformation, and Accurate Contact Handling
von: Sun, Mingyang, et al.
Veröffentlicht: (2024)
von: Sun, Mingyang, et al.
Veröffentlicht: (2024)
Entity Re-identification in Visual Storytelling via Contrastive Reinforcement Learning
von: Oliveira, Daniel A. P., et al.
Veröffentlicht: (2025)
von: Oliveira, Daniel A. P., et al.
Veröffentlicht: (2025)
Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustainable logging and post-harvest silvicultural treatment in tropical forest using airborne robotics systems
von: Albiero, Daniel, et al.
Veröffentlicht: (2026)
von: Albiero, Daniel, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evaluating gesture generation in a large-scale open challenge: The GENEA Challenge 2022
von: Kucherenko, Taras, et al.
Veröffentlicht: (2023) -
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024) -
Matcha-TTS: A fast TTS architecture with conditional flow matching
von: Mehta, Shivam, et al.
Veröffentlicht: (2023) -
Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
von: Mehta, Shivam, et al.
Veröffentlicht: (2024) -
Unified speech and gesture synthesis using flow matching
von: Mehta, Shivam, et al.
Veröffentlicht: (2023)