Evaluating the Evaluators: Towards Human-aligned Metrics for Missing Markers Reconstruction
Fuente:
arXiv
Saved in:
| Main Authors: | Kucherenko, Taras, Peristy, Derek, Bütepage, Judith |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
by: Nagy, Rajmund, et al.
Published: (2024)
by: Nagy, Rajmund, et al.
Published: (2024)
Where Does My Model Underperform? A Human Evaluation of Slice Discovery Algorithms
by: Johnson, Nari, et al.
Published: (2023)
by: Johnson, Nari, et al.
Published: (2023)
A General Model for Detecting Learner Engagement: Implementation and Evaluation
by: Malekshahi, Somayeh, et al.
Published: (2024)
by: Malekshahi, Somayeh, et al.
Published: (2024)
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
by: Zhang, Dongping, et al.
Published: (2024)
by: Zhang, Dongping, et al.
Published: (2024)
An Evaluation of Hybrid Annotation Workflows on High-Ambiguity Spatiotemporal Video Footage
by: Gutiérrez, Juan, et al.
Published: (2025)
by: Gutiérrez, Juan, et al.
Published: (2025)
Seeing Eye to AI? Applying Deep-Feature-Based Similarity Metrics to Information Visualization
by: Long, Sheng, et al.
Published: (2025)
by: Long, Sheng, et al.
Published: (2025)
Evaluating how interactive visualizations can assist in finding samples where and how computer vision models make mistakes
by: Song, Hayeong, et al.
Published: (2023)
by: Song, Hayeong, et al.
Published: (2023)
Continuous Human Action Recognition for Human-Machine Interaction: A Review
by: Gammulle, Harshala, et al.
Published: (2022)
by: Gammulle, Harshala, et al.
Published: (2022)
Human Motion Synthesis_ A Diffusion Approach for Motion Stitching and In-Betweening
by: Adewole, Michael, et al.
Published: (2024)
by: Adewole, Michael, et al.
Published: (2024)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition
by: Zhou, Zishu, et al.
Published: (2026)
by: Zhou, Zishu, et al.
Published: (2026)
From Model Uncertainty to Human Attention: Localization-Aware Visual Cues for Scalable Annotation Review
by: Sbeyti, Moussa Kassem, et al.
Published: (2026)
by: Sbeyti, Moussa Kassem, et al.
Published: (2026)
Intervening in Black Box: Concept Bottleneck Model for Enhancing Human Neural Network Mutual Understanding
by: Xiong, Nuoye, et al.
Published: (2025)
by: Xiong, Nuoye, et al.
Published: (2025)
MiMICRI: Towards Domain-centered Counterfactual Explanations of Cardiovascular Image Classification Models
by: Guo, Grace, et al.
Published: (2024)
by: Guo, Grace, et al.
Published: (2024)
Supporting Experts with a Multimodal Machine-Learning-Based Tool for Human Behavior Analysis of Conversational Videos
by: Arakawa, Riku, et al.
Published: (2024)
by: Arakawa, Riku, et al.
Published: (2024)
Evaluating Deep Human-in-the-Loop Optimization for Retinal Implants Using Sighted Participants
by: Schoinas, Eirini, et al.
Published: (2025)
by: Schoinas, Eirini, et al.
Published: (2025)
VRMN-bD: A Multi-modal Natural Behavior Dataset of Immersive Human Fear Responses in VR Stand-up Interactive Games
by: Zhang, He, et al.
Published: (2024)
by: Zhang, He, et al.
Published: (2024)
Classification Metrics for Image Explanations: Towards Building Reliable XAI-Evaluations
by: Fresz, Benjamin, et al.
Published: (2024)
by: Fresz, Benjamin, et al.
Published: (2024)
Testing Human-Hand Segmentation on In-Distribution and Out-of-Distribution Data in Human-Robot Interactions Using a Deep Ensemble Model
by: Jalayer, Reza, et al.
Published: (2025)
by: Jalayer, Reza, et al.
Published: (2025)
A Survey of Body and Face Motion: Datasets, Performance Evaluation Metrics and Generative Techniques
by: Sookha, Lownish Rai, et al.
Published: (2025)
by: Sookha, Lownish Rai, et al.
Published: (2025)
Human-in-the-Loop Segmentation of Multi-species Coral Imagery
by: Raine, Scarlett, et al.
Published: (2024)
by: Raine, Scarlett, et al.
Published: (2024)
Benchmarking Adaptive Intelligence and Computer Vision on Human-Robot Collaboration
by: Saraj, Salaar, et al.
Published: (2024)
by: Saraj, Salaar, et al.
Published: (2024)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
by: Verma, Arnav, et al.
Published: (2025)
by: Verma, Arnav, et al.
Published: (2025)
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
Toward a Surgeon-in-the-Loop Ophthalmic Robotic Apprentice using Reinforcement and Imitation Learning
by: Gomaa, Amr, et al.
Published: (2023)
by: Gomaa, Amr, et al.
Published: (2023)
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
by: Sezer, Berk, et al.
Published: (2026)
by: Sezer, Berk, et al.
Published: (2026)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
by: Bujack, Roxana, et al.
Published: (2026)
by: Bujack, Roxana, et al.
Published: (2026)
Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark
by: Nagy, Rajmund, et al.
Published: (2025)
by: Nagy, Rajmund, et al.
Published: (2025)
No Need to Sacrifice Data Quality for Quantity: Crowd-Informed Machine Annotation for Cost-Effective Understanding of Visual Data
by: Klugmann, Christopher, et al.
Published: (2024)
by: Klugmann, Christopher, et al.
Published: (2024)
Helios: An extremely low power event-based gesture recognition for always-on smart eyewear
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
by: Bhattacharyya, Prarthana, et al.
Published: (2024)
SLIMBRAIN: Augmented Reality Real-Time Acquisition and Processing System For Hyperspectral Classification Mapping with Depth Information for In-Vivo Surgical Procedures
by: Sancho, Jaime, et al.
Published: (2024)
by: Sancho, Jaime, et al.
Published: (2024)
Improve accessibility for Low Vision and Blind people using Machine Learning and Computer Vision
by: Shukurov, Jasur
Published: (2024)
by: Shukurov, Jasur
Published: (2024)
ExeChecker: Where Did I Go Wrong?
by: Gu, Yiwen, et al.
Published: (2024)
by: Gu, Yiwen, et al.
Published: (2024)
Using a CNN Model to Assess Paintings' Creativity
by: Zhang, Zhehan, et al.
Published: (2024)
by: Zhang, Zhehan, et al.
Published: (2024)
Methodology to Deploy CNN-Based Computer Vision Models on Immersive Wearable Devices
by: Malek, Kaveh, et al.
Published: (2024)
by: Malek, Kaveh, et al.
Published: (2024)
emg2pose: A Large and Diverse Benchmark for Surface Electromyographic Hand Pose Estimation
by: Salter, Sasha, et al.
Published: (2024)
by: Salter, Sasha, et al.
Published: (2024)
Random Channel Ablation for Robust Hand Gesture Classification with Multimodal Biosignals
by: Bimbraw, Keshav, et al.
Published: (2024)
by: Bimbraw, Keshav, et al.
Published: (2024)
ConvMixFormer- A Resource-efficient Convolution Mixer for Transformer-based Dynamic Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2024)
by: Garg, Mallika, et al.
Published: (2024)
IVISIT: An Interactive Visual Simulation Tool for system simulation, visualization, optimization, and parameter management
by: Knoblauch, Andreas
Published: (2024)
by: Knoblauch, Andreas
Published: (2024)
Similar Items
-
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
by: Nagy, Rajmund, et al.
Published: (2024) -
Where Does My Model Underperform? A Human Evaluation of Slice Discovery Algorithms
by: Johnson, Nari, et al.
Published: (2023) -
A General Model for Detecting Learner Engagement: Implementation and Evaluation
by: Malekshahi, Somayeh, et al.
Published: (2024) -
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
by: Zhang, Dongping, et al.
Published: (2024) -
An Evaluation of Hybrid Annotation Workflows on High-Ambiguity Spatiotemporal Video Footage
by: Gutiérrez, Juan, et al.
Published: (2025)