Where Does My Model Underperform? A Human Evaluation of Slice Discovery Algorithms
Fuente:
arXiv
Saved in:
| Main Authors: | Johnson, Nari, Cabrera, Ángel Alexander, Plumb, Gregory, Talwalkar, Ameet |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
by: Yan, Xinyuan, et al.
Published: (2025)
by: Yan, Xinyuan, et al.
Published: (2025)
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
by: Xuan, Xiwei, et al.
Published: (2024)
by: Xuan, Xiwei, et al.
Published: (2024)
When, Where, and What? A Novel Benchmark for Accident Anticipation and Localization with Large Language Models
by: Liao, Haicheng, et al.
Published: (2024)
by: Liao, Haicheng, et al.
Published: (2024)
EyeTrAES: Fine-grained, Low-Latency Eye Tracking via Adaptive Event Slicing
by: Sen, Argha, et al.
Published: (2024)
by: Sen, Argha, et al.
Published: (2024)
Scene-Aware Urban Design: A Human-AI Recommendation Framework Using Co-Occurrence Embeddings and Vision-Language Models
by: Gallardo, Rodrigo, et al.
Published: (2025)
by: Gallardo, Rodrigo, et al.
Published: (2025)
VFA: Vision Frequency Analysis of Foundation Models and Human
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
Human-AI Collaboration and Explainability for 2D/3D Registration Quality Assurance
by: Cho, Sue Min, et al.
Published: (2025)
by: Cho, Sue Min, et al.
Published: (2025)
Probabilistic Human Intent Prediction for Mobile Manipulation: An Evaluation with Human-Inspired Constraints
by: Contreras, Cesar Alan, et al.
Published: (2025)
by: Contreras, Cesar Alan, et al.
Published: (2025)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
by: Verma, Arnav, et al.
Published: (2025)
by: Verma, Arnav, et al.
Published: (2025)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
L-WISE: Boosting Human Visual Category Learning Through Model-Based Image Selection and Enhancement
by: Talbot, Morgan B., et al.
Published: (2024)
by: Talbot, Morgan B., et al.
Published: (2024)
MicroBi-ConvLSTM: An Ultra-Lightweight Efficient Model for Human Activity Recognition on Resource Constrained Devices
by: Mandal, Mridankan
Published: (2026)
by: Mandal, Mridankan
Published: (2026)
BabyMamba-HAR: Lightweight Selective State Space Models for Efficient Human Activity Recognition on Resource Constrained Devices
by: Mandal, Mridankan
Published: (2026)
by: Mandal, Mridankan
Published: (2026)
Multimodal Attention-Aware Fusion for Diagnosing Distal Myopathy: Evaluating Model Interpretability and Clinician Trust
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
Algorithmic Ways of Seeing: Using Object Detection to Facilitate Art Exploration
by: Meyer, Louie Søs, et al.
Published: (2024)
by: Meyer, Louie Søs, et al.
Published: (2024)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
ExeChecker: Where Did I Go Wrong?
by: Gu, Yiwen, et al.
Published: (2024)
by: Gu, Yiwen, et al.
Published: (2024)
Deep Learning-based Lightweight RGB Object Tracking for Augmented Reality Devices
by: Smith, Alice, et al.
Published: (2025)
by: Smith, Alice, et al.
Published: (2025)
"It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language Models
by: Garg, Kapil, et al.
Published: (2025)
by: Garg, Kapil, et al.
Published: (2025)
What They Saw, Not Just Where They Looked: Semantic Scanpath Similarity via VLMs and NLP metric
by: Kerkouri, Mohamed Amine, et al.
Published: (2026)
by: Kerkouri, Mohamed Amine, et al.
Published: (2026)
Extracting Human Attention through Crowdsourced Patch Labeling
by: Chang, Minsuk, et al.
Published: (2024)
by: Chang, Minsuk, et al.
Published: (2024)
Evaluating the Evaluators: Towards Human-aligned Metrics for Missing Markers Reconstruction
by: Kucherenko, Taras, et al.
Published: (2024)
by: Kucherenko, Taras, et al.
Published: (2024)
Benchmarking XAI Explanations with Human-Aligned Evaluations
by: Kazmierczak, Rémi, et al.
Published: (2024)
by: Kazmierczak, Rémi, et al.
Published: (2024)
Do Object Detection Localization Errors Affect Human Performance and Trust?
by: de Witte, Sven, et al.
Published: (2024)
by: de Witte, Sven, et al.
Published: (2024)
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026)
by: Shaw, Richard, et al.
Published: (2026)
MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation
by: Li, Jia, et al.
Published: (2024)
by: Li, Jia, et al.
Published: (2024)
LiDAR-based Human Activity Recognition through Laplacian Spectral Analysis
by: Sharifipour, Sasan, et al.
Published: (2025)
by: Sharifipour, Sasan, et al.
Published: (2025)
Analyzing Data Efficiency and Performance of Machine Learning Algorithms for Assessing Low Back Pain Physical Rehabilitation Exercises
by: Marusic, Aleksa, et al.
Published: (2024)
by: Marusic, Aleksa, et al.
Published: (2024)
Text-to-Image Representativity Fairness Evaluation Framework
by: Yamani, Asma, et al.
Published: (2024)
by: Yamani, Asma, et al.
Published: (2024)
SOS: A Shuffle Order Strategy for Data Augmentation in Industrial Human Activity Recognition
by: Ha, Anh Tuan, et al.
Published: (2025)
by: Ha, Anh Tuan, et al.
Published: (2025)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
Shape vs. Context: Examining Human--AI Gaps in Ambiguous Japanese Character Recognition
by: Haraguchi, Daichi
Published: (2026)
by: Haraguchi, Daichi
Published: (2026)
Design and Evaluation of Camera-Centric Mobile Crowdsourcing Applications
by: Stylianou, Abby, et al.
Published: (2024)
by: Stylianou, Abby, et al.
Published: (2024)
GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
by: Konrad, Robert, et al.
Published: (2024)
by: Konrad, Robert, et al.
Published: (2024)
Real-Time Cellist Postural Evaluation With On-Device Computer Vision
by: Wang, Paolo, et al.
Published: (2026)
by: Wang, Paolo, et al.
Published: (2026)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
by: Hall, Melissa, et al.
Published: (2023)
by: Hall, Melissa, et al.
Published: (2023)
Extended Reality for Mental Health Evaluation -A Scoping Review
by: Olatunji, Omisore, et al.
Published: (2022)
by: Olatunji, Omisore, et al.
Published: (2022)
JAX-IK: Real-Time Inverse Kinematics for Generating Multi-Constrained Movements of Virtual Human Characters
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
Similar Items
-
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
by: Yan, Xinyuan, et al.
Published: (2025) -
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
by: Xuan, Xiwei, et al.
Published: (2024) -
When, Where, and What? A Novel Benchmark for Accident Anticipation and Localization with Large Language Models
by: Liao, Haicheng, et al.
Published: (2024) -
EyeTrAES: Fine-grained, Low-Latency Eye Tracking via Adaptive Event Slicing
by: Sen, Argha, et al.
Published: (2024) -
Scene-Aware Urban Design: A Human-AI Recommendation Framework Using Co-Occurrence Embeddings and Vision-Language Models
by: Gallardo, Rodrigo, et al.
Published: (2025)