Embracing Diversity: Interpretable Zero-shot classification beyond one vector per class
Fuente:
arXiv
Saved in:
| Main Authors: | Moayeri, Mazda, Rabbat, Michael, Ibrahim, Mark, Bouchacourt, Diane |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
Learning Annotation Consensus for Continuous Emotion Recognition
by: Shoer, Ibrahim, et al.
Published: (2025)
by: Shoer, Ibrahim, et al.
Published: (2025)
PRIME: Prioritizing Interpretability in Failure Mode Extraction
by: Rezaei, Keivan, et al.
Published: (2023)
by: Rezaei, Keivan, et al.
Published: (2023)
Personalized Interpretability -- Interactive Alignment of Prototypical Parts Networks
by: Michalski, Tomasz, et al.
Published: (2025)
by: Michalski, Tomasz, et al.
Published: (2025)
Improving multidimensional projection quality with user-specific metrics and optimal scaling
by: Ibrahim, Maniru
Published: (2024)
by: Ibrahim, Maniru
Published: (2024)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
by: Hall, Melissa, et al.
Published: (2023)
by: Hall, Melissa, et al.
Published: (2023)
Allowing humans to interactively guide machines where to look does not always improve human-AI team's classification accuracy
by: Nguyen, Giang, et al.
Published: (2024)
by: Nguyen, Giang, et al.
Published: (2024)
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
by: Sezer, Berk, et al.
Published: (2026)
by: Sezer, Berk, et al.
Published: (2026)
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
Multimodal Attention-Aware Fusion for Diagnosing Distal Myopathy: Evaluating Model Interpretability and Clinician Trust
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
by: Onari, Mohsen Abbaspour, et al.
Published: (2025)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models
by: Meinecke, Christofer, et al.
Published: (2025)
by: Meinecke, Christofer, et al.
Published: (2025)
Dream360: Diverse and Immersive Outdoor Virtual Scene Creation via Transformer-Based 360 Image Outpainting
by: Ai, Hao, et al.
Published: (2024)
by: Ai, Hao, et al.
Published: (2024)
Facial Movement Dynamics Reveal Workload During Complex Multitasking
by: Sale, Carter, et al.
Published: (2026)
by: Sale, Carter, et al.
Published: (2026)
Plug-and-Play Clarifier: A Zero-Shot Multimodal Framework for Egocentric Intent Disambiguation
by: Yang, Sicheng, et al.
Published: (2025)
by: Yang, Sicheng, et al.
Published: (2025)
SymbolSight: Minimizing Inter-Symbol Interference for Reading with Prosthetic Vision
by: Lesner, Jasmine, et al.
Published: (2026)
by: Lesner, Jasmine, et al.
Published: (2026)
Establishing a Baseline for Gaze-driven Authentication Performance in VR: A Breadth-First Investigation on a Very Large Dataset
by: Lohr, Dillon, et al.
Published: (2024)
by: Lohr, Dillon, et al.
Published: (2024)
Weak-Annotation of HAR Datasets using Vision Foundation Models
by: Bock, Marius, et al.
Published: (2024)
by: Bock, Marius, et al.
Published: (2024)
WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition
by: Bock, Marius, et al.
Published: (2023)
by: Bock, Marius, et al.
Published: (2023)
Zero-shot Emotion Annotation in Facial Images Using Large Multimodal Models: Benchmarking and Prospects for Multi-Class, Multi-Frame Approaches
by: Zhang, He, et al.
Published: (2025)
by: Zhang, He, et al.
Published: (2025)
Efficient 3D Reconstruction, Streaming and Visualization of Static and Dynamic Scene Parts for Multi-client Live-telepresence in Large-scale Environments
by: Van Holland, Leif, et al.
Published: (2022)
by: Van Holland, Leif, et al.
Published: (2022)
Ocular Authentication: Fusion of Gaze and Periocular Modalities
by: Lohr, Dillon, et al.
Published: (2025)
by: Lohr, Dillon, et al.
Published: (2025)
CHiQPM: Calibrated Hierarchical Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
Interpretable facial dynamics as behavioral and perceptual traces of deepfakes
by: Murphy, Timothy Joseph, et al.
Published: (2026)
by: Murphy, Timothy Joseph, et al.
Published: (2026)
Deep Neural Encoder-Decoder Model to Relate fMRI Brain Activity with Naturalistic Stimuli
by: David, Florian, et al.
Published: (2025)
by: David, Florian, et al.
Published: (2025)
Designing Multi-Robot Ground Video Sensemaking with Public Safety Professionals
by: Zhou, Puqi, et al.
Published: (2026)
by: Zhou, Puqi, et al.
Published: (2026)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
by: Chen, Tianrun, et al.
Published: (2024)
by: Chen, Tianrun, et al.
Published: (2024)
Eye Gaze as a Signal for Conveying User Attention in Contextual AI Systems
by: Wilson, Ethan, et al.
Published: (2025)
by: Wilson, Ethan, et al.
Published: (2025)
STAR: Smartphone-analogous Typing in Augmented Reality
by: Kim, Taejun, et al.
Published: (2025)
by: Kim, Taejun, et al.
Published: (2025)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
by: Huang, Jinbin, et al.
Published: (2024)
by: Huang, Jinbin, et al.
Published: (2024)
Zero-Shot Segmentation of Eye Features Using the Segment Anything Model (SAM)
by: Maquiling, Virmarie, et al.
Published: (2023)
by: Maquiling, Virmarie, et al.
Published: (2023)
Do humans and Convolutional Neural Networks attend to similar areas during scene classification: Effects of task and image type
by: Müller, Romy, et al.
Published: (2023)
by: Müller, Romy, et al.
Published: (2023)
DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification
by: Zimmermann, Robert, et al.
Published: (2026)
by: Zimmermann, Robert, et al.
Published: (2026)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
by: Luthra, Prerna
Published: (2025)
by: Luthra, Prerna
Published: (2025)
Context-Awareness and Interpretability of Rare Occurrences for Discovery and Formalization of Critical Failure Modes
by: Polavaram, Sridevi, et al.
Published: (2025)
by: Polavaram, Sridevi, et al.
Published: (2025)
The Visual Experience Dataset: Over 200 Recorded Hours of Integrated Eye Movement, Odometry, and Egocentric Video
by: Greene, Michelle R., et al.
Published: (2024)
by: Greene, Michelle R., et al.
Published: (2024)
Real-Time Cellist Postural Evaluation With On-Device Computer Vision
by: Wang, Paolo, et al.
Published: (2026)
by: Wang, Paolo, et al.
Published: (2026)
Zero-Shot Pupil Segmentation with SAM 2: A Case Study of Over 14 Million Images
by: Maquiling, Virmarie, et al.
Published: (2024)
by: Maquiling, Virmarie, et al.
Published: (2024)
emg2pose: A Large and Diverse Benchmark for Surface Electromyographic Hand Pose Estimation
by: Salter, Sasha, et al.
Published: (2024)
by: Salter, Sasha, et al.
Published: (2024)
Similar Items
-
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
by: Karnoor, Sahil Bhandary, et al.
Published: (2025) -
Learning Annotation Consensus for Continuous Emotion Recognition
by: Shoer, Ibrahim, et al.
Published: (2025) -
PRIME: Prioritizing Interpretability in Failure Mode Extraction
by: Rezaei, Keivan, et al.
Published: (2023) -
Personalized Interpretability -- Interactive Alignment of Prototypical Parts Networks
by: Michalski, Tomasz, et al.
Published: (2025) -
Improving multidimensional projection quality with user-specific metrics and optimal scaling
by: Ibrahim, Maniru
Published: (2024)