ConfusionBench: An Expert-Validated Benchmark for Confusion Recognition and Localization in Educational Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Lu, Wang, Xiao, Frank, Mark, Setlur, Srirangaraj, Govindaraju, Venu, Nwogu, Ifeoma |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ig3D: Integrating 3D Face Representations in Facial Expression Inference
by: Dong, Lu, et al.
Published: (2024)
by: Dong, Lu, et al.
Published: (2024)
InterventionLens: A Multi-Agent Framework for Detecting ASD Intervention Strategies in Parent-Child Shared Reading
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
LLM Augmented Intervenable Multimodal Adaptor for Post-operative Complication Prediction in Lung Cancer Surgery
by: Pandey, Shubham, et al.
Published: (2026)
by: Pandey, Shubham, et al.
Published: (2026)
MistyPilot: An Agentic Fast-Slow Thinking LLM Framework for Misty Social Robots
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
AutoMisty: A Multi-Agent LLM Framework for Automated Code Generation in the Misty Social Robot
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Ridgeformer: Mutli-Stage Contrastive Training For Fine-grained Cross-Domain Fingerprint Recognition
by: Pandey, Shubham, et al.
Published: (2025)
by: Pandey, Shubham, et al.
Published: (2025)
RealCQA-V2: A Diagnostic Benchmark for Structured Visual Entailment over Scientific Charts
by: Ahmed, Saleem, et al.
Published: (2024)
by: Ahmed, Saleem, et al.
Published: (2024)
Cross-Attention Based Influence Model for Manual and Nonmanual Sign Language Analysis
by: Chaudhary, Lipisha, et al.
Published: (2024)
by: Chaudhary, Lipisha, et al.
Published: (2024)
SignAvatar: Sign Language 3D Motion Reconstruction and Generation
by: Dong, Lu, et al.
Published: (2024)
by: Dong, Lu, et al.
Published: (2024)
Audio Match Cutting: Finding and Creating Matching Audio Transitions in Movies and Videos
by: Fedorishin, Dennis, et al.
Published: (2024)
by: Fedorishin, Dennis, et al.
Published: (2024)
SCOT: Self-Supervised Contrastive Pretraining For Zero-Shot Compositional Retrieval
by: Jawade, Bhavin, et al.
Published: (2025)
by: Jawade, Bhavin, et al.
Published: (2025)
Mix from Failure: Confusion-Pairing Mixup for Long-Tailed Recognition
by: Yoon, Youngseok, et al.
Published: (2024)
by: Yoon, Youngseok, et al.
Published: (2024)
Dynamic Inter-Class Confusion-Aware Encoder for Audio-Visual Fusion in Human Activity Recognition
by: Cong, Kaixuan, et al.
Published: (2025)
by: Cong, Kaixuan, et al.
Published: (2025)
Evaluating Attribute Confusion in Fashion Text-to-Image Generation
by: Liu, Ziyue, et al.
Published: (2025)
by: Liu, Ziyue, et al.
Published: (2025)
Backdooring CLIP through Concept Confusion
by: Hu, Lijie, et al.
Published: (2025)
by: Hu, Lijie, et al.
Published: (2025)
Logits DeConfusion with CLIP for Few-Shot Learning
by: Li, Shuo, et al.
Published: (2025)
by: Li, Shuo, et al.
Published: (2025)
Learning Action Hierarchies via Hybrid Geometric Diffusion
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
Vision Language Models are Confused Tourists
by: Irawan, Patrick Amadeus, et al.
Published: (2025)
by: Irawan, Patrick Amadeus, et al.
Published: (2025)
DeconfuseTrack:Dealing with Confusion for Multi-Object Tracking
by: Huang, Cheng, et al.
Published: (2024)
by: Huang, Cheng, et al.
Published: (2024)
The Comparability of Model Fusion to Measured Data in Confuser Rejection
by: Flynn, Conor, et al.
Published: (2025)
by: Flynn, Conor, et al.
Published: (2025)
Uncovering Entity Identity Confusion in Multimodal Knowledge Editing
by: Wu, Shu, et al.
Published: (2026)
by: Wu, Shu, et al.
Published: (2026)
De-Confusing Pseudo-Labels in Source-Free Domain Adaptation
by: Diamant, Idit, et al.
Published: (2024)
by: Diamant, Idit, et al.
Published: (2024)
Towards Open Domain Text-Driven Synthesis of Multi-Person Motions
by: Shan, Mengyi, et al.
Published: (2024)
by: Shan, Mengyi, et al.
Published: (2024)
Closing the Confusion Loop: CLIP-Guided Alignment for Source-Free Domain Adaptation
by: Wang, Shanshan, et al.
Published: (2026)
by: Wang, Shanshan, et al.
Published: (2026)
Reliable Multimodal Learning Via Multi-Level Adaptive DeConfusion
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
Resolving Multi-Condition Confusion for Finetuning-Free Personalized Image Generation
by: Huang, Qihan, et al.
Published: (2024)
by: Huang, Qihan, et al.
Published: (2024)
Confusing Pair Correction Based on Category Prototype for Domain Adaptation under Noisy Environments
by: Zhi, Churan, et al.
Published: (2024)
by: Zhi, Churan, et al.
Published: (2024)
CUE: Concept-Aware Multi-Label Expansion to Mitigate Concept Confusion in Long-Tailed Learning
by: Zhang, Ruichi, et al.
Published: (2026)
by: Zhang, Ruichi, et al.
Published: (2026)
Regularizing CNNs using Confusion Penalty Based Label Smoothing for Histopathology Images
by: Kuiry, Somenath, et al.
Published: (2024)
by: Kuiry, Somenath, et al.
Published: (2024)
When Eyes and Ears Disagree: Can MLLMs Discern Audio-Visual Confusion?
by: Ye, Qilang, et al.
Published: (2025)
by: Ye, Qilang, et al.
Published: (2025)
Delve into Base-Novel Confusion: Redundancy Exploration for Few-Shot Class-Incremental Learning
by: Zhou, Haichen, et al.
Published: (2024)
by: Zhou, Haichen, et al.
Published: (2024)
Fusion or Confusion? Assessing the impact of visible-thermal image fusion for automated wildlife detection
by: Dionne-Pierre, Camille, et al.
Published: (2025)
by: Dionne-Pierre, Camille, et al.
Published: (2025)
Universal Incremental Learning: Mitigating Confusion from Inter- and Intra-task Distribution Randomness
by: Luo, Sheng, et al.
Published: (2025)
by: Luo, Sheng, et al.
Published: (2025)
Combining Discrepancy-Confusion Uncertainty and Calibration Diversity for Active Fine-Grained Image Classification
by: Jin, Yinghao, et al.
Published: (2025)
by: Jin, Yinghao, et al.
Published: (2025)
CoCoGaussian: Leveraging Circle of Confusion for Gaussian Splatting from Defocused Images
by: Lee, Jungho, et al.
Published: (2024)
by: Lee, Jungho, et al.
Published: (2024)
CAPT: Confusion-Aware Prompt Tuning for Reducing Vision-Language Misalignment
by: Shao, Maoyuan, et al.
Published: (2026)
by: Shao, Maoyuan, et al.
Published: (2026)
Intensity Confusion Matters: An Intensity-Distance Guided Loss for Bronchus Segmentation
by: Gong, Haifan, et al.
Published: (2024)
by: Gong, Haifan, et al.
Published: (2024)
AIGCs Confuse AI Too: Investigating and Explaining Synthetic Image-induced Hallucinations in Large Vision-Language Models
by: Gao, Yifei, et al.
Published: (2024)
by: Gao, Yifei, et al.
Published: (2024)
ConceptGuard: Continual Personalized Text-to-Image Generation with Forgetting and Confusion Mitigation
by: Guo, Zirun, et al.
Published: (2025)
by: Guo, Zirun, et al.
Published: (2025)
Visual Confused Deputy: Exploiting and Defending Perception Failures in Computer-Using Agents
by: Liu, Xunzhuo, et al.
Published: (2026)
by: Liu, Xunzhuo, et al.
Published: (2026)
Similar Items
-
Ig3D: Integrating 3D Face Representations in Facial Expression Inference
by: Dong, Lu, et al.
Published: (2024) -
InterventionLens: A Multi-Agent Framework for Detecting ASD Intervention Strategies in Parent-Child Shared Reading
by: Wang, Xiao, et al.
Published: (2026) -
LLM Augmented Intervenable Multimodal Adaptor for Post-operative Complication Prediction in Lung Cancer Surgery
by: Pandey, Shubham, et al.
Published: (2026) -
MistyPilot: An Agentic Fast-Slow Thinking LLM Framework for Misty Social Robots
by: Wang, Xiao, et al.
Published: (2026) -
AutoMisty: A Multi-Agent LLM Framework for Automated Code Generation in the Misty Social Robot
by: Wang, Xiao, et al.
Published: (2025)