An Evaluation of Hybrid Annotation Workflows on High-Ambiguity Spatiotemporal Video Footage
Fuente:
arXiv
Saved in:
| Main Authors: | Gutiérrez, Juan, Gutiérrez, Victor, Mora, Ángel, Rodriguez, Silvia, Blanco, José Luis |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos
by: Hamdi, Abdullah, et al.
Published: (2026)
by: Hamdi, Abdullah, et al.
Published: (2026)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
Learning Annotation Consensus for Continuous Emotion Recognition
by: Shoer, Ibrahim, et al.
Published: (2025)
by: Shoer, Ibrahim, et al.
Published: (2025)
Confidence Contours: Uncertainty-Aware Annotation for Medical Semantic Segmentation
by: Ye, Andre, et al.
Published: (2023)
by: Ye, Andre, et al.
Published: (2023)
Weak-Annotation of HAR Datasets using Vision Foundation Models
by: Bock, Marius, et al.
Published: (2024)
by: Bock, Marius, et al.
Published: (2024)
Not all Blends are Equal: The BLEMORE Dataset of Blended Emotion Expressions with Relative Salience Annotations
by: Lachmann, Tim, et al.
Published: (2026)
by: Lachmann, Tim, et al.
Published: (2026)
Detecting Activities of Daily Living in Egocentric Video to Contextualize Hand Use at Home in Outpatient Neurorehabilitation Settings
by: Kadambi, Adesh, et al.
Published: (2024)
by: Kadambi, Adesh, et al.
Published: (2024)
UST-Hand: An Uncertainty-aware Spatiotemporal Point Cloud Interaction Network for 3D Self-supervised Hand Pose Estimation
by: Han, Tianhao, et al.
Published: (2026)
by: Han, Tianhao, et al.
Published: (2026)
VideoMix: Aggregating How-To Videos for Task-Oriented Learning
by: Yang, Saelyne, et al.
Published: (2025)
by: Yang, Saelyne, et al.
Published: (2025)
Is Medieval Distant Viewing Possible? : Extending and Enriching Annotation of Legacy Image Collections using Visual Analytics
by: Meinecke, Christofer, et al.
Published: (2022)
by: Meinecke, Christofer, et al.
Published: (2022)
VideoA11y: Method and Dataset for Accessible Video Description
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
by: Sun, Qiushi, et al.
Published: (2025)
by: Sun, Qiushi, et al.
Published: (2025)
ChildCI Framework: Analysis of Motor and Cognitive Development in Children-Computer Interaction for Age Detection
by: Ruiz-Garcia, Juan Carlos, et al.
Published: (2022)
by: Ruiz-Garcia, Juan Carlos, et al.
Published: (2022)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
NarrativeBridge: Enhancing Video Captioning with Causal-Temporal Narrative
by: Nadeem, Asmar, et al.
Published: (2024)
by: Nadeem, Asmar, et al.
Published: (2024)
Reframe Anything: LLM Agent for Open World Video Reframing
by: Cao, Jiawang, et al.
Published: (2024)
by: Cao, Jiawang, et al.
Published: (2024)
Vid2Coach: Transforming How-To Videos into Task Assistants
by: Huh, Mina, et al.
Published: (2025)
by: Huh, Mina, et al.
Published: (2025)
Analyzing Swimming Performance Using Drone Captured Aerial Videos
by: Tran, Thu, et al.
Published: (2025)
by: Tran, Thu, et al.
Published: (2025)
Video Joint-Embedding Predictive Architectures for Facial Expression Recognition
by: Eing, Lennart, et al.
Published: (2026)
by: Eing, Lennart, et al.
Published: (2026)
VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents
by: Mazumdar, Amrita, et al.
Published: (2026)
by: Mazumdar, Amrita, et al.
Published: (2026)
ConceptFactory: Facilitate 3D Object Knowledge Annotation with Object Conceptualization
by: Sun, Jianhua, et al.
Published: (2024)
by: Sun, Jianhua, et al.
Published: (2024)
OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows
by: Sun, Qiushi, et al.
Published: (2025)
by: Sun, Qiushi, et al.
Published: (2025)
MVTN: A Multiscale Video Transformer Network for Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2024)
by: Garg, Mallika, et al.
Published: (2024)
Designing Multi-Robot Ground Video Sensemaking with Public Safety Professionals
by: Zhou, Puqi, et al.
Published: (2026)
by: Zhou, Puqi, et al.
Published: (2026)
Multiscaled Multi-Head Attention-based Video Transformer Network for Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2025)
by: Garg, Mallika, et al.
Published: (2025)
LLM4Brain: Training a Large Language Model for Brain Video Understanding
by: Zheng, Ruizhe, et al.
Published: (2024)
by: Zheng, Ruizhe, et al.
Published: (2024)
CinePreGen: Camera Controllable Video Previsualization via Engine-powered Diffusion
by: Chen, Yiran, et al.
Published: (2024)
by: Chen, Yiran, et al.
Published: (2024)
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
by: He, Yuchen, et al.
Published: (2025)
by: He, Yuchen, et al.
Published: (2025)
SpriteHand: Real-Time Versatile Hand-Object Interaction with Autoregressive Video Generation
by: Li, Zisu, et al.
Published: (2025)
by: Li, Zisu, et al.
Published: (2025)
Human-in-the-Loop Annotation for Image-Based Engagement Estimation: Assessing the Impact of Model Reliability on Annotation Accuracy
by: Subramanya, Sahana Yadnakudige, et al.
Published: (2025)
by: Subramanya, Sahana Yadnakudige, et al.
Published: (2025)
SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos
by: Huang, Xiyang, et al.
Published: (2026)
by: Huang, Xiyang, et al.
Published: (2026)
A Comparison of Bounding Box and Landmark Detection Methods for Video-Based Heart Rate Estimation
by: Liang, Laurence
Published: (2023)
by: Liang, Laurence
Published: (2023)
EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
by: Foteinopoulou, Niki Maria, et al.
Published: (2023)
Text-to-Image Representativity Fairness Evaluation Framework
by: Yamani, Asma, et al.
Published: (2024)
by: Yamani, Asma, et al.
Published: (2024)
EduGage: Methods and Dataset for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
by: Leng, Zikang, et al.
Published: (2026)
by: Leng, Zikang, et al.
Published: (2026)
The Visual Experience Dataset: Over 200 Recorded Hours of Integrated Eye Movement, Odometry, and Egocentric Video
by: Greene, Michelle R., et al.
Published: (2024)
by: Greene, Michelle R., et al.
Published: (2024)
Design and Evaluation of Camera-Centric Mobile Crowdsourcing Applications
by: Stylianou, Abby, et al.
Published: (2024)
by: Stylianou, Abby, et al.
Published: (2024)
Surgment: Segmentation-enabled Semantic Search and Creation of Visual Question and Feedback to Support Video-Based Surgery Learning
by: Wang, Jingying, et al.
Published: (2024)
by: Wang, Jingying, et al.
Published: (2024)
Real-Time Cellist Postural Evaluation With On-Device Computer Vision
by: Wang, Paolo, et al.
Published: (2026)
by: Wang, Paolo, et al.
Published: (2026)
Similar Items
-
Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos
by: Hamdi, Abdullah, et al.
Published: (2026) -
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025) -
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024) -
Learning Annotation Consensus for Continuous Emotion Recognition
by: Shoer, Ibrahim, et al.
Published: (2025) -
Confidence Contours: Uncertainty-Aware Annotation for Medical Semantic Segmentation
by: Ye, Andre, et al.
Published: (2023)