Technical Report for Egocentric Mistake Detection for the HoloAssist Challenge
Fuente:
arXiv
Saved in:
| Main Authors: | Patsch, Constantin, Zakour, Marsil, Wu, Yuankai, Steinbach, Eckehard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ADL4D: Towards A Contextually Rich Dataset for 4D Activities of Daily Living
by: Zakour, Marsil, et al.
Published: (2024)
by: Zakour, Marsil, et al.
Published: (2024)
REVNET: Rotation-Equivariant Point Cloud Completion via Vector Neuron Anchor Transformer
by: Ni, Zhifan, et al.
Published: (2026)
by: Ni, Zhifan, et al.
Published: (2026)
Watch and Learn: Leveraging Expert Knowledge and Language for Surgical Video Understanding
by: Gastager, David, et al.
Published: (2025)
by: Gastager, David, et al.
Published: (2025)
DinoComplete: 3D Shape Completion with Distilled Semantic Priors and State Space Models
by: Algan, Furkan Mert, et al.
Published: (2026)
by: Algan, Furkan Mert, et al.
Published: (2026)
BIR-Adapter: A parameter-efficient diffusion adapter for blind image restoration
by: Eteke, Cem, et al.
Published: (2025)
by: Eteke, Cem, et al.
Published: (2025)
How to Correctly Make Mistakes: A Framework for Constructing and Benchmarking Mistake Aware Egocentric Procedural Videos
by: Loginova, Olga, et al.
Published: (2026)
by: Loginova, Olga, et al.
Published: (2026)
HOOD: Real-Time Human Presence and Out-of-Distribution Detection Using FMCW Radar
by: Kahya, Sabri Mustafa, et al.
Published: (2023)
by: Kahya, Sabri Mustafa, et al.
Published: (2023)
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
by: Mazzamuto, Michele, et al.
Published: (2024)
by: Mazzamuto, Michele, et al.
Published: (2024)
Differentiable Task Graph Learning: Procedural Activity Representation and Online Mistake Detection from Egocentric Videos
by: Seminara, Luigi, et al.
Published: (2024)
by: Seminara, Luigi, et al.
Published: (2024)
Understanding-Enhanced Model Collaboration for Long-Tailed Egocentric Mistake Detection
by: Han, Boyu, et al.
Published: (2026)
by: Han, Boyu, et al.
Published: (2026)
LEMON: Localized Editing with Mesh Optimization and Neural Shaders
by: Algan, Furkan Mert, et al.
Published: (2024)
by: Algan, Furkan Mert, et al.
Published: (2024)
Dual-Stage Reweighted MoE for Long-Tailed Egocentric Mistake Detection
by: Han, Boyu, et al.
Published: (2025)
by: Han, Boyu, et al.
Published: (2025)
A Causal Diffusion Model for Video Reconstruction from Ultra-Low-Bitrate Representations
by: Eteke, Cem, et al.
Published: (2026)
by: Eteke, Cem, et al.
Published: (2026)
FERT: Real-Time Facial Expression Recognition with Short-Range FMCW Radar
by: Kahya, Sabri Mustafa, et al.
Published: (2024)
by: Kahya, Sabri Mustafa, et al.
Published: (2024)
FOOD: Facial Authentication and Out-of-Distribution Detection with Short-Range FMCW Radar
by: Kahya, Sabri Mustafa, et al.
Published: (2024)
by: Kahya, Sabri Mustafa, et al.
Published: (2024)
Outside Knowledge Conversational Video (OKCV) Dataset -- Dialoguing over Videos
by: Reichman, Benjamin, et al.
Published: (2025)
by: Reichman, Benjamin, et al.
Published: (2025)
EgoOops: A Dataset for Mistake Action Detection from Egocentric Videos referring to Procedural Texts
by: Haneji, Yuto, et al.
Published: (2024)
by: Haneji, Yuto, et al.
Published: (2024)
RadarCNN: Learning-based Indoor Object Classification from IQ Imaging Radar Data
by: Hägele, Stefan, et al.
Published: (2026)
by: Hägele, Stefan, et al.
Published: (2026)
Mistake Attribution: Fine-Grained Mistake Understanding in Egocentric Videos
by: Li, Yayuan, et al.
Published: (2025)
by: Li, Yayuan, et al.
Published: (2025)
Procedural Mistake Detection via Action Effect Modeling
by: Guo, Wenliang, et al.
Published: (2025)
by: Guo, Wenliang, et al.
Published: (2025)
HoloGS: Instant Depth-based 3D Gaussian Splatting with Microsoft HoloLens 2
by: Jäger, Miriam, et al.
Published: (2024)
by: Jäger, Miriam, et al.
Published: (2024)
FARE: A Deep Learning-Based Framework for Radar-based Face Recognition and Out-of-distribution Detection
by: Kahya, Sabri Mustafa, et al.
Published: (2025)
by: Kahya, Sabri Mustafa, et al.
Published: (2025)
Vision-Based Mistake Analysis in Procedural Activities: A Review of Advances and Challenges
by: Bacharidis, Konstantinos, et al.
Published: (2025)
by: Bacharidis, Konstantinos, et al.
Published: (2025)
Building Egocentric Procedural AI Assistant: Methods, Benchmarks, and Challenges
by: Li, Junlong, et al.
Published: (2025)
by: Li, Junlong, et al.
Published: (2025)
MistExit: Learning to Exit for Early Mistake Detection in Procedural Videos
by: Majumder, Sagnik, et al.
Published: (2026)
by: Majumder, Sagnik, et al.
Published: (2026)
Efficient Contextformer: Spatio-Channel Window Attention for Fast Context Modeling in Learned Image Compression
by: Koyuncu, A. Burakhan, et al.
Published: (2023)
by: Koyuncu, A. Burakhan, et al.
Published: (2023)
Multi-Modal UAV Detection, Classification and Tracking Algorithm -- Technical Report for CVPR 2024 UG2 Challenge
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
by: Chen, Shimin, et al.
Published: (2024)
by: Chen, Shimin, et al.
Published: (2024)
Technical Report for SoccerNet Challenge 2022 -- Replay Grounding Task
by: Chen, Shimin, et al.
Published: (2024)
by: Chen, Shimin, et al.
Published: (2024)
HoloPart: Generative 3D Part Amodal Segmentation
by: Yang, Yunhan, et al.
Published: (2025)
by: Yang, Yunhan, et al.
Published: (2025)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
by: Fan, Zicong, et al.
Published: (2024)
by: Fan, Zicong, et al.
Published: (2024)
HOLa: HoloLens Object Labeling
by: Schwimmbeck, Michael, et al.
Published: (2024)
by: Schwimmbeck, Michael, et al.
Published: (2024)
Scene Text Detection and Recognition "in light of" Challenging Environmental Conditions using Aria Glasses Egocentric Vision Cameras
by: De Mathia, Joseph, et al.
Published: (2025)
by: De Mathia, Joseph, et al.
Published: (2025)
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
by: Plini, Leonardo, et al.
Published: (2024)
by: Plini, Leonardo, et al.
Published: (2024)
FOODER: Real-time Facial Authentication and Expression Recognition
by: Kahya, Sabri Mustafa, et al.
Published: (2025)
by: Kahya, Sabri Mustafa, et al.
Published: (2025)
HoloDx: Knowledge- and Data-Driven Multimodal Diagnosis of Alzheimer's Disease
by: Chen, Qiuhui, et al.
Published: (2025)
by: Chen, Qiuhui, et al.
Published: (2025)
Adapting Learned Image Codecs to Screen Content via Adjustable Transformations
by: Dogaroglu, H. Burak, et al.
Published: (2024)
by: Dogaroglu, H. Burak, et al.
Published: (2024)
Logics-Parsing Technical Report
by: Chen, Xiangyang, et al.
Published: (2025)
by: Chen, Xiangyang, et al.
Published: (2025)
Qwen-Image Technical Report
by: Wu, Chenfei, et al.
Published: (2025)
by: Wu, Chenfei, et al.
Published: (2025)
HoloHisto: End-to-end Gigapixel WSI Segmentation with 4K Resolution Sequential Tokenization
by: Tang, Yucheng, et al.
Published: (2024)
by: Tang, Yucheng, et al.
Published: (2024)
Similar Items
-
ADL4D: Towards A Contextually Rich Dataset for 4D Activities of Daily Living
by: Zakour, Marsil, et al.
Published: (2024) -
REVNET: Rotation-Equivariant Point Cloud Completion via Vector Neuron Anchor Transformer
by: Ni, Zhifan, et al.
Published: (2026) -
Watch and Learn: Leveraging Expert Knowledge and Language for Surgical Video Understanding
by: Gastager, David, et al.
Published: (2025) -
DinoComplete: 3D Shape Completion with Distilled Semantic Priors and State Space Models
by: Algan, Furkan Mert, et al.
Published: (2026) -
BIR-Adapter: A parameter-efficient diffusion adapter for blind image restoration
by: Eteke, Cem, et al.
Published: (2025)