Leveraging Synthetic Data for Enhancing Egocentric Hand-Object Interaction Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Leonardi, Rosario, Furnari, Antonino, Ragusa, Francesco, Farinella, Giovanni Maria |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are Synthetic Data Useful for Egocentric Hand-Object Interaction Detection?
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
Exploiting Multimodal Synthetic Data for Egocentric Human-Object Interaction Detection in an Industrial Scenario
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
GlovEgo-HOI: Bridging the Synthetic-to-Real Gap for Industrial Egocentric Human-Object Interaction Detection
von: Spoto, Alfio, et al.
Veröffentlicht: (2026)
von: Spoto, Alfio, et al.
Veröffentlicht: (2026)
A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipation
von: Ragusa, Francesco, et al.
Veröffentlicht: (2023)
von: Ragusa, Francesco, et al.
Veröffentlicht: (2023)
Leveraging Gaze and Set-of-Mark in VLLMs for Human-Object Interaction Anticipation from Egocentric Videos
von: Materia, Daniele, et al.
Veröffentlicht: (2026)
von: Materia, Daniele, et al.
Veröffentlicht: (2026)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
von: Messina, Nicola, et al.
Veröffentlicht: (2025)
von: Messina, Nicola, et al.
Veröffentlicht: (2025)
EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation
von: Leonardi, Rosario, et al.
Veröffentlicht: (2026)
von: Leonardi, Rosario, et al.
Veröffentlicht: (2026)
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
Differentiable Task Graph Learning: Procedural Activity Representation and Online Mistake Detection from Egocentric Videos
von: Seminara, Luigi, et al.
Veröffentlicht: (2024)
von: Seminara, Luigi, et al.
Veröffentlicht: (2024)
Task Graph Maximum Likelihood Estimation for Procedural Activity Understanding in Egocentric Videos
von: Seminara, Luigi, et al.
Veröffentlicht: (2025)
von: Seminara, Luigi, et al.
Veröffentlicht: (2025)
Mamba-OTR: a Mamba-based Solution for Online Take and Release Detection from Untrimmed Egocentric Video
von: Catinello, Alessandro Sebastiano, et al.
Veröffentlicht: (2025)
von: Catinello, Alessandro Sebastiano, et al.
Veröffentlicht: (2025)
An Outlook into the Future of Egocentric Vision
von: Plizzari, Chiara, et al.
Veröffentlicht: (2023)
von: Plizzari, Chiara, et al.
Veröffentlicht: (2023)
Ego-EXTRA: video-language Egocentric Dataset for EXpert-TRAinee assistance
von: Ragusa, Francesco, et al.
Veröffentlicht: (2025)
von: Ragusa, Francesco, et al.
Veröffentlicht: (2025)
Calisthenics Skills Temporal Video Segmentation
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
Efficient Calisthenics Skills Classification through Foreground Instance Selection and Depth Estimation
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
How Far Can Off-the-Shelf Multimodal Large Language Models Go in Online Episodic Memory Question Answering?
von: Lando, Giuseppe, et al.
Veröffentlicht: (2025)
von: Lando, Giuseppe, et al.
Veröffentlicht: (2025)
Online Episodic Memory Visual Query Localization with Egocentric Streaming Object Memory
von: Manigrasso, Zaira, et al.
Veröffentlicht: (2024)
von: Manigrasso, Zaira, et al.
Veröffentlicht: (2024)
AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
ENIGMA-360: An Ego-Exo Dataset for Human Behavior Understanding in Industrial Scenarios
von: Ragusa, Francesco, et al.
Veröffentlicht: (2026)
von: Ragusa, Francesco, et al.
Veröffentlicht: (2026)
Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2026)
von: Santos-Villafranca, Maria, et al.
Veröffentlicht: (2026)
Synchronization is All You Need: Exocentric-to-Egocentric Transfer for Temporal Action Segmentation with Unlabeled Synchronized Video Pairs
von: Quattrocchi, Camillo, et al.
Veröffentlicht: (2023)
von: Quattrocchi, Camillo, et al.
Veröffentlicht: (2023)
Integrating Affordances and Attention models for Short-Term Object Interaction Anticipation
von: Labadia, Lorenzo Mur, et al.
Veröffentlicht: (2026)
von: Labadia, Lorenzo Mur, et al.
Veröffentlicht: (2026)
EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision
von: Forte, Rosario, et al.
Veröffentlicht: (2026)
von: Forte, Rosario, et al.
Veröffentlicht: (2026)
EASG-Bench: Video Q&A Benchmark with Egocentric Action Scene Graphs
von: Rodin, Ivan, et al.
Veröffentlicht: (2025)
von: Rodin, Ivan, et al.
Veröffentlicht: (2025)
SignIT: A Comprehensive Dataset and Multimodal Analysis for Italian Sign Language Recognition
von: Micieli, Alessia, et al.
Veröffentlicht: (2025)
von: Micieli, Alessia, et al.
Veröffentlicht: (2025)
Semantically Guided Action Anticipation
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
von: Diko, Anxhelo, et al.
Veröffentlicht: (2024)
ProSkill: Segment-Level Skill Assessment in Procedural Videos
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2026)
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2026)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
Exploring Multimodal LMMs for Online Episodic Memory Question Answering on the Edge
von: Lando, Giuseppe, et al.
Veröffentlicht: (2026)
von: Lando, Giuseppe, et al.
Veröffentlicht: (2026)
ZARRIO @ Ego4D Short Term Object Interaction Anticipation Challenge: Leveraging Affordances and Attention-based models for STA
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2024)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026)
von: Li, Dayou, et al.
Veröffentlicht: (2026)
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
von: Plini, Leonardo, et al.
Veröffentlicht: (2024)
von: Plini, Leonardo, et al.
Veröffentlicht: (2024)
Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
PREGO: online mistake detection in PRocedural EGOcentric videos
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2024)
von: Flaborea, Alessandro, et al.
Veröffentlicht: (2024)
RECIPE: Procedural Planning via Grounding in Instructional Video
von: Seminara, Luigi, et al.
Veröffentlicht: (2026)
von: Seminara, Luigi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Are Synthetic Data Useful for Egocentric Hand-Object Interaction Detection?
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023) -
Exploiting Multimodal Synthetic Data for Egocentric Human-Object Interaction Detection in an Industrial Scenario
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023) -
GlovEgo-HOI: Bridging the Synthetic-to-Real Gap for Industrial Egocentric Human-Object Interaction Detection
von: Spoto, Alfio, et al.
Veröffentlicht: (2026) -
A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025) -
StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipation
von: Ragusa, Francesco, et al.
Veröffentlicht: (2023)