Efficient Calisthenics Skills Classification through Foreground Instance Selection and Depth Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Finocchiaro, Antonio, Farinella, Giovanni Maria, Furnari, Antonino |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Calisthenics Skills Temporal Video Segmentation
por: Finocchiaro, Antonio, et al.
Publicado: (2025)
por: Finocchiaro, Antonio, et al.
Publicado: (2025)
Task Graph Maximum Likelihood Estimation for Procedural Activity Understanding in Egocentric Videos
por: Seminara, Luigi, et al.
Publicado: (2025)
por: Seminara, Luigi, et al.
Publicado: (2025)
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
por: Mazzamuto, Michele, et al.
Publicado: (2024)
por: Mazzamuto, Michele, et al.
Publicado: (2024)
Differentiable Task Graph Learning: Procedural Activity Representation and Online Mistake Detection from Egocentric Videos
por: Seminara, Luigi, et al.
Publicado: (2024)
por: Seminara, Luigi, et al.
Publicado: (2024)
StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipation
por: Ragusa, Francesco, et al.
Publicado: (2023)
por: Ragusa, Francesco, et al.
Publicado: (2023)
ProSkill: Segment-Level Skill Assessment in Procedural Videos
por: Mazzamuto, Michele, et al.
Publicado: (2026)
por: Mazzamuto, Michele, et al.
Publicado: (2026)
A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains
por: Finocchiaro, Antonio, et al.
Publicado: (2025)
por: Finocchiaro, Antonio, et al.
Publicado: (2025)
Mamba-OTR: a Mamba-based Solution for Online Take and Release Detection from Untrimmed Egocentric Video
por: Catinello, Alessandro Sebastiano, et al.
Publicado: (2025)
por: Catinello, Alessandro Sebastiano, et al.
Publicado: (2025)
How Far Can Off-the-Shelf Multimodal Large Language Models Go in Online Episodic Memory Question Answering?
por: Lando, Giuseppe, et al.
Publicado: (2025)
por: Lando, Giuseppe, et al.
Publicado: (2025)
Leveraging Synthetic Data for Enhancing Egocentric Hand-Object Interaction Detection
por: Leonardi, Rosario, et al.
Publicado: (2026)
por: Leonardi, Rosario, et al.
Publicado: (2026)
Exploiting Multimodal Synthetic Data for Egocentric Human-Object Interaction Detection in an Industrial Scenario
por: Leonardi, Rosario, et al.
Publicado: (2023)
por: Leonardi, Rosario, et al.
Publicado: (2023)
Are Synthetic Data Useful for Egocentric Hand-Object Interaction Detection?
por: Leonardi, Rosario, et al.
Publicado: (2023)
por: Leonardi, Rosario, et al.
Publicado: (2023)
Semantically Guided Action Anticipation
por: Diko, Anxhelo, et al.
Publicado: (2024)
por: Diko, Anxhelo, et al.
Publicado: (2024)
Online Episodic Memory Visual Query Localization with Egocentric Streaming Object Memory
por: Manigrasso, Zaira, et al.
Publicado: (2024)
por: Manigrasso, Zaira, et al.
Publicado: (2024)
Synchronization is All You Need: Exocentric-to-Egocentric Transfer for Temporal Action Segmentation with Unlabeled Synchronized Video Pairs
por: Quattrocchi, Camillo, et al.
Publicado: (2023)
por: Quattrocchi, Camillo, et al.
Publicado: (2023)
AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation
por: Mur-Labadia, Lorenzo, et al.
Publicado: (2024)
por: Mur-Labadia, Lorenzo, et al.
Publicado: (2024)
Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark
por: Santos-Villafranca, Maria, et al.
Publicado: (2026)
por: Santos-Villafranca, Maria, et al.
Publicado: (2026)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
por: Messina, Nicola, et al.
Publicado: (2025)
por: Messina, Nicola, et al.
Publicado: (2025)
Integrating Affordances and Attention models for Short-Term Object Interaction Anticipation
por: Labadia, Lorenzo Mur, et al.
Publicado: (2026)
por: Labadia, Lorenzo Mur, et al.
Publicado: (2026)
EASG-Bench: Video Q&A Benchmark with Egocentric Action Scene Graphs
por: Rodin, Ivan, et al.
Publicado: (2025)
por: Rodin, Ivan, et al.
Publicado: (2025)
An Outlook into the Future of Egocentric Vision
por: Plizzari, Chiara, et al.
Publicado: (2023)
por: Plizzari, Chiara, et al.
Publicado: (2023)
Ego-EXTRA: video-language Egocentric Dataset for EXpert-TRAinee assistance
por: Ragusa, Francesco, et al.
Publicado: (2025)
por: Ragusa, Francesco, et al.
Publicado: (2025)
ENIGMA-360: An Ego-Exo Dataset for Human Behavior Understanding in Industrial Scenarios
por: Ragusa, Francesco, et al.
Publicado: (2026)
por: Ragusa, Francesco, et al.
Publicado: (2026)
PREGO: online mistake detection in PRocedural EGOcentric videos
por: Flaborea, Alessandro, et al.
Publicado: (2024)
por: Flaborea, Alessandro, et al.
Publicado: (2024)
RECIPE: Procedural Planning via Grounding in Instructional Video
por: Seminara, Luigi, et al.
Publicado: (2026)
por: Seminara, Luigi, et al.
Publicado: (2026)
Exploring Multimodal LMMs for Online Episodic Memory Question Answering on the Edge
por: Lando, Giuseppe, et al.
Publicado: (2026)
por: Lando, Giuseppe, et al.
Publicado: (2026)
EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision
por: Forte, Rosario, et al.
Publicado: (2026)
por: Forte, Rosario, et al.
Publicado: (2026)
ViterbiPlanNet: Injecting Procedural Knowledge via Differentiable Viterbi for Planning in Instructional Videos
por: Seminara, Luigi, et al.
Publicado: (2026)
por: Seminara, Luigi, et al.
Publicado: (2026)
TI-PREGO: Chain of Thought and In-Context Learning for Online Mistake Detection in PRocedural EGOcentric Videos
por: Plini, Leonardo, et al.
Publicado: (2024)
por: Plini, Leonardo, et al.
Publicado: (2024)
SignIT: A Comprehensive Dataset and Multimodal Analysis for Italian Sign Language Recognition
por: Micieli, Alessia, et al.
Publicado: (2025)
por: Micieli, Alessia, et al.
Publicado: (2025)
Leveraging Gaze and Set-of-Mark in VLLMs for Human-Object Interaction Anticipation from Egocentric Videos
por: Materia, Daniele, et al.
Publicado: (2026)
por: Materia, Daniele, et al.
Publicado: (2026)
Rethinking Transparent Object Grasping: Depth Completion with Monocular Depth Estimation and Instance Mask
por: Cheng, Yaofeng, et al.
Publicado: (2025)
por: Cheng, Yaofeng, et al.
Publicado: (2025)
GlovEgo-HOI: Bridging the Synthetic-to-Real Gap for Industrial Egocentric Human-Object Interaction Detection
por: Spoto, Alfio, et al.
Publicado: (2026)
por: Spoto, Alfio, et al.
Publicado: (2026)
Background Fades, Foreground Leads: Curriculum-Guided Background Pruning for Efficient Foreground-Centric Collaborative Perception
por: Wu, Yuheng, et al.
Publicado: (2025)
por: Wu, Yuheng, et al.
Publicado: (2025)
ZARRIO @ Ego4D Short Term Object Interaction Anticipation Challenge: Leveraging Affordances and Attention-based models for STA
por: Mur-Labadia, Lorenzo, et al.
Publicado: (2024)
por: Mur-Labadia, Lorenzo, et al.
Publicado: (2024)
MT-Depth: Multi-task Instance feature analysis for the Depth Completion
por: Nizamani, Abdul Haseeb, et al.
Publicado: (2025)
por: Nizamani, Abdul Haseeb, et al.
Publicado: (2025)
Depth-Guided Semi-Supervised Instance Segmentation
por: Chen, Xin, et al.
Publicado: (2024)
por: Chen, Xin, et al.
Publicado: (2024)
RoCo-Sim: Enhancing Roadside Collaborative Perception through Foreground Simulation
por: Du, Yuwen, et al.
Publicado: (2025)
por: Du, Yuwen, et al.
Publicado: (2025)
SemanticStitch: Enhancing Image Coherence through Foreground-Aware Seam Carving
por: Jin, Ji-Ping, et al.
Publicado: (2025)
por: Jin, Ji-Ping, et al.
Publicado: (2025)
Instance-Guided Radar Depth Estimation for 3D Object Detection
por: Lo, Chen-Chou, et al.
Publicado: (2026)
por: Lo, Chen-Chou, et al.
Publicado: (2026)
Ejemplares similares
-
Calisthenics Skills Temporal Video Segmentation
por: Finocchiaro, Antonio, et al.
Publicado: (2025) -
Task Graph Maximum Likelihood Estimation for Procedural Activity Understanding in Egocentric Videos
por: Seminara, Luigi, et al.
Publicado: (2025) -
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
por: Mazzamuto, Michele, et al.
Publicado: (2024) -
Differentiable Task Graph Learning: Procedural Activity Representation and Online Mistake Detection from Egocentric Videos
por: Seminara, Luigi, et al.
Publicado: (2024) -
StillFast: An End-to-End Approach for Short-Term Object Interaction Anticipation
por: Ragusa, Francesco, et al.
Publicado: (2023)