Learning reusable concepts across different egocentric video understanding tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peirone, Simone Alberto, Pistilli, Francesca, Alliegro, Antonio, Tommasi, Tatiana, Averta, Giuseppe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025)
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
von: Alliegro, Antonio, et al.
Veröffentlicht: (2025)
von: Alliegro, Antonio, et al.
Veröffentlicht: (2025)
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
von: Zenotto, Andrea, et al.
Veröffentlicht: (2026)
von: Zenotto, Andrea, et al.
Veröffentlicht: (2026)
Egocentric zone-aware action recognition across environments
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
Domain Generalization using Action Sequences for Egocentric Action Recognition
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
von: Nasirimajd, Amirshayan, et al.
Veröffentlicht: (2025)
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2024)
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2024)
AMEGO: Active Memory from long EGOcentric videos
von: Goletto, Gabriele, et al.
Veröffentlicht: (2024)
von: Goletto, Gabriele, et al.
Veröffentlicht: (2024)
Transient Fault Tolerant Semantic Segmentation for Autonomous Driving
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
The BabyView dataset: High-resolution egocentric videos of infants' and young children's everyday experiences
von: Long, Bria, et al.
Veröffentlicht: (2024)
von: Long, Bria, et al.
Veröffentlicht: (2024)
A Modern Take on Visual Relationship Reasoning for Grasp Planning
von: Rabino, Paolo, et al.
Veröffentlicht: (2024)
von: Rabino, Paolo, et al.
Veröffentlicht: (2024)
MaskPlanner: Learning-Based Object-Centric Motion Generation from 3D Point Clouds
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2025)
von: Tiboni, Gabriele, et al.
Veröffentlicht: (2025)
SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation
von: Cuttano, Claudia, et al.
Veröffentlicht: (2025)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2025)
What does CLIP know about peeling a banana?
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
Efficient Odd-One-Out Anomaly Detection
von: Chito, Silvio, et al.
Veröffentlicht: (2025)
von: Chito, Silvio, et al.
Veröffentlicht: (2025)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
von: Cuttano, Claudia, et al.
Veröffentlicht: (2024)
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2024)
von: Rosi, Gabriele, et al.
Veröffentlicht: (2024)
A Second-Order Perspective on Pruning at Initialization and Knowledge Transfer
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
Finding Lottery Tickets in Vision Models via Data-driven Spectral Foresight Pruning
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2024)
Fixed External Cameras as Common Prior Maps for Active 3D Scene Graph Generation
von: Modi, Giorgia, et al.
Veröffentlicht: (2026)
von: Modi, Giorgia, et al.
Veröffentlicht: (2026)
RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots
von: Modi, Giorgia, et al.
Veröffentlicht: (2026)
von: Modi, Giorgia, et al.
Veröffentlicht: (2026)
MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware
von: Palanisamy, Senthil, et al.
Veröffentlicht: (2026)
von: Palanisamy, Senthil, et al.
Veröffentlicht: (2026)
Efficient Model Editing with Task-Localized Sparse Fine-tuning
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
von: Iurada, Leonardo, et al.
Veröffentlicht: (2025)
HUP-3D: A 3D multi-view synthetic dataset for assisted-egocentric hand-ultrasound pose estimation
von: Birlo, Manuel, et al.
Veröffentlicht: (2024)
von: Birlo, Manuel, et al.
Veröffentlicht: (2024)
MultiGraspNet: A Multitask 3D Vision Model for Multi-gripper Robotic Grasping
von: Ortuno-Chanelo, Stephany, et al.
Veröffentlicht: (2026)
von: Ortuno-Chanelo, Stephany, et al.
Veröffentlicht: (2026)
A generalizable foundation model for intraoperative understanding across surgical procedures
von: Park, Kanggil, et al.
Veröffentlicht: (2026)
von: Park, Kanggil, et al.
Veröffentlicht: (2026)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
Open and reusable deep learning for pathology with WSInfer and QuPath
von: Kaczmarzyk, Jakub R., et al.
Veröffentlicht: (2023)
von: Kaczmarzyk, Jakub R., et al.
Veröffentlicht: (2023)
AI-driven visual monitoring of industrial assembly tasks
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
DLM-VMTL:A Double Layer Mapper for heterogeneous data video Multi-task prompt learning
von: Bo, Zeyi, et al.
Veröffentlicht: (2024)
von: Bo, Zeyi, et al.
Veröffentlicht: (2024)
ViSTa Dataset: Do vision-language models understand sequential tasks?
von: Wybitul, Evžen, et al.
Veröffentlicht: (2024)
von: Wybitul, Evžen, et al.
Veröffentlicht: (2024)
Intuitive physics understanding emerges from self-supervised pretraining on natural videos
von: Garrido, Quentin, et al.
Veröffentlicht: (2025)
von: Garrido, Quentin, et al.
Veröffentlicht: (2025)
Did you just see that? Arbitrary view synthesis for egocentric replay of operating room workflows from ambient sensors
von: Zhang, Han, et al.
Veröffentlicht: (2025)
von: Zhang, Han, et al.
Veröffentlicht: (2025)
Benchmarking transferability of SSL pretraining to same and different modality segmentation tasks
von: Jiang, Jue, et al.
Veröffentlicht: (2026)
von: Jiang, Jue, et al.
Veröffentlicht: (2026)
An Outlook into the Future of Egocentric Vision
von: Plizzari, Chiara, et al.
Veröffentlicht: (2023)
von: Plizzari, Chiara, et al.
Veröffentlicht: (2023)
Body Segmentation Using Multi-task Learning
von: Jug, Julijan, et al.
Veröffentlicht: (2022)
von: Jug, Julijan, et al.
Veröffentlicht: (2022)
Multi-step manipulation task and motion planning guided by video demonstration
von: Zorina, Kateryna, et al.
Veröffentlicht: (2025)
von: Zorina, Kateryna, et al.
Veröffentlicht: (2025)
Multi-task Learning For Joint Action and Gesture Recognition
von: Spathis, Konstantinos, et al.
Veröffentlicht: (2025)
von: Spathis, Konstantinos, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025) -
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2025) -
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
von: Alliegro, Antonio, et al.
Veröffentlicht: (2025) -
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024) -
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
von: Zenotto, Andrea, et al.
Veröffentlicht: (2026)