Guardado en:
| Autores principales: | Peirone, Simone Alberto, Pistilli, Francesca, Alliegro, Antonio, Averta, Giuseppe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.03037 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
Learning reusable concepts across different egocentric video understanding tasks
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
por: Alliegro, Antonio, et al.
Publicado: (2025)
por: Alliegro, Antonio, et al.
Publicado: (2025)
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
por: Zenotto, Andrea, et al.
Publicado: (2026)
por: Zenotto, Andrea, et al.
Publicado: (2026)
Egocentric zone-aware action recognition across environments
por: Peirone, Simone Alberto, et al.
Publicado: (2024)
por: Peirone, Simone Alberto, et al.
Publicado: (2024)
Domain Generalization using Action Sequences for Egocentric Action Recognition
por: Nasirimajd, Amirshayan, et al.
Publicado: (2025)
por: Nasirimajd, Amirshayan, et al.
Publicado: (2025)
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding
por: Nagrani, Arsha, et al.
Publicado: (2026)
por: Nagrani, Arsha, et al.
Publicado: (2026)
Ego-VPA: Egocentric Video Understanding with Parameter-efficient Adaptation
por: Wu, Tz-Ying, et al.
Publicado: (2024)
por: Wu, Tz-Ying, et al.
Publicado: (2024)
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
FEEL (Force-Enhanced Egocentric Learning): A Dataset for Physical Action Understanding
por: Dessalene, Eadom, et al.
Publicado: (2026)
por: Dessalene, Eadom, et al.
Publicado: (2026)
LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos
por: Wang, Ying, et al.
Publicado: (2023)
por: Wang, Ying, et al.
Publicado: (2023)
Ego4OOD: Rethinking Egocentric Video Domain Generalization via Covariate Shift Scoring
por: Vaseqi, Zahra, et al.
Publicado: (2026)
por: Vaseqi, Zahra, et al.
Publicado: (2026)
Memory Storyboard: Leveraging Temporal Segmentation for Streaming Self-Supervised Learning from Egocentric Videos
por: Yang, Yanlai, et al.
Publicado: (2025)
por: Yang, Yanlai, et al.
Publicado: (2025)
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
por: Hoque, Ryan, et al.
Publicado: (2025)
por: Hoque, Ryan, et al.
Publicado: (2025)
Gradient Similarity Surgery in Multi-Task Deep Learning
por: Borsani, Thomas, et al.
Publicado: (2025)
por: Borsani, Thomas, et al.
Publicado: (2025)
$\infty$-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation
por: Santos, Saul, et al.
Publicado: (2025)
por: Santos, Saul, et al.
Publicado: (2025)
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
por: Bhalgat, Yash, et al.
Publicado: (2024)
por: Bhalgat, Yash, et al.
Publicado: (2024)
Being-H0.7: A Latent World-Action Model from Egocentric Videos
por: Luo, Hao, et al.
Publicado: (2026)
por: Luo, Hao, et al.
Publicado: (2026)
Understanding-Enhanced Model Collaboration for Long-Tailed Egocentric Mistake Detection
por: Han, Boyu, et al.
Publicado: (2026)
por: Han, Boyu, et al.
Publicado: (2026)
ECBench: Can Multi-modal Foundation Models Understand the Egocentric World? A Holistic Embodied Cognition Benchmark
por: Dang, Ronghao, et al.
Publicado: (2025)
por: Dang, Ronghao, et al.
Publicado: (2025)
MM-Ego: Towards Building Egocentric Multimodal LLMs for Video QA
por: Ye, Hanrong, et al.
Publicado: (2024)
por: Ye, Hanrong, et al.
Publicado: (2024)
What to Do Next? Memorizing skills from Egocentric Instructional Video
por: Bi, Jing, et al.
Publicado: (2025)
por: Bi, Jing, et al.
Publicado: (2025)
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
por: Liang, Zhixuan, et al.
Publicado: (2023)
por: Liang, Zhixuan, et al.
Publicado: (2023)
Whole-Body Conditioned Egocentric Video Prediction
por: Bai, Yutong, et al.
Publicado: (2025)
por: Bai, Yutong, et al.
Publicado: (2025)
Understanding Domain Generalization: A Noise Robustness Perspective
por: Qiao, Rui, et al.
Publicado: (2024)
por: Qiao, Rui, et al.
Publicado: (2024)
EgoMAGIC- An Egocentric Video Field Medicine Dataset for Training Perception Algorithms
por: VanVoorst, Brian, et al.
Publicado: (2026)
por: VanVoorst, Brian, et al.
Publicado: (2026)
Online Video Understanding: OVBench and VideoChat-Online
por: Huang, Zhenpeng, et al.
Publicado: (2024)
por: Huang, Zhenpeng, et al.
Publicado: (2024)
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2024)
por: Rosi, Gabriele, et al.
Publicado: (2024)
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs
por: Chung, Hyungjin, et al.
Publicado: (2025)
por: Chung, Hyungjin, et al.
Publicado: (2025)
Understanding Task Transfer in Vision-Language Models
por: Sachdeva, Bhuvan, et al.
Publicado: (2025)
por: Sachdeva, Bhuvan, et al.
Publicado: (2025)
PEDESTRIAN: An Egocentric Vision Dataset for Obstacle Detection on Pavements
por: Thoma, Marios, et al.
Publicado: (2025)
por: Thoma, Marios, et al.
Publicado: (2025)
EgoSurgery-HTS: A Dataset for Egocentric Hand-Tool Segmentation in Open Surgery Videos
por: Darjana, Nathan, et al.
Publicado: (2025)
por: Darjana, Nathan, et al.
Publicado: (2025)
EasyVideoR1: Easier RL for Video Understanding
por: Qin, Chuanyu, et al.
Publicado: (2026)
por: Qin, Chuanyu, et al.
Publicado: (2026)
Towards Sparse Video Understanding and Reasoning
por: Xu, Chenwei, et al.
Publicado: (2026)
por: Xu, Chenwei, et al.
Publicado: (2026)
Agentic Very Long Video Understanding
por: Rege, Aniket, et al.
Publicado: (2026)
por: Rege, Aniket, et al.
Publicado: (2026)
EgoCogNav: Cognition-aware Human Egocentric Navigation
por: Qiu, Zhiwen, et al.
Publicado: (2025)
por: Qiu, Zhiwen, et al.
Publicado: (2025)
Benchmarking Egocentric Multimodal Goal Inference for Assistive Wearable Agents
por: Veerabadran, Vijay, et al.
Publicado: (2025)
por: Veerabadran, Vijay, et al.
Publicado: (2025)
EgoSurgery-Phase: A Dataset of Surgical Phase Recognition from Egocentric Open Surgery Videos
por: Fujii, Ryo, et al.
Publicado: (2024)
por: Fujii, Ryo, et al.
Publicado: (2024)
Advancing Egocentric Video Question Answering with Multimodal Large Language Models
por: Patel, Alkesh, et al.
Publicado: (2025)
por: Patel, Alkesh, et al.
Publicado: (2025)
Ejemplares similares
-
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2025) -
Learning reusable concepts across different egocentric video understanding tasks
por: Peirone, Simone Alberto, et al.
Publicado: (2025) -
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
por: Peirone, Simone Alberto, et al.
Publicado: (2025) -
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
por: Alliegro, Antonio, et al.
Publicado: (2025) -
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
por: Zenotto, Andrea, et al.
Publicado: (2026)