Egocentric zone-aware action recognition across environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Peirone, Simone Alberto, Goletto, Gabriele, Planamente, Mirco, Bottino, Andrea, Caputo, Barbara, Averta, Giuseppe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Domain Generalization using Action Sequences for Egocentric Action Recognition
por: Nasirimajd, Amirshayan, et al.
Publicado: (2025)
por: Nasirimajd, Amirshayan, et al.
Publicado: (2025)
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2024)
por: Peirone, Simone Alberto, et al.
Publicado: (2024)
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
por: Zenotto, Andrea, et al.
Publicado: (2026)
por: Zenotto, Andrea, et al.
Publicado: (2026)
AMEGO: Active Memory from long EGOcentric videos
por: Goletto, Gabriele, et al.
Publicado: (2024)
por: Goletto, Gabriele, et al.
Publicado: (2024)
Learning reusable concepts across different egocentric video understanding tasks
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
por: Peirone, Simone Alberto, et al.
Publicado: (2025)
FS-SAM2: Adapting Segment Anything Model 2 for Few-Shot Semantic Segmentation via Low-Rank Adaptation
por: Forni, Bernardo, et al.
Publicado: (2025)
por: Forni, Bernardo, et al.
Publicado: (2025)
EarthMatch: Iterative Coregistration for Fine-grained Localization of Astronaut Photography
por: Berton, Gabriele, et al.
Publicado: (2024)
por: Berton, Gabriele, et al.
Publicado: (2024)
Cross-Domain Transfer Learning with CoRTe: Consistent and Reliable Transfer from Black-Box to Lightweight Segmentation Model
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
An Outlook into the Future of Egocentric Vision
por: Plizzari, Chiara, et al.
Publicado: (2023)
por: Plizzari, Chiara, et al.
Publicado: (2023)
What does CLIP know about peeling a banana?
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation
por: Cuttano, Claudia, et al.
Publicado: (2025)
por: Cuttano, Claudia, et al.
Publicado: (2025)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
The revenge of BiSeNet: Efficient Multi-Task Image Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2024)
por: Rosi, Gabriele, et al.
Publicado: (2024)
JIST: Joint Image and Sequence Training for Sequential Visual Place Recognition
por: Berton, Gabriele, et al.
Publicado: (2024)
por: Berton, Gabriele, et al.
Publicado: (2024)
FORESCENE: FOREcasting human activity via latent SCENE graphs diffusion
por: Alliegro, Antonio, et al.
Publicado: (2025)
por: Alliegro, Antonio, et al.
Publicado: (2025)
SMART: Scene-motion-aware human action recognition framework for mental disorder group
por: Lai, Zengyuan, et al.
Publicado: (2024)
por: Lai, Zengyuan, et al.
Publicado: (2024)
EarthLoc: Astronaut Photography Localization by Indexing Earth from Space
por: Berton, Gabriele, et al.
Publicado: (2024)
por: Berton, Gabriele, et al.
Publicado: (2024)
The Unreasonable Effectiveness of Pre-Trained Features for Camera Pose Refinement
por: Trivigno, Gabriele, et al.
Publicado: (2024)
por: Trivigno, Gabriele, et al.
Publicado: (2024)
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
por: Rosi, Gabriele, et al.
Publicado: (2026)
por: Rosi, Gabriele, et al.
Publicado: (2026)
PEM: Prototype-based Efficient MaskFormer for Image Segmentation
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
por: Cavagnero, Niccolò, et al.
Publicado: (2024)
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
por: Wang, Weizhuo, et al.
Publicado: (2024)
por: Wang, Weizhuo, et al.
Publicado: (2024)
Data-driven Exploration of Mobility Interaction Patterns
por: Galatolo, Gabriele, et al.
Publicado: (2025)
por: Galatolo, Gabriele, et al.
Publicado: (2025)
MeshVPR: Citywide Visual Place Recognition Using 3D Meshes
por: Berton, Gabriele, et al.
Publicado: (2024)
por: Berton, Gabriele, et al.
Publicado: (2024)
Egocentric Action-aware Inertial Localization in Point Clouds with Vision-Language Guidance
por: Zhang, Mingfang, et al.
Publicado: (2025)
por: Zhang, Mingfang, et al.
Publicado: (2025)
Fixed External Cameras as Common Prior Maps for Active 3D Scene Graph Generation
por: Modi, Giorgia, et al.
Publicado: (2026)
por: Modi, Giorgia, et al.
Publicado: (2026)
RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots
por: Modi, Giorgia, et al.
Publicado: (2026)
por: Modi, Giorgia, et al.
Publicado: (2026)
EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision
por: Forte, Rosario, et al.
Publicado: (2026)
por: Forte, Rosario, et al.
Publicado: (2026)
Multi-model learning by sequential reading of untrimmed videos for action recognition
por: Kamiya, Kodai, et al.
Publicado: (2024)
por: Kamiya, Kodai, et al.
Publicado: (2024)
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
por: Su, Yuejiao, et al.
Publicado: (2026)
por: Su, Yuejiao, et al.
Publicado: (2026)
Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
por: Zhou, Weijie, et al.
Publicado: (2026)
por: Zhou, Weijie, et al.
Publicado: (2026)
Can VLMs be used on videos for action recognition? LLMs are Visual Reasoning Coordinators
por: Lunia, Harsh
Publicado: (2024)
por: Lunia, Harsh
Publicado: (2024)
Real-time 3D human action recognition based on Hyperpoint sequence
por: Li, Xing, et al.
Publicado: (2021)
por: Li, Xing, et al.
Publicado: (2021)
HMD^2: Environment-aware Motion Generation from Single Egocentric Head-Mounted Device
por: Guzov, Vladimir, et al.
Publicado: (2024)
por: Guzov, Vladimir, et al.
Publicado: (2024)
CaRe-Ego: Contact-aware Relationship Modeling for Egocentric Interactive Hand-object Segmentation
por: Su, Yuejiao, et al.
Publicado: (2024)
por: Su, Yuejiao, et al.
Publicado: (2024)
Robust Egocentric Visual Attention Prediction Through Language-guided Scene Context-aware Learning
por: Park, Sungjune, et al.
Publicado: (2026)
por: Park, Sungjune, et al.
Publicado: (2026)
EgoCogNav: Cognition-aware Human Egocentric Navigation
por: Qiu, Zhiwen, et al.
Publicado: (2025)
por: Qiu, Zhiwen, et al.
Publicado: (2025)
Community-aware evaluation and threshold calibration for open-set plankton image recognition
por: Chen, Xi, et al.
Publicado: (2026)
por: Chen, Xi, et al.
Publicado: (2026)
From Videos to Conversations: Egocentric Instructions for Task Assistance
por: Aggarwal, Lavisha, et al.
Publicado: (2026)
por: Aggarwal, Lavisha, et al.
Publicado: (2026)
Ejemplares similares
-
Domain Generalization using Action Sequences for Egocentric Action Recognition
por: Nasirimajd, Amirshayan, et al.
Publicado: (2025) -
A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2024) -
Hier-EgoPack: Hierarchical Egocentric Video Understanding with Diverse Task Perspectives
por: Peirone, Simone Alberto, et al.
Publicado: (2025) -
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
por: Peirone, Simone Alberto, et al.
Publicado: (2025) -
HiERO-StepG @ Ego4D Step Grounding Challenge: hierarchical activity understanding enables zero-shot step grounding
por: Zenotto, Andrea, et al.
Publicado: (2026)