Guardado en:
| Autores principales: | Darkhalil, Ahmad, Guerrier, Rhodri, Harley, Adam W., Damen, Dima |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2412.04592 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PointSt3R: Point Tracking through 3D Grounded Correspondence
por: Guerrier, Rhodri, et al.
Publicado: (2025)
por: Guerrier, Rhodri, et al.
Publicado: (2025)
HD-EPIC: A Highly-Detailed Egocentric Video Dataset
por: Perrett, Toby, et al.
Publicado: (2025)
por: Perrett, Toby, et al.
Publicado: (2025)
Get a Grip: Reconstructing Hand-Object Stable Grasps in Egocentric Videos
por: Zhu, Zhifan, et al.
Publicado: (2023)
por: Zhu, Zhifan, et al.
Publicado: (2023)
EPIC Fields: Marrying 3D Geometry and Video Understanding
por: Tschernezki, Vadim, et al.
Publicado: (2023)
por: Tschernezki, Vadim, et al.
Publicado: (2023)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
por: Zhu, Zhifan, et al.
Publicado: (2025)
por: Zhu, Zhifan, et al.
Publicado: (2025)
The N-Body Problem: Parallel Execution from Single-Person Egocentric Video
por: Zhu, Zhifan, et al.
Publicado: (2025)
por: Zhu, Zhifan, et al.
Publicado: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
por: Bansal, Siddhant, et al.
Publicado: (2024)
por: Bansal, Siddhant, et al.
Publicado: (2024)
The Invisible EgoHand: 3D Hand Forecasting through EgoBody Pose Estimation
por: Hatano, Masashi, et al.
Publicado: (2025)
por: Hatano, Masashi, et al.
Publicado: (2025)
Segmenting Collision Sound Sources in Egocentric Videos
por: Parida, Kranti Kumar, et al.
Publicado: (2025)
por: Parida, Kranti Kumar, et al.
Publicado: (2025)
Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind
por: Plizzari, Chiara, et al.
Publicado: (2024)
por: Plizzari, Chiara, et al.
Publicado: (2024)
Generative Point Tracking with Flow Matching
por: Tesfaldet, Mattie, et al.
Publicado: (2025)
por: Tesfaldet, Mattie, et al.
Publicado: (2025)
Leveraging Auxiliary Information in Text-to-Video Retrieval: A Review
por: Fragomeni, Adriano, et al.
Publicado: (2025)
por: Fragomeni, Adriano, et al.
Publicado: (2025)
Leveraging Modality Tags for Enhanced Cross-Modal Video Retrieval
por: Fragomeni, Adriano, et al.
Publicado: (2025)
por: Fragomeni, Adriano, et al.
Publicado: (2025)
Moment of Untruth: Dealing with Negative Queries in Video Moment Retrieval
por: Flanagan, Kevin, et al.
Publicado: (2025)
por: Flanagan, Kevin, et al.
Publicado: (2025)
An Outlook into the Future of Egocentric Vision
por: Plizzari, Chiara, et al.
Publicado: (2023)
por: Plizzari, Chiara, et al.
Publicado: (2023)
It's Just Another Day: Unique Video Captioning by Discriminative Prompting
por: Perrett, Toby, et al.
Publicado: (2024)
por: Perrett, Toby, et al.
Publicado: (2024)
Every Shot Counts: Using Exemplars for Repetition Counting in Videos
por: Sinha, Saptarshi, et al.
Publicado: (2024)
por: Sinha, Saptarshi, et al.
Publicado: (2024)
TAPIP3D: Tracking Any Point in Persistent 3D Geometry
por: Zhang, Bowei, et al.
Publicado: (2025)
por: Zhang, Bowei, et al.
Publicado: (2025)
Video Editing for Video Retrieval
por: Zhu, Bin, et al.
Publicado: (2024)
por: Zhu, Bin, et al.
Publicado: (2024)
EgoSound: Benchmarking Sound Understanding in Egocentric Videos
por: Zhu, Bingwen, et al.
Publicado: (2026)
por: Zhu, Bingwen, et al.
Publicado: (2026)
GenHowTo: Learning to Generate Actions and State Transformations from Instructional Videos
por: Souček, Tomáš, et al.
Publicado: (2023)
por: Souček, Tomáš, et al.
Publicado: (2023)
Beyond Caption-Based Queries for Video Moment Retrieval
por: Pujol-Perich, David, et al.
Publicado: (2026)
por: Pujol-Perich, David, et al.
Publicado: (2026)
EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
por: Pei, Baoqi, et al.
Publicado: (2024)
por: Pei, Baoqi, et al.
Publicado: (2024)
EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning
por: Kulkarni, Yogesh, et al.
Publicado: (2025)
por: Kulkarni, Yogesh, et al.
Publicado: (2025)
EgoLCD: Egocentric Video Generation with Long Context Diffusion
por: Zhang, Liuzhou, et al.
Publicado: (2025)
por: Zhang, Liuzhou, et al.
Publicado: (2025)
EgoGraph: Temporal Knowledge Graph for Egocentric Video Understanding
por: Sun, Shitong, et al.
Publicado: (2026)
por: Sun, Shitong, et al.
Publicado: (2026)
LookOut: Real-World Humanoid Egocentric Navigation
por: Pan, Boxiao, et al.
Publicado: (2025)
por: Pan, Boxiao, et al.
Publicado: (2025)
EgoX: Egocentric Video Generation from a Single Exocentric Video
por: Kang, Taewoong, et al.
Publicado: (2025)
por: Kang, Taewoong, et al.
Publicado: (2025)
AMEGO: Active Memory from long EGOcentric videos
por: Goletto, Gabriele, et al.
Publicado: (2024)
por: Goletto, Gabriele, et al.
Publicado: (2024)
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding
por: Nagrani, Arsha, et al.
Publicado: (2026)
por: Nagrani, Arsha, et al.
Publicado: (2026)
EgoVLM: Policy Optimization for Egocentric Video Understanding
por: Vinod, Ashwin, et al.
Publicado: (2025)
por: Vinod, Ashwin, et al.
Publicado: (2025)
Animal Pose Labeling Using General-Purpose Point Trackers
por: Pan, Zhuoyang, et al.
Publicado: (2025)
por: Pan, Zhuoyang, et al.
Publicado: (2025)
EgoMimic: Scaling Imitation Learning via Egocentric Video
por: Kareer, Simar, et al.
Publicado: (2024)
por: Kareer, Simar, et al.
Publicado: (2024)
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
por: Hummel, Thomas, et al.
Publicado: (2024)
por: Hummel, Thomas, et al.
Publicado: (2024)
EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation
por: Leonardi, Rosario, et al.
Publicado: (2026)
por: Leonardi, Rosario, et al.
Publicado: (2026)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
por: Heyward, Joseph, et al.
Publicado: (2024)
por: Heyward, Joseph, et al.
Publicado: (2024)
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
por: Xiao, Junbin, et al.
Publicado: (2026)
por: Xiao, Junbin, et al.
Publicado: (2026)
AllTracker: Efficient Dense Point Tracking at High Resolution
por: Harley, Adam W., et al.
Publicado: (2025)
por: Harley, Adam W., et al.
Publicado: (2025)
Ego-VPA: Egocentric Video Understanding with Parameter-efficient Adaptation
por: Wu, Tz-Ying, et al.
Publicado: (2024)
por: Wu, Tz-Ying, et al.
Publicado: (2024)
Estimating Ego-Body Pose from Doubly Sparse Egocentric Video Data
por: Chi, Seunggeun, et al.
Publicado: (2024)
por: Chi, Seunggeun, et al.
Publicado: (2024)
Ejemplares similares
-
PointSt3R: Point Tracking through 3D Grounded Correspondence
por: Guerrier, Rhodri, et al.
Publicado: (2025) -
HD-EPIC: A Highly-Detailed Egocentric Video Dataset
por: Perrett, Toby, et al.
Publicado: (2025) -
Get a Grip: Reconstructing Hand-Object Stable Grasps in Egocentric Videos
por: Zhu, Zhifan, et al.
Publicado: (2023) -
EPIC Fields: Marrying 3D Geometry and Video Understanding
por: Tschernezki, Vadim, et al.
Publicado: (2023) -
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
por: Zhu, Zhifan, et al.
Publicado: (2025)