Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Yuping, Huang, Yifei, Chen, Guo, Lu, Lidong, Pei, Baoqi, Xu, Jilan, Lu, Tong, Sato, Yoichi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs
von: He, Yuping, et al.
Veröffentlicht: (2025)
von: He, Yuping, et al.
Veröffentlicht: (2025)
EgoThinker: Unveiling Egocentric Reasoning with Spatio-Temporal CoT
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
von: Pei, Baoqi, et al.
Veröffentlicht: (2024)
von: Pei, Baoqi, et al.
Veröffentlicht: (2024)
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023)
CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
von: Chen, Guo, et al.
Veröffentlicht: (2024)
von: Chen, Guo, et al.
Veröffentlicht: (2024)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
Egocentric and Exocentric Methods: A Short Survey
von: Thatipelli, Anirudh, et al.
Veröffentlicht: (2024)
von: Thatipelli, Anirudh, et al.
Veröffentlicht: (2024)
Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model
von: Huang, Yifei, et al.
Veröffentlicht: (2024)
von: Huang, Yifei, et al.
Veröffentlicht: (2024)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
An Egocentric Vision-Language Model based Portable Real-time Smart Assistant
von: Huang, Yifei, et al.
Veröffentlicht: (2025)
von: Huang, Yifei, et al.
Veröffentlicht: (2025)
The Audio-Visual Conversational Graph: From an Egocentric-Exocentric Perspective
von: Jia, Wenqi, et al.
Veröffentlicht: (2023)
von: Jia, Wenqi, et al.
Veröffentlicht: (2023)
EgoExoLearn: A Dataset for Bridging Asynchronous Ego- and Exo-centric View of Procedural Activities in Real World
von: Huang, Yifei, et al.
Veröffentlicht: (2024)
von: Huang, Yifei, et al.
Veröffentlicht: (2024)
Masked Video and Body-worn IMU Autoencoder for Egocentric Action Recognition
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding
von: Chen, Guo, et al.
Veröffentlicht: (2024)
von: Chen, Guo, et al.
Veröffentlicht: (2024)
Egocentric Action-aware Inertial Localization in Point Clouds with Vision-Language Guidance
von: Zhang, Mingfang, et al.
Veröffentlicht: (2025)
von: Zhang, Mingfang, et al.
Veröffentlicht: (2025)
The N-Body Problem: Parallel Execution from Single-Person Egocentric Video
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Egocentric Video Captioning
von: Xu, Jilan, et al.
Veröffentlicht: (2024)
von: Xu, Jilan, et al.
Veröffentlicht: (2024)
Egocentric Gaze Estimation via Neck-Mounted Camera
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
von: Huang, Haoyu, et al.
Veröffentlicht: (2026)
Put Myself in Your Shoes: Lifting the Egocentric Perspective from Exocentric Videos
von: Luo, Mi, et al.
Veröffentlicht: (2024)
von: Luo, Mi, et al.
Veröffentlicht: (2024)
EgoWorld: Translating Exocentric View to Egocentric View using Rich Exocentric Observations
von: Park, Junho, et al.
Veröffentlicht: (2025)
von: Park, Junho, et al.
Veröffentlicht: (2025)
SFHand: Learning Embodied Manipulation by Streaming Egocentric 3D Hand Forecasting
von: Liu, Ruicong, et al.
Veröffentlicht: (2025)
von: Liu, Ruicong, et al.
Veröffentlicht: (2025)
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
von: Xu, Jilan, et al.
Veröffentlicht: (2025)
von: Xu, Jilan, et al.
Veröffentlicht: (2025)
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
Unlocking Exocentric Video-Language Data for Egocentric Video Representation Learning
von: Dou, Zi-Yi, et al.
Veröffentlicht: (2024)
von: Dou, Zi-Yi, et al.
Veröffentlicht: (2024)
EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2024)
von: Li, Yuan-Ming, et al.
Veröffentlicht: (2024)
EgoX: Egocentric Video Generation from a Single Exocentric Video
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
Exo2Ego: Exocentric Knowledge Guided MLLM for Egocentric Video Understanding
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2025)
O-MaMa: Learning Object Mask Matching between Egocentric and Exocentric Views
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2025)
von: Mur-Labadia, Lorenzo, et al.
Veröffentlicht: (2025)
Single-to-Dual-View Adaptation for Egocentric 3D Hand Pose Estimation
von: Liu, Ruicong, et al.
Veröffentlicht: (2024)
von: Liu, Ruicong, et al.
Veröffentlicht: (2024)
Exo2EgoSyn: Unlocking Foundation Video Generation Models for Exocentric-to-Egocentric Video Synthesis
von: Mahdi, Mohammad, et al.
Veröffentlicht: (2025)
von: Mahdi, Mohammad, et al.
Veröffentlicht: (2025)
Synchronization is All You Need: Exocentric-to-Egocentric Transfer for Temporal Action Segmentation with Unlabeled Synchronized Video Pairs
von: Quattrocchi, Camillo, et al.
Veröffentlicht: (2023)
von: Quattrocchi, Camillo, et al.
Veröffentlicht: (2023)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
von: Deichler, Anna, et al.
Veröffentlicht: (2025)
von: Deichler, Anna, et al.
Veröffentlicht: (2025)
EgoInstruct: An Egocentric Video Dataset of Face-to-face Instructional Interactions with Multi-modal LLM Benchmarking
von: Sakai, Yuki, et al.
Veröffentlicht: (2025)
von: Sakai, Yuki, et al.
Veröffentlicht: (2025)
AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs
von: Lu, Lidong, et al.
Veröffentlicht: (2025)
von: Lu, Lidong, et al.
Veröffentlicht: (2025)
Learning Visual Affordance from Audio
von: Lu, Lidong, et al.
Veröffentlicht: (2025)
von: Lu, Lidong, et al.
Veröffentlicht: (2025)
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
Egocentric Vision Language Planning
von: Fang, Zhirui, et al.
Veröffentlicht: (2024)
von: Fang, Zhirui, et al.
Veröffentlicht: (2024)
Towards Multimodal Lifelong Understanding: A Dataset and Agentic Baseline
von: Chen, Guo, et al.
Veröffentlicht: (2026)
von: Chen, Guo, et al.
Veröffentlicht: (2026)
Challenges and Trends in Egocentric Vision: A Survey
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
ActionVOS: Actions as Prompts for Video Object Segmentation
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2024)
von: Ouyang, Liangyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs
von: He, Yuping, et al.
Veröffentlicht: (2025) -
EgoThinker: Unveiling Egocentric Reasoning with Spatio-Temporal CoT
von: Pei, Baoqi, et al.
Veröffentlicht: (2025) -
EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
von: Pei, Baoqi, et al.
Veröffentlicht: (2024) -
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
von: Truong, Thanh-Dat, et al.
Veröffentlicht: (2023) -
CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding
von: Chen, Guo, et al.
Veröffentlicht: (2024)