Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lall, Vishakha, Liu, Yisi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024)
von: An, Joungbin, et al.
Veröffentlicht: (2024)
Dynamic Stress Detection: A Study of Temporal Progression Modelling of Stress in Speech
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
Object-Shot Enhanced Grounding Network for Egocentric Video
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
Gaze-VLM:Bridging Gaze and VLMs through Attention Regularization for Egocentric Understanding
von: Pani, Anupam, et al.
Veröffentlicht: (2025)
von: Pani, Anupam, et al.
Veröffentlicht: (2025)
AI Meets Maritime Training: Precision Analytics for Enhanced Safety and Performance
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
Toward Gaze Target Detection of Young Autistic Children
von: Deng, Shijian, et al.
Veröffentlicht: (2025)
von: Deng, Shijian, et al.
Veröffentlicht: (2025)
EgoDTM: Towards 3D-Aware Egocentric Video-Language Pretraining
von: Xu, Boshen, et al.
Veröffentlicht: (2025)
von: Xu, Boshen, et al.
Veröffentlicht: (2025)
Gazing Into Missteps: Leveraging Eye-Gaze for Unsupervised Mistake Detection in Egocentric Videos of Skilled Human Activities
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
von: Mazzamuto, Michele, et al.
Veröffentlicht: (2024)
Gaze-Guided 3D Hand Motion Prediction for Detecting Intent in Egocentric Grasping Tasks
von: He, Yufei, et al.
Veröffentlicht: (2025)
von: He, Yufei, et al.
Veröffentlicht: (2025)
GazeMoE: Perception of Gaze Target with Mixture-of-Experts
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2026)
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2026)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
von: Li, Yiwei, et al.
Veröffentlicht: (2026)
von: Li, Yiwei, et al.
Veröffentlicht: (2026)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
von: Kumar, Yogesh, et al.
Veröffentlicht: (2025)
von: Kumar, Yogesh, et al.
Veröffentlicht: (2025)
Prompt-and-Check: Using Large Language Models to Evaluate Communication Protocol Compliance in Simulation-Based Training
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
In the Eye of MLLM: Benchmarking Egocentric Video Intent Understanding with Gaze-Guided Prompting
von: Peng, Taiying, et al.
Veröffentlicht: (2025)
von: Peng, Taiying, et al.
Veröffentlicht: (2025)
When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Detection
von: Kim, Jihyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jihyeon, et al.
Veröffentlicht: (2026)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
EgoVLM: Policy Optimization for Egocentric Video Understanding
von: Vinod, Ashwin, et al.
Veröffentlicht: (2025)
von: Vinod, Ashwin, et al.
Veröffentlicht: (2025)
Continual Multimodal Egocentric Activity Recognition via Modality-Aware Novel Detection
von: Lim, Wonseon, et al.
Veröffentlicht: (2026)
von: Lim, Wonseon, et al.
Veröffentlicht: (2026)
Object-fabrication Targeted Attack for Object Detection
von: Zhang, Xuchong, et al.
Veröffentlicht: (2022)
von: Zhang, Xuchong, et al.
Veröffentlicht: (2022)
GazeQwen: Lightweight Gaze-Conditioned LLM Modulation for Streaming Video Understanding
von: Pham, Trong Thang, et al.
Veröffentlicht: (2026)
von: Pham, Trong Thang, et al.
Veröffentlicht: (2026)
EyeCue: Driver Cognitive Distraction Detection via Gaze-Empowered Egocentric Video Understanding
von: Zhang, Lang, et al.
Veröffentlicht: (2026)
von: Zhang, Lang, et al.
Veröffentlicht: (2026)
DeepFake Detection in Dyadic Video Calls using Point of Gaze Tracking
von: Kohler, Odin, et al.
Veröffentlicht: (2025)
von: Kohler, Odin, et al.
Veröffentlicht: (2025)
Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball Rotation
von: Choi, YoungChan, et al.
Veröffentlicht: (2025)
von: Choi, YoungChan, et al.
Veröffentlicht: (2025)
EGOILLUSION: Benchmarking Hallucinations in Egocentric Video Understanding
von: Seth, Ashish, et al.
Veröffentlicht: (2025)
von: Seth, Ashish, et al.
Veröffentlicht: (2025)
Comparing Learning Paradigms for Egocentric Video Summarization
von: Wen, Daniel
Veröffentlicht: (2025)
von: Wen, Daniel
Veröffentlicht: (2025)
Exploring Audio Hallucination in Egocentric Video Understanding
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
von: Bhalgat, Yash, et al.
Veröffentlicht: (2024)
von: Bhalgat, Yash, et al.
Veröffentlicht: (2024)
In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation
von: Lai, Bolin, et al.
Veröffentlicht: (2022)
von: Lai, Bolin, et al.
Veröffentlicht: (2022)
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
von: Zhao, Xinyuan, et al.
Veröffentlicht: (2026)
von: Zhao, Xinyuan, et al.
Veröffentlicht: (2026)
Contextual Biasing to Improve Domain-specific Custom Vocabulary Audio Transcription without Explicit Fine-Tuning of Whisper Model
von: Lall, Vishakha, et al.
Veröffentlicht: (2024)
von: Lall, Vishakha, et al.
Veröffentlicht: (2024)
SViTT-Ego: A Sparse Video-Text Transformer for Egocentric Video
von: Valdez, Hector A., et al.
Veröffentlicht: (2024)
von: Valdez, Hector A., et al.
Veröffentlicht: (2024)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
Gazing at Rewards: Eye Movements as a Lens into Human and AI Decision-Making in Hybrid Visual Foraging
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos
von: Lee, Daeun, et al.
Veröffentlicht: (2025)
von: Lee, Daeun, et al.
Veröffentlicht: (2025)
StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
Context-Aware Temporal Embedding of Objects in Video Data
von: Farhan, Ahnaf, et al.
Veröffentlicht: (2024)
von: Farhan, Ahnaf, et al.
Veröffentlicht: (2024)
PolarBEVDet: Exploring Polar Representation for Multi-View 3D Object Detection in Bird's-Eye-View
von: Yu, Zichen, et al.
Veröffentlicht: (2024)
von: Yu, Zichen, et al.
Veröffentlicht: (2024)
From Scene to Object: Text-Guided Dual-Gaze Prediction
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
Toddlers' Active Gaze Behavior Supports Self-Supervised Object Learning
von: Yu, Zhengyang, et al.
Veröffentlicht: (2024)
von: Yu, Zhengyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024) -
Dynamic Stress Detection: A Study of Temporal Progression Modelling of Stress in Speech
von: Lall, Vishakha, et al.
Veröffentlicht: (2025) -
Object-Shot Enhanced Grounding Network for Egocentric Video
von: Feng, Yisen, et al.
Veröffentlicht: (2025) -
Gaze-VLM:Bridging Gaze and VLMs through Attention Regularization for Egocentric Understanding
von: Pani, Anupam, et al.
Veröffentlicht: (2025) -
AI Meets Maritime Training: Precision Analytics for Enhanced Safety and Performance
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)