Gespeichert in:
| Hauptverfasser: | Li, Yifan, Zhou, Xinyu, Ge, Yunhao, Kong, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.20085 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Medical Visual Grounding via Knowledge-guided Spatial Prompts
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
Attention to Trajectory: Trajectory-Aware Open-Vocabulary Tracking
von: Li, Yunhao, et al.
Veröffentlicht: (2025)
von: Li, Yunhao, et al.
Veröffentlicht: (2025)
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
von: Wang, Weizhuo, et al.
Veröffentlicht: (2024)
von: Wang, Weizhuo, et al.
Veröffentlicht: (2024)
Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision
von: Yoshida, Tomoya, et al.
Veröffentlicht: (2025)
von: Yoshida, Tomoya, et al.
Veröffentlicht: (2025)
Egocentric Event-Based Vision for Ping Pong Ball Trajectory Prediction
von: Alberico, Ivan, et al.
Veröffentlicht: (2025)
von: Alberico, Ivan, et al.
Veröffentlicht: (2025)
DreamDistribution: Learning Prompt Distribution for Diverse In-distribution Generation
von: Zhao, Brian Nlong, et al.
Veröffentlicht: (2023)
von: Zhao, Brian Nlong, et al.
Veröffentlicht: (2023)
I-Scene: 3D Instance Models are Implicit Generalizable Spatial Learners
von: Ling, Lu, et al.
Veröffentlicht: (2025)
von: Ling, Lu, et al.
Veröffentlicht: (2025)
Spatial-Conditioned Reasoning in Long-Egocentric Videos
von: Tribble, James, et al.
Veröffentlicht: (2026)
von: Tribble, James, et al.
Veröffentlicht: (2026)
EgoPrompt: Prompt Learning for Egocentric Action Recognition
von: Lyu, Huaihai, et al.
Veröffentlicht: (2025)
von: Lyu, Huaihai, et al.
Veröffentlicht: (2025)
MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
HEADS-UP: Head-Mounted Egocentric Dataset for Trajectory Prediction in Blind Assistance Systems
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2024)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2024)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
von: Guo, Pinxue, et al.
Veröffentlicht: (2024)
Visual Intention Grounding for Egocentric Assistants
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
EgoActor: Grounding Task Planning into Spatial-aware Egocentric Actions for Humanoid Robots via Visual-Language Models
von: Bai, Yu, et al.
Veröffentlicht: (2026)
von: Bai, Yu, et al.
Veröffentlicht: (2026)
How You Move Tells What You'll Do: Trajectory-Conditioned Egocentric Prediction
von: Jun, Sejoon, et al.
Veröffentlicht: (2026)
von: Jun, Sejoon, et al.
Veröffentlicht: (2026)
Intention Enhanced Diffusion Model for Multimodal Pedestrian Trajectory Prediction
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
Attention to the Burstiness in Visual Prompt Tuning!
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
von: Wang, Yuzhu, et al.
Veröffentlicht: (2025)
VPN: Visual Prompt Navigation
von: Feng, Shuo, et al.
Veröffentlicht: (2025)
von: Feng, Shuo, et al.
Veröffentlicht: (2025)
STF: Spatial Temporal Fusion for Trajectory Prediction
von: Han, Pengqian, et al.
Veröffentlicht: (2023)
von: Han, Pengqian, et al.
Veröffentlicht: (2023)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
von: Chen, Mingfei, et al.
Veröffentlicht: (2025)
von: Chen, Mingfei, et al.
Veröffentlicht: (2025)
Diagnose, Correct, and Learn from Manipulation Failures via Visual Symbols
von: Zeng, Xianchao, et al.
Veröffentlicht: (2025)
von: Zeng, Xianchao, et al.
Veröffentlicht: (2025)
EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices
von: Yu, Liuchuan, et al.
Veröffentlicht: (2026)
von: Yu, Liuchuan, et al.
Veröffentlicht: (2026)
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
von: Rong, Jintao, et al.
Veröffentlicht: (2023)
von: Rong, Jintao, et al.
Veröffentlicht: (2023)
Visual Fact Checker: Enabling High-Fidelity Detailed Caption Generation
von: Ge, Yunhao, et al.
Veröffentlicht: (2024)
von: Ge, Yunhao, et al.
Veröffentlicht: (2024)
EgoAVU: Egocentric Audio-Visual Understanding
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
3D Skew Gaussian Splatting with Any Camera Trajectory Visualization Engine
von: Zhao, Beizhen, et al.
Veröffentlicht: (2026)
von: Zhao, Beizhen, et al.
Veröffentlicht: (2026)
VRP-SAM: SAM with Visual Reference Prompt
von: Sun, Yanpeng, et al.
Veröffentlicht: (2024)
von: Sun, Yanpeng, et al.
Veröffentlicht: (2024)
Robust Egocentric Visual Attention Prediction Through Language-guided Scene Context-aware Learning
von: Park, Sungjune, et al.
Veröffentlicht: (2026)
von: Park, Sungjune, et al.
Veröffentlicht: (2026)
Visual Attention Prompted Prediction and Learning
von: Zhang, Yifei, et al.
Veröffentlicht: (2023)
von: Zhang, Yifei, et al.
Veröffentlicht: (2023)
TP-DRSeg: Improving Diabetic Retinopathy Lesion Segmentation with Explicit Text-Prompts Assisted SAM
von: Li, Wenxue, et al.
Veröffentlicht: (2024)
von: Li, Wenxue, et al.
Veröffentlicht: (2024)
Visual Trajectory Prediction of Vessels for Inland Navigation
von: Puzicha, Alexander, et al.
Veröffentlicht: (2025)
von: Puzicha, Alexander, et al.
Veröffentlicht: (2025)
Spherical World-Locking for Audio-Visual Localization in Egocentric Videos
von: Yun, Heeseung, et al.
Veröffentlicht: (2024)
von: Yun, Heeseung, et al.
Veröffentlicht: (2024)
UnAC: Adaptive Visual Prompting with Abstraction and Stepwise Checking for Complex Multimodal Reasoning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
von: Chen, Hongyu, et al.
Veröffentlicht: (2024)
von: Chen, Hongyu, et al.
Veröffentlicht: (2024)
Point-It-Out: Benchmarking Embodied Reasoning for Vision Language Models in Multi-Stage Visual Grounding
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
von: Xue, Haotian, et al.
Veröffentlicht: (2025)
FedMGP: Personalized Federated Learning with Multi-Group Text-Visual Prompts
von: Bo, Weihao, et al.
Veröffentlicht: (2025)
von: Bo, Weihao, et al.
Veröffentlicht: (2025)
Instance Tracking in 3D Scenes from Egocentric Videos
von: Zhao, Yunhan, et al.
Veröffentlicht: (2023)
von: Zhao, Yunhan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Enhancing Medical Visual Grounding via Knowledge-guided Spatial Prompts
von: Gao, Yifan, et al.
Veröffentlicht: (2026) -
Attention to Trajectory: Trajectory-Aware Open-Vocabulary Tracking
von: Li, Yunhao, et al.
Veröffentlicht: (2025) -
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
von: Wang, Weizhuo, et al.
Veröffentlicht: (2024) -
Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision
von: Yoshida, Tomoya, et al.
Veröffentlicht: (2025) -
Egocentric Event-Based Vision for Ping Pong Ball Trajectory Prediction
von: Alberico, Ivan, et al.
Veröffentlicht: (2025)