Adaptive Visual Navigation Assistant in 3D RPGs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Kaijie, Verbrugge, Clark |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NavigateDiff: Visual Predictors are Zero-Shot Navigation Assistants
von: Qin, Yiran, et al.
Veröffentlicht: (2025)
von: Qin, Yiran, et al.
Veröffentlicht: (2025)
MR.NAVI: Mixed-Reality Navigation Assistant for the Visually Impaired
von: Pfitzer, Nicolas, et al.
Veröffentlicht: (2025)
von: Pfitzer, Nicolas, et al.
Veröffentlicht: (2025)
HRVDA: High-Resolution Visual Document Assistant
von: Liu, Chaohu, et al.
Veröffentlicht: (2024)
von: Liu, Chaohu, et al.
Veröffentlicht: (2024)
SceneAssistant: A Visual Feedback Agent for Open-Vocabulary 3D Scene Generation
von: Luo, Jun, et al.
Veröffentlicht: (2026)
von: Luo, Jun, et al.
Veröffentlicht: (2026)
Visual Intention Grounding for Egocentric Assistants
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
Neural Radiance and Gaze Fields for Visual Attention Modeling in 3D Environments
von: Chubarau, Andrei, et al.
Veröffentlicht: (2025)
von: Chubarau, Andrei, et al.
Veröffentlicht: (2025)
Implicit Discriminative Knowledge Learning for Visible-Infrared Person Re-Identification
von: Ren, Kaijie, et al.
Veröffentlicht: (2024)
von: Ren, Kaijie, et al.
Veröffentlicht: (2024)
RAVE: Rate-Adaptive Visual Encoding for 3D Gaussian Splatting
von: Tran, Hoang-Nhat, et al.
Veröffentlicht: (2025)
von: Tran, Hoang-Nhat, et al.
Veröffentlicht: (2025)
Adaptive Scaling with Geometric and Visual Continuity of completed 3D objects
von: Vermandere, Jelle, et al.
Veröffentlicht: (2026)
von: Vermandere, Jelle, et al.
Veröffentlicht: (2026)
NeRF Is a Valuable Assistant for 3D Gaussian Splatting
von: Fang, Shuangkang, et al.
Veröffentlicht: (2025)
von: Fang, Shuangkang, et al.
Veröffentlicht: (2025)
Resolving Ambiguity in Gaze-Facilitated Visual Assistant Interaction Paradigm
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
von: Wang, Zeyu, et al.
Veröffentlicht: (2025)
Recursive Visual Imagination and Adaptive Linguistic Grounding for Vision Language Navigation
von: Chen, Bolei, et al.
Veröffentlicht: (2025)
von: Chen, Bolei, et al.
Veröffentlicht: (2025)
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
von: Chen, Huiyi, et al.
Veröffentlicht: (2025)
CLOVA: A Closed-Loop Visual Assistant with Tool Usage and Update
von: Gao, Zhi, et al.
Veröffentlicht: (2023)
von: Gao, Zhi, et al.
Veröffentlicht: (2023)
VSAS-Bench: Real-Time Evaluation of Visual Streaming Assistant Models
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2026)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2026)
EmoAssist: Emotional Assistant for Visual Impairment Community
von: Qi, Xingyu, et al.
Veröffentlicht: (2025)
von: Qi, Xingyu, et al.
Veröffentlicht: (2025)
VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic Memory
von: Wang, Shaoan, et al.
Veröffentlicht: (2026)
von: Wang, Shaoan, et al.
Veröffentlicht: (2026)
VPN: Visual Prompt Navigation
von: Feng, Shuo, et al.
Veröffentlicht: (2025)
von: Feng, Shuo, et al.
Veröffentlicht: (2025)
Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation
von: Zhu, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhu, Ziyu, et al.
Veröffentlicht: (2025)
A Database-Driven Framework for 3D Level Generation with LLMs
von: Xu, Kaijie, et al.
Veröffentlicht: (2025)
von: Xu, Kaijie, et al.
Veröffentlicht: (2025)
IPFormer: Visual 3D Panoptic Scene Completion with Context-Adaptive Instance Proposals
von: Gross, Markus, et al.
Veröffentlicht: (2025)
von: Gross, Markus, et al.
Veröffentlicht: (2025)
MonoTAKD: Teaching Assistant Knowledge Distillation for Monocular 3D Object Detection
von: Liu, Hou-I, et al.
Veröffentlicht: (2024)
von: Liu, Hou-I, et al.
Veröffentlicht: (2024)
Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaboration
von: Wang, Junyang, et al.
Veröffentlicht: (2024)
von: Wang, Junyang, et al.
Veröffentlicht: (2024)
PixWizard: Versatile Image-to-Image Visual Assistant with Open-Language Instructions
von: Lin, Weifeng, et al.
Veröffentlicht: (2024)
von: Lin, Weifeng, et al.
Veröffentlicht: (2024)
Empowering Visual Creativity: A Vision-Language Assistant to Image Editing Recommendations
von: Shen, Tiancheng, et al.
Veröffentlicht: (2024)
von: Shen, Tiancheng, et al.
Veröffentlicht: (2024)
Pathfinder for Low-altitude Aircraft with Binary Neural Network
von: Yin, Kaijie, et al.
Veröffentlicht: (2024)
von: Yin, Kaijie, et al.
Veröffentlicht: (2024)
Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations
von: Zhang, Xuesong, et al.
Veröffentlicht: (2024)
von: Zhang, Xuesong, et al.
Veröffentlicht: (2024)
Memory Proxy Maps for Visual Navigation
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
VS-Assistant: Versatile Surgery Assistant on the Demand of Surgeons
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences
von: Zhi, Hongyan, et al.
Veröffentlicht: (2024)
von: Zhi, Hongyan, et al.
Veröffentlicht: (2024)
PD-APE: A Parallel Decoding Framework with Adaptive Position Encoding for 3D Visual Grounding
von: Hou, Chenshu, et al.
Veröffentlicht: (2024)
von: Hou, Chenshu, et al.
Veröffentlicht: (2024)
Visual Trajectory Prediction of Vessels for Inland Navigation
von: Puzicha, Alexander, et al.
Veröffentlicht: (2025)
von: Puzicha, Alexander, et al.
Veröffentlicht: (2025)
Turn-by-Turn Indoor Navigation for the Visually Impaired
von: Srinivasaiah, Santosh, et al.
Veröffentlicht: (2024)
von: Srinivasaiah, Santosh, et al.
Veröffentlicht: (2024)
GaussNav: Gaussian Splatting for Visual Navigation
von: Lei, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lei, Xiaohan, et al.
Veröffentlicht: (2024)
Weakly-Supervised 3D Visual Grounding based on Visual Language Alignment
von: Xu, Xiaoxu, et al.
Veröffentlicht: (2023)
von: Xu, Xiaoxu, et al.
Veröffentlicht: (2023)
AgriDoctor: A Multimodal Intelligent Assistant for Agriculture
von: Zhang, Mingqing, et al.
Veröffentlicht: (2025)
von: Zhang, Mingqing, et al.
Veröffentlicht: (2025)
PerLA: Perceptive 3D Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
High-Fidelity Differential-information Driven Binary Vision Transformer
von: Gao, Tian, et al.
Veröffentlicht: (2025)
von: Gao, Tian, et al.
Veröffentlicht: (2025)
Towards Physically Executable 3D Gaussian for Embodied Navigation
von: Miao, Bingchen, et al.
Veröffentlicht: (2025)
von: Miao, Bingchen, et al.
Veröffentlicht: (2025)
Task-oriented Sequential Grounding and Navigation in 3D Scenes
von: Zhang, Zhuofan, et al.
Veröffentlicht: (2024)
von: Zhang, Zhuofan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
NavigateDiff: Visual Predictors are Zero-Shot Navigation Assistants
von: Qin, Yiran, et al.
Veröffentlicht: (2025) -
MR.NAVI: Mixed-Reality Navigation Assistant for the Visually Impaired
von: Pfitzer, Nicolas, et al.
Veröffentlicht: (2025) -
HRVDA: High-Resolution Visual Document Assistant
von: Liu, Chaohu, et al.
Veröffentlicht: (2024) -
SceneAssistant: A Visual Feedback Agent for Open-Vocabulary 3D Scene Generation
von: Luo, Jun, et al.
Veröffentlicht: (2026) -
Visual Intention Grounding for Egocentric Assistants
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)