DAP: Domain-aware Prompt Learning for Vision-and-Language Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Ting, Hu, Yue, Wu, Wansen, Wang, Youkai, Xu, Kai, Yin, Quanjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding
von: Liu, Ting, et al.
Veröffentlicht: (2024)
von: Liu, Ting, et al.
Veröffentlicht: (2024)
Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model
von: Liu, Ting, et al.
Veröffentlicht: (2024)
von: Liu, Ting, et al.
Veröffentlicht: (2024)
DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition
von: Lin, Jiaying, et al.
Veröffentlicht: (2026)
von: Lin, Jiaying, et al.
Veröffentlicht: (2026)
SwimVG: Step-wise Multimodal Fusion and Adaption for Visual Grounding
von: Shi, Liangtao, et al.
Veröffentlicht: (2025)
von: Shi, Liangtao, et al.
Veröffentlicht: (2025)
Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation
von: Liu, Kuanghong, et al.
Veröffentlicht: (2025)
von: Liu, Kuanghong, et al.
Veröffentlicht: (2025)
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
DAP-MAE: Domain-Adaptive Point Cloud Masked Autoencoder for Effective Cross-Domain Learning
von: Gao, Ziqi, et al.
Veröffentlicht: (2025)
von: Gao, Ziqi, et al.
Veröffentlicht: (2025)
Transitive Vision-Language Prompt Learning for Domain Generalization
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
FedDAP: Domain-Aware Prototype Learning for Federated Learning under Domain Shift
von: Le, Huy Q., et al.
Veröffentlicht: (2026)
von: Le, Huy Q., et al.
Veröffentlicht: (2026)
MaPPER: Multimodal Prior-guided Parameter Efficient Tuning for Referring Expression Comprehension
von: Liu, Ting, et al.
Veröffentlicht: (2024)
von: Liu, Ting, et al.
Veröffentlicht: (2024)
Plug-and-play Class-aware Knowledge Injection for Prompt Learning with Visual-Language Model
von: Yin, Junhui, et al.
Veröffentlicht: (2026)
von: Yin, Junhui, et al.
Veröffentlicht: (2026)
Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation
von: Hu, Zixuan, et al.
Veröffentlicht: (2026)
von: Hu, Zixuan, et al.
Veröffentlicht: (2026)
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models
von: Yin, Xiaojie, et al.
Veröffentlicht: (2025)
von: Yin, Xiaojie, et al.
Veröffentlicht: (2025)
M2IST: Multi-Modal Interactive Side-Tuning for Efficient Referring Expression Comprehension
von: Liu, Xuyang, et al.
Veröffentlicht: (2024)
von: Liu, Xuyang, et al.
Veröffentlicht: (2024)
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
von: Lei, Ting, et al.
Veröffentlicht: (2025)
von: Lei, Ting, et al.
Veröffentlicht: (2025)
LG-Gaze: Learning Geometry-aware Continuous Prompts for Language-Guided Gaze Estimation
von: Yin, Pengwei, et al.
Veröffentlicht: (2024)
von: Yin, Pengwei, et al.
Veröffentlicht: (2024)
Semantics-aware Motion Retargeting with Vision-Language Models
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
In-context Prompt Learning for Test-time Vision Recognition with Frozen Vision-language Model
von: Yin, Junhui, et al.
Veröffentlicht: (2024)
von: Yin, Junhui, et al.
Veröffentlicht: (2024)
Integrated Structural Prompt Learning for Vision-Language Models
von: Wang, Jiahui, et al.
Veröffentlicht: (2025)
von: Wang, Jiahui, et al.
Veröffentlicht: (2025)
Modeling Variants of Prompts for Vision-Language Models
von: Li, Ao, et al.
Veröffentlicht: (2025)
von: Li, Ao, et al.
Veröffentlicht: (2025)
PanoGen++: Domain-Adapted Text-Guided Panoramic Environment Generation for Vision-and-Language Navigation
von: Wang, Sen, et al.
Veröffentlicht: (2025)
von: Wang, Sen, et al.
Veröffentlicht: (2025)
Why Only Text: Empowering Vision-and-Language Navigation with Multi-modal Prompts
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
CityCube: Benchmarking Cross-view Spatial Reasoning on Vision-Language Models in Urban Environments
von: Xu, Haotian, et al.
Veröffentlicht: (2026)
von: Xu, Haotian, et al.
Veröffentlicht: (2026)
Domain-Invariant Prompt Learning for Vision-Language Models
von: Khoee, Arsham Gholamzadeh, et al.
Veröffentlicht: (2026)
von: Khoee, Arsham Gholamzadeh, et al.
Veröffentlicht: (2026)
Dynamic Topology Awareness: Breaking the Granularity Rigidity in Vision-Language Navigation
von: Peng, Jiankun, et al.
Veröffentlicht: (2026)
von: Peng, Jiankun, et al.
Veröffentlicht: (2026)
Multi-modal Attribute Prompting for Vision-Language Models
von: Liu, Xin, et al.
Veröffentlicht: (2024)
von: Liu, Xin, et al.
Veröffentlicht: (2024)
VaMP: Variational Multi-Modal Prompt Learning for Vision-Language Models
von: Cheng, Silin, et al.
Veröffentlicht: (2025)
von: Cheng, Silin, et al.
Veröffentlicht: (2025)
TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation
von: Liu, Jiaxing, et al.
Veröffentlicht: (2026)
von: Liu, Jiaxing, et al.
Veröffentlicht: (2026)
DAP-LED: Learning Degradation-Aware Priors with CLIP for Joint Low-light Enhancement and Deblurring
von: Wang, Ling, et al.
Veröffentlicht: (2024)
von: Wang, Ling, et al.
Veröffentlicht: (2024)
AVION: Aerial Vision-Language Instruction from Offline Teacher to Prompt-Tuned Network
von: Hu, Yu, et al.
Veröffentlicht: (2026)
von: Hu, Yu, et al.
Veröffentlicht: (2026)
Closed-Loop Bidirectional Prompting for Adversarial Robustness of Vision Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
von: Liu, Xiao, et al.
Veröffentlicht: (2026)
Volumetric Environment Representation for Vision-Language Navigation
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Vision-Language Navigation with Energy-Based Policy
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
AeroDuo: Aerial Duo for UAV-based Vision and Language Navigation
von: Wu, Ruipu, et al.
Veröffentlicht: (2025)
von: Wu, Ruipu, et al.
Veröffentlicht: (2025)
UAOR: Uncertainty-aware Observation Reinjection for Vision-Language-Action Models
von: Yang, Jiabing, et al.
Veröffentlicht: (2026)
von: Yang, Jiabing, et al.
Veröffentlicht: (2026)
Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
von: Hong, Haodong, et al.
Veröffentlicht: (2024)
Cascade Prompt Learning for Vision-Language Model Adaptation
von: Wu, Ge, et al.
Veröffentlicht: (2024)
von: Wu, Ge, et al.
Veröffentlicht: (2024)
Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology
von: Wang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiangyu, et al.
Veröffentlicht: (2024)
Causality-based Cross-Modal Representation Learning for Vision-and-Language Navigation
von: Wang, Liuyi, et al.
Veröffentlicht: (2024)
von: Wang, Liuyi, et al.
Veröffentlicht: (2024)
LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation
von: Peng, Jiankun, et al.
Veröffentlicht: (2026)
von: Peng, Jiankun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding
von: Liu, Ting, et al.
Veröffentlicht: (2024) -
Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model
von: Liu, Ting, et al.
Veröffentlicht: (2024) -
DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition
von: Lin, Jiaying, et al.
Veröffentlicht: (2026) -
SwimVG: Step-wise Multimodal Fusion and Adaption for Visual Grounding
von: Shi, Liangtao, et al.
Veröffentlicht: (2025) -
Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation
von: Liu, Kuanghong, et al.
Veröffentlicht: (2025)