SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Pengna, Wu, Kangyi, Xu, Shaoqing, Li, Fang, Li, Hanbing, Zhao, Lin, Lyu, Kailin, Chen, Long, Yang, Zhi-Xin, Zheng, Nanning |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Think before Go: Hierarchical Reasoning for Image-goal Navigation
by: Li, Pengna, et al.
Published: (2026)
by: Li, Pengna, et al.
Published: (2026)
Dual-Anchoring: Addressing State Drift in Vision-Language Navigation
by: Wu, Kangyi, et al.
Published: (2026)
by: Wu, Kangyi, et al.
Published: (2026)
REGNav: Room Expert Guided Image-Goal Navigation
by: Li, Pengna, et al.
Published: (2025)
by: Li, Pengna, et al.
Published: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026)
by: Lyu, Kailin, et al.
Published: (2026)
DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale
by: Zuo, Sicheng, et al.
Published: (2026)
by: Zuo, Sicheng, et al.
Published: (2026)
Camera-aware Label Refinement for Unsupervised Person Re-identification
by: Li, Pengna, et al.
Published: (2024)
by: Li, Pengna, et al.
Published: (2024)
Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
SUM-AgriVLN: Spatial Understanding Memory for Agricultural Vision-and-Language Navigation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
Enhancing Continuous Domain Adaptation with Multi-Path Transfer Curriculum
by: Liu, Hanbing, et al.
Published: (2024)
by: Liu, Hanbing, et al.
Published: (2024)
Vision-Guided MPPI for Agile Drone Racing: Navigating Arbitrary Gate Poses via Neural Signed Distance Fields
by: Zhao, Fangguo, et al.
Published: (2026)
by: Zhao, Fangguo, et al.
Published: (2026)
CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation
by: Wu, Kangyi, et al.
Published: (2025)
by: Wu, Kangyi, et al.
Published: (2025)
SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic Data
by: Ogezi, Michael, et al.
Published: (2025)
by: Ogezi, Michael, et al.
Published: (2025)
T-araVLN: Translator for Agricultural Robotic Agents on Vision-and-Language Navigation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
AudioSpa: Spatializing Sound Events with Text
by: Feng, Linfeng, et al.
Published: (2025)
by: Feng, Linfeng, et al.
Published: (2025)
MDE-AgriVLN: Agricultural Vision-and-Language Navigation with Monocular Depth Estimation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
All-day Multi-scenes Lifelong Vision-and-Language Navigation with Tucker Adaptation
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
AgriVLN: Vision-and-Language Navigation for Agricultural Robots
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
On the structure of noncollapsed Ricci flow limit spaces
by: Fang, Hanbing, et al.
Published: (2025)
by: Fang, Hanbing, et al.
Published: (2025)
Singular sets in noncollapsed Ricci flow limit spaces
by: Fang, Hanbing, et al.
Published: (2025)
by: Fang, Hanbing, et al.
Published: (2025)
Volume estimates for the singular sets of mean curvature flows
by: Fang, Hanbing, et al.
Published: (2025)
by: Fang, Hanbing, et al.
Published: (2025)
Strong uniqueness and rectifiability of generalized cylindrical singularities in Ricci flow
by: Fang, Hanbing, et al.
Published: (2026)
by: Fang, Hanbing, et al.
Published: (2026)
Strong uniqueness of tangent flows at cylindrical singularities in Ricci flow
by: Fang, Hanbing, et al.
Published: (2025)
by: Fang, Hanbing, et al.
Published: (2025)
Semantic-aware Representation Learning for Homography Estimation
by: Liu, Yuhan, et al.
Published: (2024)
by: Liu, Yuhan, et al.
Published: (2024)
IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?
by: Zhao, Xiaobei, et al.
Published: (2026)
by: Zhao, Xiaobei, et al.
Published: (2026)
SpaRC and SpaRP: Spatial Reasoning Characterization and Path Generation for Understanding Spatial Reasoning Capability of Large Language Models
by: Rizvi, Md Imbesat Hassan, et al.
Published: (2024)
by: Rizvi, Md Imbesat Hassan, et al.
Published: (2024)
SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence
by: Gong, Ziyang, et al.
Published: (2025)
by: Gong, Ziyang, et al.
Published: (2025)
TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation
by: Li, Dingbang, et al.
Published: (2024)
by: Li, Dingbang, et al.
Published: (2024)
FedSpaLLM: Federated Pruning of Large Language Models
by: Bai, Guangji, et al.
Published: (2024)
by: Bai, Guangji, et al.
Published: (2024)
SpaDA: A Spatial Dataflow Architecture Programming Language
by: Gianinazzi, Lukas, et al.
Published: (2025)
by: Gianinazzi, Lukas, et al.
Published: (2025)
MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation
by: Lin, Xi, et al.
Published: (2026)
by: Lin, Xi, et al.
Published: (2026)
Cascade Prompt Learning for Vision-Language Model Adaptation
by: Wu, Ge, et al.
Published: (2024)
by: Wu, Ge, et al.
Published: (2024)
SpaCE: The Spatial Confounding Environment
by: Tec, Mauricio, et al.
Published: (2023)
by: Tec, Mauricio, et al.
Published: (2023)
User-Feedback-Driven Adaptation for Vision-and-Language Navigation
by: Yu, Yongqiang, et al.
Published: (2025)
by: Yu, Yongqiang, et al.
Published: (2025)
BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation
by: Lyu, Wenqi, et al.
Published: (2025)
by: Lyu, Wenqi, et al.
Published: (2025)
VILTA: A VLM-in-the-Loop Adversary for Enhancing Driving Policy Robustness
by: Chen, Qimao, et al.
Published: (2026)
by: Chen, Qimao, et al.
Published: (2026)
RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real World
by: Liu, Hanbing, et al.
Published: (2026)
by: Liu, Hanbing, et al.
Published: (2026)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
Cluster-Aware Prompt Ensemble Learning for Few-Shot Vision-Language Model Adaptation
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
Optimal-Horizon Social Robot Navigation in Heterogeneous Crowds
by: Shi, Jiamin, et al.
Published: (2026)
by: Shi, Jiamin, et al.
Published: (2026)
Flatness Guided Test-Time Adaptation for Vision-Language Models
by: Li, Aodi, et al.
Published: (2025)
by: Li, Aodi, et al.
Published: (2025)
Similar Items
-
Think before Go: Hierarchical Reasoning for Image-goal Navigation
by: Li, Pengna, et al.
Published: (2026) -
Dual-Anchoring: Addressing State Drift in Vision-Language Navigation
by: Wu, Kangyi, et al.
Published: (2026) -
REGNav: Room Expert Guided Image-Goal Navigation
by: Li, Pengna, et al.
Published: (2025) -
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026) -
DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale
by: Zuo, Sicheng, et al.
Published: (2026)