OmniVLN: Omnidirectional 3D Perception and Token-Efficient LLM Reasoning for Visual-Language Navigation across Air and Ground Platforms
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Zhongyuang, He, Min, Yu, Shaonan, Xu, Xinhang, Cao, Muqing, Li, Jianping, Yang, Jianfei, Xie, Lihua |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AirCrab: A Hybrid Aerial-Ground Manipulator with An Active Wheel
by: Cao, Muqing, et al.
Published: (2024)
by: Cao, Muqing, et al.
Published: (2024)
AToM: Adaptive Theory-of-Mind-Based Human Motion Prediction in Long-Term Human-Robot Interactions
by: Liao, Yuwen, et al.
Published: (2025)
by: Liao, Yuwen, et al.
Published: (2025)
AEOS: Active Environment-aware Optimal Scanning Control for UAV LiDAR-Inertial Odometry in Complex Scenes
by: Li, Jianping, et al.
Published: (2025)
by: Li, Jianping, et al.
Published: (2025)
Learning Energy-Efficient Air--Ground Actuation for Hybrid Robots on Stair-Like Terrain
by: Li, Jiaxing, et al.
Published: (2026)
by: Li, Jiaxing, et al.
Published: (2026)
AV-PedAware: Self-Supervised Audio-Visual Fusion for Dynamic Pedestrian Awareness
by: Yang, Yizhuo, et al.
Published: (2024)
by: Yang, Yizhuo, et al.
Published: (2024)
Following Is All You Need: Robot Crowd Navigation Using People As Planners
by: Liao, Yuwen, et al.
Published: (2025)
by: Liao, Yuwen, et al.
Published: (2025)
Learning Dynamic Weight Adjustment for Spatial-Temporal Trajectory Planning in Crowd Navigation
by: Cao, Muqing, et al.
Published: (2024)
by: Cao, Muqing, et al.
Published: (2024)
VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
S3KF: Spherical State-Space Kalman Filtering for Panoramic 3D Multi-Object Tracking
by: Liu, Zhongyuan, et al.
Published: (2026)
by: Liu, Zhongyuan, et al.
Published: (2026)
Handle Object Navigation as Weighted Traveling Repairman Problem
by: Liu, Ruimeng, et al.
Published: (2025)
by: Liu, Ruimeng, et al.
Published: (2025)
A Cost-Effective Cooperative Exploration and Inspection Strategy for Heterogeneous Aerial System
by: Xu, Xinhang, et al.
Published: (2024)
by: Xu, Xinhang, et al.
Published: (2024)
UA-MPC: Uncertainty-Aware Model Predictive Control for Motorized LiDAR Odometry
by: Li, Jianping, et al.
Published: (2024)
by: Li, Jianping, et al.
Published: (2024)
HelmetPoser: A Helmet-Mounted IMU Dataset for Data-Driven Estimation of Human Head Motion in Diverse Conditions
by: Li, Jianping, et al.
Published: (2024)
by: Li, Jianping, et al.
Published: (2024)
HCTO: Optimality-Aware LiDAR Inertial Odometry with Hybrid Continuous Time Optimization for Compact Wearable Mapping System
by: Li, Jianping, et al.
Published: (2024)
by: Li, Jianping, et al.
Published: (2024)
OmniNxt: A Fully Open-source and Compact Aerial Robot with Omnidirectional Visual Perception
by: Liu, Peize, et al.
Published: (2024)
by: Liu, Peize, et al.
Published: (2024)
Spatial-VLN: Zero-Shot Vision-and-Language Navigation With Explicit Spatial Perception and Exploration
by: Yue, Lu, et al.
Published: (2026)
by: Yue, Lu, et al.
Published: (2026)
Omni-Perception: Omnidirectional Collision Avoidance for Legged Locomotion in Dynamic Environments
by: Wang, Zifan, et al.
Published: (2025)
by: Wang, Zifan, et al.
Published: (2025)
Omni$^2$: Unifying Omnidirectional Image Generation and Editing in an Omni Model
by: Yang, Liu, et al.
Published: (2025)
by: Yang, Liu, et al.
Published: (2025)
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
by: Guo, Wenxuan, et al.
Published: (2026)
by: Guo, Wenxuan, et al.
Published: (2026)
Efficient-VLN: A Training-Efficient Vision-Language Navigation Model
by: Zheng, Duo, et al.
Published: (2025)
by: Zheng, Duo, et al.
Published: (2025)
DecoVLN: Decoupling Observation, Reasoning, and Correction for Vision-and-Language Navigation
by: Xin, Zihao, et al.
Published: (2026)
by: Xin, Zihao, et al.
Published: (2026)
Graph Optimality-Aware Stochastic LiDAR Bundle Adjustment with Progressive Spatial Smoothing
by: Li, Jianping, et al.
Published: (2024)
by: Li, Jianping, et al.
Published: (2024)
VLN-MME: Diagnosing MLLMs as Language-guided Visual Navigation agents
by: Zhao, Xunyi, et al.
Published: (2025)
by: Zhao, Xunyi, et al.
Published: (2025)
Eigen Is All You Need: Efficient Lidar-Inertial Continuous-Time Odometry with Internal Association
by: Nguyen, Thien-Minh, et al.
Published: (2024)
by: Nguyen, Thien-Minh, et al.
Published: (2024)
Topological Motion Planning Diffusion: Generative Tangle-Free Path Planning for Tethered Robots in Obstacle-Rich Environments
by: Tian, Yifu, et al.
Published: (2026)
by: Tian, Yifu, et al.
Published: (2026)
HoloLLM: Multisensory Foundation Model for Language-Grounded Human Sensing and Reasoning
by: Zhou, Chuhao, et al.
Published: (2025)
by: Zhou, Chuhao, et al.
Published: (2025)
Omni Differential Drive for Simultaneous Reconfiguration and Omnidirectional Mobility of Wheeled Robots
by: Zhao, Ziqi, et al.
Published: (2024)
by: Zhao, Ziqi, et al.
Published: (2024)
MiniVLN: Efficient Vision-and-Language Navigation by Progressive Knowledge Distillation
by: Zhu, Junyou, et al.
Published: (2024)
by: Zhu, Junyou, et al.
Published: (2024)
FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-Language Navigation
by: Zuo, Jing, et al.
Published: (2026)
by: Zuo, Jing, et al.
Published: (2026)
DV-VLN: Dual Verification for Reliable LLM-Based Vision-and-Language Navigation
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
OmniLocalRF: Omnidirectional Local Radiance Fields from Dynamic Videos
by: Choi, Dongyoung, et al.
Published: (2024)
by: Choi, Dongyoung, et al.
Published: (2024)
Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification
by: Wen, Jiawen, et al.
Published: (2026)
by: Wen, Jiawen, et al.
Published: (2026)
ViSA-Enhanced Aerial VLN: A Visual-Spatial Reasoning Enhanced Framework for Aerial Vision-Language Navigation
by: Tong, Haoyu, et al.
Published: (2026)
by: Tong, Haoyu, et al.
Published: (2026)
OmniDP: Beyond-FOV Large-Workspace Humanoid Manipulation with Omnidirectional 3D Perception
by: Qu, Pei, et al.
Published: (2026)
by: Qu, Pei, et al.
Published: (2026)
ODD: Omni Differential Drive for Simultaneous Reconfiguration and Omnidirectional Mobility of Wheeled Robots
by: Zhao, Ziqi, et al.
Published: (2024)
by: Zhao, Ziqi, et al.
Published: (2024)
AirSLAM: An Efficient and Illumination-Robust Point-Line Visual SLAM System
by: Xu, Kuan, et al.
Published: (2024)
by: Xu, Kuan, et al.
Published: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
by: Chung, Jiwan, et al.
Published: (2025)
by: Chung, Jiwan, et al.
Published: (2025)
GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation
by: Yang, Jiahao, et al.
Published: (2026)
by: Yang, Jiahao, et al.
Published: (2026)
Zero-Shot Open-Vocabulary Human Motion Grounding with Test-Time Training
by: Zhou, Yunjiao, et al.
Published: (2025)
by: Zhou, Yunjiao, et al.
Published: (2025)
AdaVLN: Towards Visual Language Navigation in Continuous Indoor Environments with Moving Humans
by: Loh, Dillon, et al.
Published: (2024)
by: Loh, Dillon, et al.
Published: (2024)
Similar Items
-
AirCrab: A Hybrid Aerial-Ground Manipulator with An Active Wheel
by: Cao, Muqing, et al.
Published: (2024) -
AToM: Adaptive Theory-of-Mind-Based Human Motion Prediction in Long-Term Human-Robot Interactions
by: Liao, Yuwen, et al.
Published: (2025) -
AEOS: Active Environment-aware Optimal Scanning Control for UAV LiDAR-Inertial Odometry in Complex Scenes
by: Li, Jianping, et al.
Published: (2025) -
Learning Energy-Efficient Air--Ground Actuation for Hybrid Robots on Stair-Like Terrain
by: Li, Jiaxing, et al.
Published: (2026) -
AV-PedAware: Self-Supervised Audio-Visual Fusion for Dynamic Pedestrian Awareness
by: Yang, Yizhuo, et al.
Published: (2024)