Saved in:
| Main Authors: | Zhang, Arthur, Sikchi, Harshit, Zhang, Amy, Biswas, Joydeep |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.03921 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lift, Splat, Map: Lifting Foundation Masks for Label-Free Semantic Scene Completion
by: Zhang, Arthur, et al.
Published: (2024)
by: Zhang, Arthur, et al.
Published: (2024)
IFG: Internet-Scale Guidance for Functional Grasping Generation
by: Liu, Ray Muxin, et al.
Published: (2025)
by: Liu, Ray Muxin, et al.
Published: (2025)
ComposableNav: Instruction-Following Navigation in Dynamic Environments via Composable Diffusion
by: Hu, Zichao, et al.
Published: (2025)
by: Hu, Zichao, et al.
Published: (2025)
Image-Goal Navigation Using Refined Feature Guidance and Scene Graph Enhancement
by: Feng, Zhicheng, et al.
Published: (2025)
by: Feng, Zhicheng, et al.
Published: (2025)
MPVO: Motion-Prior based Visual Odometry for PointGoal Navigation
by: Paul, Sayan, et al.
Published: (2024)
by: Paul, Sayan, et al.
Published: (2024)
VENTURA: Adapting Image Diffusion Models for Unified Task Conditioned Navigation
by: Zhang, Arthur, et al.
Published: (2025)
by: Zhang, Arthur, et al.
Published: (2025)
Virtual Guidance as a Mid-level Representation for Navigation with Augmented Reality
by: Yang, Hsuan-Kung, et al.
Published: (2023)
by: Yang, Hsuan-Kung, et al.
Published: (2023)
MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
by: Kim, Yejin, et al.
Published: (2026)
by: Kim, Yejin, et al.
Published: (2026)
General Flow as Foundation Affordance for Scalable Robot Learning
by: Yuan, Chengbo, et al.
Published: (2024)
by: Yuan, Chengbo, et al.
Published: (2024)
MOSE: Monocular Semantic Reconstruction Using NeRF-Lifted Noisy Priors
by: Du, Zhenhua, et al.
Published: (2024)
by: Du, Zhenhua, et al.
Published: (2024)
MemoNav: Working Memory Model for Visual Navigation
by: Li, Hongxin, et al.
Published: (2024)
by: Li, Hongxin, et al.
Published: (2024)
OctoNav: Towards Generalist Embodied Navigation
by: Gao, Chen, et al.
Published: (2025)
by: Gao, Chen, et al.
Published: (2025)
Towards Autonomous Micromobility through Scalable Urban Simulation
by: Wu, Wayne, et al.
Published: (2025)
by: Wu, Wayne, et al.
Published: (2025)
Large Trajectory Models are Scalable Motion Predictors and Planners
by: Sun, Qiao, et al.
Published: (2023)
by: Sun, Qiao, et al.
Published: (2023)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
by: He, Yu, et al.
Published: (2025)
by: He, Yu, et al.
Published: (2025)
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
by: Guo, Jun, et al.
Published: (2026)
by: Guo, Jun, et al.
Published: (2026)
Tag Map: A Text-Based Map for Spatial Reasoning and Navigation with Large Language Models
by: Zhang, Mike, et al.
Published: (2024)
by: Zhang, Mike, et al.
Published: (2024)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
MapDream: Task-Driven Map Learning for Vision-Language Navigation
by: Lian, Guoxin, et al.
Published: (2026)
by: Lian, Guoxin, et al.
Published: (2026)
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
by: Qi, Carl, et al.
Published: (2024)
by: Qi, Carl, et al.
Published: (2024)
RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation
by: Han, Mingfei, et al.
Published: (2024)
by: Han, Mingfei, et al.
Published: (2024)
ActiveVLN: Towards Active Exploration via Multi-Turn RL in Vision-and-Language Navigation
by: Zhang, Zekai, et al.
Published: (2025)
by: Zhang, Zekai, et al.
Published: (2025)
Enhanced Safety in Autonomous Driving: Integrating Latent State Diffusion Model for End-to-End Navigation
by: Chu, Detian, et al.
Published: (2024)
by: Chu, Detian, et al.
Published: (2024)
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
Embodied Navigation at the Art Gallery
by: Bigazzi, Roberto, et al.
Published: (2022)
by: Bigazzi, Roberto, et al.
Published: (2022)
Seeing Through Uncertainty: A Free-Energy Approach for Real-Time Perceptual Adaptation in Robust Visual Navigation
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
Leveraging Unknown Objects to Construct Labeled-Unlabeled Meta-Relationships for Zero-Shot Object Navigation
by: Zheng, Yanwei, et al.
Published: (2024)
by: Zheng, Yanwei, et al.
Published: (2024)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
by: Jin, Qunchao, et al.
Published: (2025)
by: Jin, Qunchao, et al.
Published: (2025)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
by: Zhong, Linqing, et al.
Published: (2024)
by: Zhong, Linqing, et al.
Published: (2024)
Learning 3D Robotics Perception using Inductive Priors
by: Irshad, Muhammad Zubair
Published: (2024)
by: Irshad, Muhammad Zubair
Published: (2024)
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
Language-Conditioned World Modeling for Visual Navigation
by: Dong, Yifei, et al.
Published: (2026)
by: Dong, Yifei, et al.
Published: (2026)
SocialNav: Training Human-Inspired Foundation Model for Socially-Aware Embodied Navigation
by: Chen, Ziyi, et al.
Published: (2025)
by: Chen, Ziyi, et al.
Published: (2025)
What Limits Vision-and-Language Navigation ?
by: Wang, Yunheng, et al.
Published: (2026)
by: Wang, Yunheng, et al.
Published: (2026)
ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
by: Chu, Zedong, et al.
Published: (2026)
by: Chu, Zedong, et al.
Published: (2026)
SLAG: Scalable Language-Augmented Gaussian Splatting
by: Szilagyi, Laszlo, et al.
Published: (2025)
by: Szilagyi, Laszlo, et al.
Published: (2025)
Dynamic Sensor Matching based on Geomagnetic Inertial Navigation
by: Müller, Simone, et al.
Published: (2022)
by: Müller, Simone, et al.
Published: (2022)
Vision-Language Navigation with Embodied Intelligence: A Survey
by: Gao, Peng, et al.
Published: (2024)
by: Gao, Peng, et al.
Published: (2024)
A Navigation Framework Utilizing Vision-Language Models
by: Duan, Yicheng, et al.
Published: (2025)
by: Duan, Yicheng, et al.
Published: (2025)
AgriVLN: Vision-and-Language Navigation for Agricultural Robots
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
Similar Items
-
Lift, Splat, Map: Lifting Foundation Masks for Label-Free Semantic Scene Completion
by: Zhang, Arthur, et al.
Published: (2024) -
IFG: Internet-Scale Guidance for Functional Grasping Generation
by: Liu, Ray Muxin, et al.
Published: (2025) -
ComposableNav: Instruction-Following Navigation in Dynamic Environments via Composable Diffusion
by: Hu, Zichao, et al.
Published: (2025) -
Image-Goal Navigation Using Refined Feature Guidance and Scene Graph Enhancement
by: Feng, Zhicheng, et al.
Published: (2025) -
MPVO: Motion-Prior based Visual Odometry for PointGoal Navigation
by: Paul, Sayan, et al.
Published: (2024)