FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Yokoyama, Naoki, Ha, Sehoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
by: Yokoyama, Naoki, et al.
Published: (2024)
by: Yokoyama, Naoki, et al.
Published: (2024)
PPF: Pre-training and Preservative Fine-tuning of Humanoid Locomotion via Model-Assumption-based Regularization
by: Jung, Hyunyoung, et al.
Published: (2025)
by: Jung, Hyunyoung, et al.
Published: (2025)
STSM-FiLM: A FiLM-Conditioned Neural Architecture for Time-Scale Modification of Speech
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
by: Serpiva, Valerii, et al.
Published: (2026)
by: Serpiva, Valerii, et al.
Published: (2026)
SoraNav: Adaptive UAV Task-Centric Navigation via Zeroshot VLM Reasoning
by: Song, Hongyu, et al.
Published: (2025)
by: Song, Hongyu, et al.
Published: (2025)
Safe Navigation of Bipedal Robots via Koopman Operator-Based Model Predictive Control
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models
by: Song, Daeun, et al.
Published: (2024)
by: Song, Daeun, et al.
Published: (2024)
BAT: Balancing Agility and Stability via Online Policy Switching for Long-Horizon Whole-Body Humanoid Control
by: Baek, Donghoon, et al.
Published: (2026)
by: Baek, Donghoon, et al.
Published: (2026)
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
by: Cho, Daesol, et al.
Published: (2026)
by: Cho, Daesol, et al.
Published: (2026)
E-SocialNav: Efficient Socially Compliant Navigation with Language Models
by: Xiao, Ling, et al.
Published: (2026)
by: Xiao, Ling, et al.
Published: (2026)
FlowNav: Combining Flow Matching and Depth Priors for Efficient Navigation
by: Gode, Samiran, et al.
Published: (2024)
by: Gode, Samiran, et al.
Published: (2024)
StepNav: Structured Trajectory Priors for Efficient and Multimodal Visual Navigation
by: Luo, Xubo, et al.
Published: (2026)
by: Luo, Xubo, et al.
Published: (2026)
MoReFlow: Motion Retargeting Learning through Unsupervised Flow Matching
by: Kim, Wontaek, et al.
Published: (2025)
by: Kim, Wontaek, et al.
Published: (2025)
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
by: Abreu, Wallace, et al.
Published: (2026)
by: Abreu, Wallace, et al.
Published: (2026)
LOG-Nav: Efficient Layout-Aware Object-Goal Navigation with Hierarchical Planning
by: Hou, Jiawei, et al.
Published: (2025)
by: Hou, Jiawei, et al.
Published: (2025)
DRIVE-Nav: Directional Reasoning, Inspection, and Verification for Efficient Open-Vocabulary Navigation
by: Gao, Maoguo, et al.
Published: (2026)
by: Gao, Maoguo, et al.
Published: (2026)
Head-Pose-Aware Visual Speech Recognition with FiLM Modulation
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
by: Yang, Zebin, et al.
Published: (2025)
by: Yang, Zebin, et al.
Published: (2025)
Hydra-Nav: Object Navigation via Adaptive Dual-Process Reasoning
by: Wang, Zixuan, et al.
Published: (2026)
by: Wang, Zixuan, et al.
Published: (2026)
A Design Co-Pilot for Task-Tailored Manipulators
by: Külz, Jonathan, et al.
Published: (2025)
by: Külz, Jonathan, et al.
Published: (2025)
Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
by: Liu, Yihong, et al.
Published: (2025)
by: Liu, Yihong, et al.
Published: (2025)
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals
by: Yang, Chieh-Chi, et al.
Published: (2026)
by: Yang, Chieh-Chi, et al.
Published: (2026)
TopoNav: Topological Navigation for Efficient Exploration in Sparse Reward Environments
by: Hossain, Jumman, et al.
Published: (2024)
by: Hossain, Jumman, et al.
Published: (2024)
SEA-Nav: Efficient Policy Learning for Safe and Agile Quadruped Navigation in Cluttered Environments
by: Chen, Shiyi, et al.
Published: (2026)
by: Chen, Shiyi, et al.
Published: (2026)
Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface
by: Wang, Dewei, et al.
Published: (2025)
by: Wang, Dewei, et al.
Published: (2025)
The NavINST Dataset for Multi-Sensor Autonomous Navigation
by: de Araujo, Paulo Ricardo Marques, et al.
Published: (2025)
by: de Araujo, Paulo Ricardo Marques, et al.
Published: (2025)
FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation
by: Shao, Dian, et al.
Published: (2026)
by: Shao, Dian, et al.
Published: (2026)
NavOL: Navigation Policy with Online Imitation Learning
by: Wei, Xiaofei, et al.
Published: (2026)
by: Wei, Xiaofei, et al.
Published: (2026)
Learning Physical Interaction Skills from Human Demonstrations
by: Li, Tianyu, et al.
Published: (2025)
by: Li, Tianyu, et al.
Published: (2025)
Partial Motion Imitation for Learning Cart Pushing with Legged Manipulators
by: Das, Mili, et al.
Published: (2026)
by: Das, Mili, et al.
Published: (2026)
STRIVE: Structured Representation Integrating VLM Reasoning for Efficient Object Navigation
by: Zhu, Haokun, et al.
Published: (2025)
by: Zhu, Haokun, et al.
Published: (2025)
VLM-Empowered Multi-Mode System for Efficient and Safe Planetary Navigation
by: Cheng, Sinuo, et al.
Published: (2025)
by: Cheng, Sinuo, et al.
Published: (2025)
Deep Reinforcement Learning for Multi-Agent Coordination
by: Aina, Kehinde O., et al.
Published: (2025)
by: Aina, Kehinde O., et al.
Published: (2025)
MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments
by: Sima, Kuankuan, et al.
Published: (2025)
by: Sima, Kuankuan, et al.
Published: (2025)
RobotDesignGPT: Automated Robot Design Synthesis using Vision Language Models
by: Sontakke, Nitish, et al.
Published: (2026)
by: Sontakke, Nitish, et al.
Published: (2026)
ImagiNav: Scalable Embodied Navigation via Generative Visual Prediction and Inverse Dynamics
by: Chen, Jie, et al.
Published: (2026)
by: Chen, Jie, et al.
Published: (2026)
SFCo-Nav: Efficient Zero-Shot Visual Language Navigation via Collaboration of Slow LLM and Fast Attributed Graph Alignment
by: Xiong, Chaoran, et al.
Published: (2026)
by: Xiong, Chaoran, et al.
Published: (2026)
DynaNav: Dynamic Feature and Layer Selection for Efficient Visual Navigation
by: Wang, Jiahui, et al.
Published: (2025)
by: Wang, Jiahui, et al.
Published: (2025)
MOFM-Nav: On-Manifold Ordering-Flexible Multi-Robot Navigation
by: Hu, Bin-Bin, et al.
Published: (2025)
by: Hu, Bin-Bin, et al.
Published: (2025)
AdaNav: Adaptive Reasoning with Uncertainty for Vision-Language Navigation
by: Ding, Xin, et al.
Published: (2025)
by: Ding, Xin, et al.
Published: (2025)
Similar Items
-
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
by: Yokoyama, Naoki, et al.
Published: (2024) -
PPF: Pre-training and Preservative Fine-tuning of Humanoid Locomotion via Model-Assumption-based Regularization
by: Jung, Hyunyoung, et al.
Published: (2025) -
STSM-FiLM: A FiLM-Conditioned Neural Architecture for Time-Scale Modification of Speech
by: Wisnu, Dyah A. M. G., et al.
Published: (2025) -
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
by: Serpiva, Valerii, et al.
Published: (2026) -
SoraNav: Adaptive UAV Task-Centric Navigation via Zeroshot VLM Reasoning
by: Song, Hongyu, et al.
Published: (2025)