Drift-Resistant Navigation World Model with Anchored Epipolar Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Luan, Po-Chien, Xia, Zimin, Li, Wuyang, Gao, Yang, Alahi, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Transmotion: Pre-trained Model for Human Motion Prediction
by: Gao, Yang, et al.
Published: (2024)
by: Gao, Yang, et al.
Published: (2024)
Unified Human Localization and Trajectory Prediction with Monocular Vision
by: Luan, Po-Chien, et al.
Published: (2025)
by: Luan, Po-Chien, et al.
Published: (2025)
Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models
by: Luan, Po-Chien, et al.
Published: (2026)
by: Luan, Po-Chien, et al.
Published: (2026)
Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
by: Li, Wuyang, et al.
Published: (2025)
by: Li, Wuyang, et al.
Published: (2025)
FG$^2$: Fine-Grained Cross-View Localization by Fine-Grained Feature Matching
by: Xia, Zimin, et al.
Published: (2025)
by: Xia, Zimin, et al.
Published: (2025)
RAP: 3D Rasterization Augmented End-to-End Planning
by: Feng, Lan, et al.
Published: (2025)
by: Feng, Lan, et al.
Published: (2025)
Sim-to-Real Causal Transfer: A Metric Learning Approach to Causally-Aware Interaction Representations
by: Rahimi, Ahmad, et al.
Published: (2023)
by: Rahimi, Ahmad, et al.
Published: (2023)
Anchored Video Generation: Decoupling Scene Construction and Temporal Synthesis in Text-to-Video Diffusion Models
by: Hassan, Mariam, et al.
Published: (2025)
by: Hassan, Mariam, et al.
Published: (2025)
EverAnimate: Minute-Scale Human Animation via Latent Flow Restoration
by: Li, Wuyang, et al.
Published: (2026)
by: Li, Wuyang, et al.
Published: (2026)
Loc$^2$: Interpretable Cross-View Localization via Depth-Lifted Local Feature Matching
by: Xia, Zimin, et al.
Published: (2025)
by: Xia, Zimin, et al.
Published: (2025)
Social-Transmotion: Promptable Human Trajectory Prediction
by: Saadatnejad, Saeed, et al.
Published: (2023)
by: Saadatnejad, Saeed, et al.
Published: (2023)
OmniTraj: Pre-Training on Heterogeneous Data for Adaptive and Zero-Shot Human Trajectory Prediction
by: Gao, Yang, et al.
Published: (2025)
by: Gao, Yang, et al.
Published: (2025)
Visual SLAM with DEM Anchoring for Lunar Surface Navigation
by: Dai, Adam, et al.
Published: (2026)
by: Dai, Adam, et al.
Published: (2026)
Relational Epipolar Graphs for Robust Relative Camera Pose Estimation
by: Rao, Prateeth, et al.
Published: (2026)
by: Rao, Prateeth, et al.
Published: (2026)
VoxDet: Rethinking 3D Semantic Occupancy Prediction as Dense Object Detection
by: Li, Wuyang, et al.
Published: (2025)
by: Li, Wuyang, et al.
Published: (2025)
Towards Motion Forecasting with Real-World Perception Inputs: Are End-to-End Approaches Competitive?
by: Xu, Yihong, et al.
Published: (2023)
by: Xu, Yihong, et al.
Published: (2023)
Epipolar Attention Field Transformers for Bird's Eye View Semantic Segmentation
by: Witte, Christian, et al.
Published: (2024)
by: Witte, Christian, et al.
Published: (2024)
World Guidance: World Modeling in Condition Space for Action Generation
by: Su, Yue, et al.
Published: (2026)
by: Su, Yue, et al.
Published: (2026)
Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
by: Zayene, Mehdi, et al.
Published: (2024)
by: Zayene, Mehdi, et al.
Published: (2024)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
by: Zhao, Baining, et al.
Published: (2026)
by: Zhao, Baining, et al.
Published: (2026)
Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation
by: Hassan, Mariam, et al.
Published: (2026)
by: Hassan, Mariam, et al.
Published: (2026)
GeoDistill: Geometry-Guided Self-Distillation for Weakly Supervised Cross-View Localization
by: Tong, Shaowen, et al.
Published: (2025)
by: Tong, Shaowen, et al.
Published: (2025)
HHI-Assist: A Dataset and Benchmark of Human-Human Interaction in Physical Assistance Scenario
by: Saadatnejad, Saeed, et al.
Published: (2025)
by: Saadatnejad, Saeed, et al.
Published: (2025)
DynamicGlue: Epipolar and Time-Informed Data Association in Dynamic Environments using Graph Neural Networks
by: Huber, Theresa, et al.
Published: (2024)
by: Huber, Theresa, et al.
Published: (2024)
Forecast-PEFT: Parameter-Efficient Fine-Tuning for Pre-trained Motion Forecasting Models
by: Wang, Jifeng, et al.
Published: (2024)
by: Wang, Jifeng, et al.
Published: (2024)
VISTAv2: World Imagination for Indoor Vision-and-Language Navigation
by: Huang, Yanjia, et al.
Published: (2025)
by: Huang, Yanjia, et al.
Published: (2025)
Certified Human Trajectory Prediction
by: Bahari, Mohammadhossein, et al.
Published: (2024)
by: Bahari, Mohammadhossein, et al.
Published: (2024)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
by: Chen, Hongjin, et al.
Published: (2026)
by: Chen, Hongjin, et al.
Published: (2026)
WorldRFT: Latent World Model Planning with Reinforcement Fine-Tuning for Autonomous Driving
by: Yang, Pengxuan, et al.
Published: (2025)
by: Yang, Pengxuan, et al.
Published: (2025)
WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation
by: Nie, Dujun, et al.
Published: (2025)
by: Nie, Dujun, et al.
Published: (2025)
A Multi-Loss Strategy for Vehicle Trajectory Prediction: Combining Off-Road, Diversity, and Directional Consistency Losses
by: Rahimi, Ahmad, et al.
Published: (2024)
by: Rahimi, Ahmad, et al.
Published: (2024)
SCENES: Subpixel Correspondence Estimation With Epipolar Supervision
by: Kloepfer, Dominik A., et al.
Published: (2024)
by: Kloepfer, Dominik A., et al.
Published: (2024)
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
by: Ye, Weicai, et al.
Published: (2024)
by: Ye, Weicai, et al.
Published: (2024)
PAPL-SLAM: Principal Axis-Anchored Monocular Point-Line SLAM
by: Li, Guanghao, et al.
Published: (2024)
by: Li, Guanghao, et al.
Published: (2024)
Sparse Video Generation Propels Real-World Beyond-the-View Vision-Language Navigation
by: Zhang, Hai, et al.
Published: (2026)
by: Zhang, Hai, et al.
Published: (2026)
Towards Real-World Aerial Vision Guidance with Categorical 6D Pose Tracker
by: Sun, Jingtao, et al.
Published: (2024)
by: Sun, Jingtao, et al.
Published: (2024)
Causal World Modeling for Robot Control
by: Li, Lin, et al.
Published: (2026)
by: Li, Lin, et al.
Published: (2026)
POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation
by: Gong, Ruiyan, et al.
Published: (2026)
by: Gong, Ruiyan, et al.
Published: (2026)
Language-Conditioned World Modeling for Visual Navigation
by: Dong, Yifei, et al.
Published: (2026)
by: Dong, Yifei, et al.
Published: (2026)
Enhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour Videos
by: Xu, Haoxuan, et al.
Published: (2026)
by: Xu, Haoxuan, et al.
Published: (2026)
Similar Items
-
Multi-Transmotion: Pre-trained Model for Human Motion Prediction
by: Gao, Yang, et al.
Published: (2024) -
Unified Human Localization and Trajectory Prediction with Monocular Vision
by: Luan, Po-Chien, et al.
Published: (2025) -
Social-Mamba: Socially-Aware Trajectory Forecasting with State-Space Models
by: Luan, Po-Chien, et al.
Published: (2026) -
Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
by: Li, Wuyang, et al.
Published: (2025) -
FG$^2$: Fine-Grained Cross-View Localization by Fine-Grained Feature Matching
by: Xia, Zimin, et al.
Published: (2025)