Empowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Wentao, Chen, Xuweiyi, Rajagopal, Vignesh, Chen, Jeffrey, Chandra, Rohan, Cheng, Zezhou |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
by: Chen, Xuweiyi, et al.
Published: (2026)
by: Chen, Xuweiyi, et al.
Published: (2026)
Probing the Mid-level Vision Capabilities of Self-Supervised Learning
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
by: Chen, Xuweiyi, et al.
Published: (2025)
by: Chen, Xuweiyi, et al.
Published: (2025)
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
by: Wang, Boyang, et al.
Published: (2025)
by: Wang, Boyang, et al.
Published: (2025)
Open Vocabulary Monocular 3D Object Detection
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
by: Chen, Xuweiyi, et al.
Published: (2025)
by: Chen, Xuweiyi, et al.
Published: (2025)
StereoVGGT: A Training-Free Visual Geometry Transformer for Stereo Vision
by: Chen, Ziyang, et al.
Published: (2026)
by: Chen, Ziyang, et al.
Published: (2026)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
by: Xia, Tian, et al.
Published: (2024)
by: Xia, Tian, et al.
Published: (2024)
All-in-One: Transferring Vision Foundation Models into Stereo Matching
by: Zhou, Jingyi, et al.
Published: (2024)
by: Zhou, Jingyi, et al.
Published: (2024)
Next-Embedding Prediction Makes Strong Vision Learners
by: Xu, Sihan, et al.
Published: (2025)
by: Xu, Sihan, et al.
Published: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
by: Wang, Zihan, et al.
Published: (2025)
by: Wang, Zihan, et al.
Published: (2025)
VDNeRF: Vision-only Dynamic Neural Radiance Field for Urban Scenes
by: Zou, Zhengyu, et al.
Published: (2025)
by: Zou, Zhengyu, et al.
Published: (2025)
Eyes on the Streets: Leveraging Street-Level Imaging to Model Urban Crime Dynamics
by: Qi, Zhixuan, et al.
Published: (2024)
by: Qi, Zhixuan, et al.
Published: (2024)
Dual-Level Precision Edges Guided Multi-View Stereo with Accurate Planarization
by: Chen, Kehua, et al.
Published: (2024)
by: Chen, Kehua, et al.
Published: (2024)
SemStereo: Semantic-Constrained Stereo Matching Network for Remote Sensing
by: Chen, Chen, et al.
Published: (2024)
by: Chen, Chen, et al.
Published: (2024)
ZeroStereo: Zero-shot Stereo Matching from Single Images
by: Wang, Xianqi, et al.
Published: (2025)
by: Wang, Xianqi, et al.
Published: (2025)
The Role of Cyclopean-Eye in Stereo Vision
by: da Silva, Sherlon Almeida, et al.
Published: (2025)
by: da Silva, Sherlon Almeida, et al.
Published: (2025)
WakeupUrban: Unsupervised Semantic Segmentation of Mid-20$^{th}$ century Urban Landscapes with Satellite Imagery
by: Hao, Tianxiang, et al.
Published: (2025)
by: Hao, Tianxiang, et al.
Published: (2025)
Cross-Modal Urban Sensing: Evaluating Sound-Vision Alignment Across Street-Level and Aerial Imagery
by: Chen, Pengyu, et al.
Published: (2025)
by: Chen, Pengyu, et al.
Published: (2025)
MLG-Stereo: ViT Based Stereo Matching with Multi-Stage Local-Global Enhancement
by: Zhang, Haoyu, et al.
Published: (2026)
by: Zhang, Haoyu, et al.
Published: (2026)
Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching
by: Jing, Junpeng, et al.
Published: (2024)
by: Jing, Junpeng, et al.
Published: (2024)
StereoDiff: Stereo-Diffusion Synergy for Video Depth Estimation
by: Li, Haodong, et al.
Published: (2025)
by: Li, Haodong, et al.
Published: (2025)
Fisheye Stereo Vision: Depth and Range Error
by: Jiang, Leaf, et al.
Published: (2026)
by: Jiang, Leaf, et al.
Published: (2026)
Non-Learning Low-Light Stereo Vision
by: Wang, Jason, et al.
Published: (2026)
by: Wang, Jason, et al.
Published: (2026)
StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors
by: Shen, Guibao, et al.
Published: (2025)
by: Shen, Guibao, et al.
Published: (2025)
SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation
by: Wang, Yun, et al.
Published: (2026)
by: Wang, Yun, et al.
Published: (2026)
Pip-Stereo: Progressive Iterations Pruner for Iterative Optimization based Stereo Matching
by: Zheng, Jintu, et al.
Published: (2026)
by: Zheng, Jintu, et al.
Published: (2026)
Stereo-GS: Multi-View Stereo Vision Model for Generalizable 3D Gaussian Splatting Reconstruction
by: Huang, Xiufeng, et al.
Published: (2025)
by: Huang, Xiufeng, et al.
Published: (2025)
Adapting Stereo Vision From Objects To 3D Lunar Surface Reconstruction with the StereoLunar Dataset
by: Grethen, Clementine, et al.
Published: (2025)
by: Grethen, Clementine, et al.
Published: (2025)
DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos
by: Huang, Yuan, et al.
Published: (2026)
by: Huang, Yuan, et al.
Published: (2026)
MoCha-Stereo: Motif Channel Attention Network for Stereo Matching
by: Chen, Ziyang, et al.
Published: (2024)
by: Chen, Ziyang, et al.
Published: (2024)
Affine Correspondences in Stereo Vision: Theory, Practice, and Limitations
by: Hajder, Levente
Published: (2026)
by: Hajder, Levente
Published: (2026)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
by: Mittal, Mayank, et al.
Published: (2019)
by: Mittal, Mayank, et al.
Published: (2019)
Mono2Stereo: A Benchmark and Empirical Study for Stereo Conversion
by: Yu, Songsong, et al.
Published: (2025)
by: Yu, Songsong, et al.
Published: (2025)
OpenStereo: A Comprehensive Benchmark for Stereo Matching and Strong Baseline
by: Guo, Xianda, et al.
Published: (2023)
by: Guo, Xianda, et al.
Published: (2023)
Playing to Vision Foundation Model's Strengths in Stereo Matching
by: Liu, Chuang-Wei, et al.
Published: (2024)
by: Liu, Chuang-Wei, et al.
Published: (2024)
StereoWorld: Geometry-Aware Monocular-to-Stereo Video Generation
by: Xing, Ke, et al.
Published: (2025)
by: Xing, Ke, et al.
Published: (2025)
Multi-Object Hallucination in Vision-Language Models
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
StereoCarla: A High-Fidelity Driving Dataset for Generalizable Stereo
by: Guo, Xianda, et al.
Published: (2025)
by: Guo, Xianda, et al.
Published: (2025)
Similar Items
-
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
by: Chen, Xuweiyi, et al.
Published: (2026) -
Probing the Mid-level Vision Capabilities of Self-Supervised Learning
by: Chen, Xuweiyi, et al.
Published: (2024) -
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
by: Chen, Xuweiyi, et al.
Published: (2024) -
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
by: Chen, Xuweiyi, et al.
Published: (2025) -
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
by: Wang, Boyang, et al.
Published: (2025)