Empowering Dynamic Urban Navigation with Stereo and Mid-Level Vision
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhou, Wentao, Chen, Xuweiyi, Rajagopal, Vignesh, Chen, Jeffrey, Chandra, Rohan, Cheng, Zezhou |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
par: Chen, Xuweiyi, et autres
Publié: (2026)
par: Chen, Xuweiyi, et autres
Publié: (2026)
Probing the Mid-level Vision Capabilities of Self-Supervised Learning
par: Chen, Xuweiyi, et autres
Publié: (2024)
par: Chen, Xuweiyi, et autres
Publié: (2024)
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
par: Chen, Xuweiyi, et autres
Publié: (2024)
par: Chen, Xuweiyi, et autres
Publié: (2024)
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
par: Chen, Xuweiyi, et autres
Publié: (2025)
par: Chen, Xuweiyi, et autres
Publié: (2025)
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
par: Wang, Boyang, et autres
Publié: (2025)
par: Wang, Boyang, et autres
Publié: (2025)
Open Vocabulary Monocular 3D Object Detection
par: Yao, Jin, et autres
Publié: (2024)
par: Yao, Jin, et autres
Publié: (2024)
SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
par: Chen, Xuweiyi, et autres
Publié: (2025)
par: Chen, Xuweiyi, et autres
Publié: (2025)
StereoVGGT: A Training-Free Visual Geometry Transformer for Stereo Vision
par: Chen, Ziyang, et autres
Publié: (2026)
par: Chen, Ziyang, et autres
Publié: (2026)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
par: Xia, Tian, et autres
Publié: (2024)
par: Xia, Tian, et autres
Publié: (2024)
All-in-One: Transferring Vision Foundation Models into Stereo Matching
par: Zhou, Jingyi, et autres
Publié: (2024)
par: Zhou, Jingyi, et autres
Publié: (2024)
Next-Embedding Prediction Makes Strong Vision Learners
par: Xu, Sihan, et autres
Publié: (2025)
par: Xu, Sihan, et autres
Publié: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
par: Wang, Zihan, et autres
Publié: (2025)
par: Wang, Zihan, et autres
Publié: (2025)
VDNeRF: Vision-only Dynamic Neural Radiance Field for Urban Scenes
par: Zou, Zhengyu, et autres
Publié: (2025)
par: Zou, Zhengyu, et autres
Publié: (2025)
Eyes on the Streets: Leveraging Street-Level Imaging to Model Urban Crime Dynamics
par: Qi, Zhixuan, et autres
Publié: (2024)
par: Qi, Zhixuan, et autres
Publié: (2024)
Dual-Level Precision Edges Guided Multi-View Stereo with Accurate Planarization
par: Chen, Kehua, et autres
Publié: (2024)
par: Chen, Kehua, et autres
Publié: (2024)
SemStereo: Semantic-Constrained Stereo Matching Network for Remote Sensing
par: Chen, Chen, et autres
Publié: (2024)
par: Chen, Chen, et autres
Publié: (2024)
ZeroStereo: Zero-shot Stereo Matching from Single Images
par: Wang, Xianqi, et autres
Publié: (2025)
par: Wang, Xianqi, et autres
Publié: (2025)
The Role of Cyclopean-Eye in Stereo Vision
par: da Silva, Sherlon Almeida, et autres
Publié: (2025)
par: da Silva, Sherlon Almeida, et autres
Publié: (2025)
WakeupUrban: Unsupervised Semantic Segmentation of Mid-20$^{th}$ century Urban Landscapes with Satellite Imagery
par: Hao, Tianxiang, et autres
Publié: (2025)
par: Hao, Tianxiang, et autres
Publié: (2025)
Cross-Modal Urban Sensing: Evaluating Sound-Vision Alignment Across Street-Level and Aerial Imagery
par: Chen, Pengyu, et autres
Publié: (2025)
par: Chen, Pengyu, et autres
Publié: (2025)
MLG-Stereo: ViT Based Stereo Matching with Multi-Stage Local-Global Enhancement
par: Zhang, Haoyu, et autres
Publié: (2026)
par: Zhang, Haoyu, et autres
Publié: (2026)
Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching
par: Jing, Junpeng, et autres
Publié: (2024)
par: Jing, Junpeng, et autres
Publié: (2024)
StereoDiff: Stereo-Diffusion Synergy for Video Depth Estimation
par: Li, Haodong, et autres
Publié: (2025)
par: Li, Haodong, et autres
Publié: (2025)
Fisheye Stereo Vision: Depth and Range Error
par: Jiang, Leaf, et autres
Publié: (2026)
par: Jiang, Leaf, et autres
Publié: (2026)
Non-Learning Low-Light Stereo Vision
par: Wang, Jason, et autres
Publié: (2026)
par: Wang, Jason, et autres
Publié: (2026)
StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors
par: Shen, Guibao, et autres
Publié: (2025)
par: Shen, Guibao, et autres
Publié: (2025)
SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation
par: Wang, Yun, et autres
Publié: (2026)
par: Wang, Yun, et autres
Publié: (2026)
Pip-Stereo: Progressive Iterations Pruner for Iterative Optimization based Stereo Matching
par: Zheng, Jintu, et autres
Publié: (2026)
par: Zheng, Jintu, et autres
Publié: (2026)
Stereo-GS: Multi-View Stereo Vision Model for Generalizable 3D Gaussian Splatting Reconstruction
par: Huang, Xiufeng, et autres
Publié: (2025)
par: Huang, Xiufeng, et autres
Publié: (2025)
Adapting Stereo Vision From Objects To 3D Lunar Surface Reconstruction with the StereoLunar Dataset
par: Grethen, Clementine, et autres
Publié: (2025)
par: Grethen, Clementine, et autres
Publié: (2025)
DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos
par: Huang, Yuan, et autres
Publié: (2026)
par: Huang, Yuan, et autres
Publié: (2026)
MoCha-Stereo: Motif Channel Attention Network for Stereo Matching
par: Chen, Ziyang, et autres
Publié: (2024)
par: Chen, Ziyang, et autres
Publié: (2024)
Affine Correspondences in Stereo Vision: Theory, Practice, and Limitations
par: Hajder, Levente
Publié: (2026)
par: Hajder, Levente
Publié: (2026)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
par: Mittal, Mayank, et autres
Publié: (2019)
par: Mittal, Mayank, et autres
Publié: (2019)
Mono2Stereo: A Benchmark and Empirical Study for Stereo Conversion
par: Yu, Songsong, et autres
Publié: (2025)
par: Yu, Songsong, et autres
Publié: (2025)
OpenStereo: A Comprehensive Benchmark for Stereo Matching and Strong Baseline
par: Guo, Xianda, et autres
Publié: (2023)
par: Guo, Xianda, et autres
Publié: (2023)
Playing to Vision Foundation Model's Strengths in Stereo Matching
par: Liu, Chuang-Wei, et autres
Publié: (2024)
par: Liu, Chuang-Wei, et autres
Publié: (2024)
StereoWorld: Geometry-Aware Monocular-to-Stereo Video Generation
par: Xing, Ke, et autres
Publié: (2025)
par: Xing, Ke, et autres
Publié: (2025)
Multi-Object Hallucination in Vision-Language Models
par: Chen, Xuweiyi, et autres
Publié: (2024)
par: Chen, Xuweiyi, et autres
Publié: (2024)
StereoCarla: A High-Fidelity Driving Dataset for Generalizable Stereo
par: Guo, Xianda, et autres
Publié: (2025)
par: Guo, Xianda, et autres
Publié: (2025)
Documents similaires
-
WildRayZer: Self-supervised Large View Synthesis in Dynamic Environments
par: Chen, Xuweiyi, et autres
Publié: (2026) -
Probing the Mid-level Vision Capabilities of Self-Supervised Learning
par: Chen, Xuweiyi, et autres
Publié: (2024) -
Semantic-Free Procedural 3D Shapes Are Surprisingly Good Teachers
par: Chen, Xuweiyi, et autres
Publié: (2024) -
Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation
par: Chen, Xuweiyi, et autres
Publié: (2025) -
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
par: Wang, Boyang, et autres
Publié: (2025)