FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chou, Gene, Xian, Wenqi, Yang, Guandao, Abdelfattah, Mohamed, Hariharan, Bharath, Snavely, Noah, Yu, Ning, Debevec, Paul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MegaScenes: Scene-Level View Synthesis at Scale
von: Tung, Joseph, et al.
Veröffentlicht: (2024)
von: Tung, Joseph, et al.
Veröffentlicht: (2024)
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
von: Chou, Gene, et al.
Veröffentlicht: (2024)
von: Chou, Gene, et al.
Veröffentlicht: (2024)
Learning Feature Descriptors using Camera Pose Supervision
von: Wang, Qianqian, et al.
Veröffentlicht: (2020)
von: Wang, Qianqian, et al.
Veröffentlicht: (2020)
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
von: Chou, Gene, et al.
Veröffentlicht: (2026)
von: Chou, Gene, et al.
Veröffentlicht: (2026)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
von: Huang, Kuan Wei, et al.
Veröffentlicht: (2025)
von: Huang, Kuan Wei, et al.
Veröffentlicht: (2025)
Self-Calibrating Gaussian Splatting for Large Field of View Reconstruction
von: Deng, Youming, et al.
Veröffentlicht: (2025)
von: Deng, Youming, et al.
Veröffentlicht: (2025)
G3T Up! Gravity Aligned Coordinate Frames Simplify Pointmap Processing
von: Kani, Bharath Raj Nagoor, et al.
Veröffentlicht: (2026)
von: Kani, Bharath Raj Nagoor, et al.
Veröffentlicht: (2026)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Accurate Differential Operators for Hybrid Neural Fields
von: Chetan, Aditya, et al.
Veröffentlicht: (2023)
von: Chetan, Aditya, et al.
Veröffentlicht: (2023)
Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation
von: Lin, Haotong, et al.
Veröffentlicht: (2024)
von: Lin, Haotong, et al.
Veröffentlicht: (2024)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
von: Hassena, Gemmechu, et al.
Veröffentlicht: (2024)
von: Hassena, Gemmechu, et al.
Veröffentlicht: (2024)
MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
von: Sun, Yihong, et al.
Veröffentlicht: (2024)
von: Sun, Yihong, et al.
Veröffentlicht: (2024)
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
FlashVSR: Towards Real-Time Diffusion-Based Streaming Video Super-Resolution
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
CineScale: Free Lunch in High-Resolution Cinematic Visual Generation
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
Depth-Centric Dehazing and Depth-Estimation from Real-World Hazy Driving Video
von: Fan, Junkai, et al.
Veröffentlicht: (2024)
von: Fan, Junkai, et al.
Veröffentlicht: (2024)
Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise
von: Burgert, Ryan, et al.
Veröffentlicht: (2025)
von: Burgert, Ryan, et al.
Veröffentlicht: (2025)
Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
von: Chen, Sili, et al.
Veröffentlicht: (2025)
von: Chen, Sili, et al.
Veröffentlicht: (2025)
Wide-Baseline Relative Camera Pose Estimation with Directional Learning
von: Chen, Kefan, et al.
Veröffentlicht: (2021)
von: Chen, Kefan, et al.
Veröffentlicht: (2021)
Real-time Monocular Depth Estimation on Embedded Systems
von: Feng, Cheng, et al.
Veröffentlicht: (2023)
von: Feng, Cheng, et al.
Veröffentlicht: (2023)
MAViS: A Multi-Agent Framework for Long-Sequence Video Storytelling
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural Implicit Fields
von: Yu, Hao, et al.
Veröffentlicht: (2026)
von: Yu, Hao, et al.
Veröffentlicht: (2026)
Neural Control Variates with Automatic Integration
von: Li, Zilu, et al.
Veröffentlicht: (2024)
von: Li, Zilu, et al.
Veröffentlicht: (2024)
Depth Anywhere: Enhancing 360 Monocular Depth Estimation via Perspective Distillation and Unlabeled Data Augmentation
von: Wang, Ning-Hsu, et al.
Veröffentlicht: (2024)
von: Wang, Ning-Hsu, et al.
Veröffentlicht: (2024)
DepthPolyp: Pseudo-Depth Guided Lightweight Segmentation for Real-Time Colonoscopy
von: Wu, Zhuoyu, et al.
Veröffentlicht: (2026)
von: Wu, Zhuoyu, et al.
Veröffentlicht: (2026)
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2024)
von: Yasarla, Rajeev, et al.
Veröffentlicht: (2024)
Flash-VStream: Efficient Real-Time Understanding for Long Video Streams
von: Zhang, Haoji, et al.
Veröffentlicht: (2025)
von: Zhang, Haoji, et al.
Veröffentlicht: (2025)
DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
von: Dong, Yue-Jiang, et al.
Veröffentlicht: (2025)
von: Dong, Yue-Jiang, et al.
Veröffentlicht: (2025)
Instance-Guided Radar Depth Estimation for 3D Object Detection
von: Lo, Chen-Chou, et al.
Veröffentlicht: (2026)
von: Lo, Chen-Chou, et al.
Veröffentlicht: (2026)
Loss-resilient Coding of Texture and Depth for Free-viewpoint Video Conferencing
von: Macchiavello, Bruno, et al.
Veröffentlicht: (2013)
von: Macchiavello, Bruno, et al.
Veröffentlicht: (2013)
QuartDepth: Post-Training Quantization for Real-Time Depth Estimation on the Edge
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
NVDS+: Towards Efficient and Versatile Neural Stabilizer for Video Depth Estimation
von: Wang, Yiran, et al.
Veröffentlicht: (2023)
von: Wang, Yiran, et al.
Veröffentlicht: (2023)
Flash-VStream: Memory-Based Real-Time Understanding for Long Video Streams
von: Zhang, Haoji, et al.
Veröffentlicht: (2024)
von: Zhang, Haoji, et al.
Veröffentlicht: (2024)
Can Generative Video Models Help Pose Estimation?
von: Cai, Ruojin, et al.
Veröffentlicht: (2024)
von: Cai, Ruojin, et al.
Veröffentlicht: (2024)
Depth-Aware Image and Video Orientation Estimation
von: Alam, Muhammad Z., et al.
Veröffentlicht: (2026)
von: Alam, Muhammad Z., et al.
Veröffentlicht: (2026)
Live Interactive Training for Video Segmentation
von: Yang, Xinyu, et al.
Veröffentlicht: (2026)
von: Yang, Xinyu, et al.
Veröffentlicht: (2026)
UniDepth: Universal Monocular Metric Depth Estimation
von: Piccinelli, Luigi, et al.
Veröffentlicht: (2024)
von: Piccinelli, Luigi, et al.
Veröffentlicht: (2024)
Enhancing Neural Radiance Fields with Depth and Normal Completion Priors from Sparse Views
von: Guo, Jiawei, et al.
Veröffentlicht: (2024)
von: Guo, Jiawei, et al.
Veröffentlicht: (2024)
SpatioTemporal Difference Network for Video Depth Super-Resolution
von: Wang, Zhengxue, et al.
Veröffentlicht: (2025)
von: Wang, Zhengxue, et al.
Veröffentlicht: (2025)
ScaleDepth: Decomposing Metric Depth Estimation into Scale Prediction and Relative Depth Estimation
von: Zhu, Ruijie, et al.
Veröffentlicht: (2024)
von: Zhu, Ruijie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MegaScenes: Scene-Level View Synthesis at Scale
von: Tung, Joseph, et al.
Veröffentlicht: (2024) -
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
von: Chou, Gene, et al.
Veröffentlicht: (2024) -
Learning Feature Descriptors using Camera Pose Supervision
von: Wang, Qianqian, et al.
Veröffentlicht: (2020) -
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
von: Chou, Gene, et al.
Veröffentlicht: (2026) -
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
von: Huang, Kuan Wei, et al.
Veröffentlicht: (2025)