MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Yasarla, Rajeev, Cai, Hong, Jeong, Jisoo, Shi, Yunxiao, Garrepalli, Risheek, Porikli, Fatih |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2024)
by: Yasarla, Rajeev, et al.
Published: (2024)
OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
by: Jeong, Jisoo, et al.
Published: (2024)
by: Jeong, Jisoo, et al.
Published: (2024)
SciFlow: Empowering Lightweight Optical Flow Models with Self-Cleaning Iterations
by: Lin, Jamie Menjay, et al.
Published: (2024)
by: Lin, Jamie Menjay, et al.
Published: (2024)
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2025)
by: Yasarla, Rajeev, et al.
Published: (2025)
Improving Optical Flow and Stereo Depth Estimation by Leveraging Uncertainty-Based Learning Difficulties
by: Jeong, Jisoo, et al.
Published: (2025)
by: Jeong, Jisoo, et al.
Published: (2025)
DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos
by: Yasarla, Rajeev, et al.
Published: (2025)
by: Yasarla, Rajeev, et al.
Published: (2025)
ToSA: Token Selective Attention for Efficient Vision Transformers
by: Singh, Manish Kumar, et al.
Published: (2024)
by: Singh, Manish Kumar, et al.
Published: (2024)
DeCoTR: Enhancing Depth Completion with 2D and 3D Attentions
by: Shi, Yunxiao, et al.
Published: (2024)
by: Shi, Yunxiao, et al.
Published: (2024)
Generative Scenario Rollouts for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
Distilling Multi-modal Large Language Models for Autonomous Driving
by: Hegde, Deepti, et al.
Published: (2025)
by: Hegde, Deepti, et al.
Published: (2025)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
by: Garrepalli, Risheek, et al.
Published: (2024)
by: Garrepalli, Risheek, et al.
Published: (2024)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
by: Kadambi, Shreya, et al.
Published: (2025)
by: Kadambi, Shreya, et al.
Published: (2025)
BePo: Dual Representation for 3D Occupancy Prediction
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
ODG: Occupancy Prediction Using Dual Gaussians
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
H3O: Hyper-Efficient 3D Occupancy Prediction with Heterogeneous Supervision
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
HexaGen3D: StableDiffusion is just one step away from Fast and Diverse Text-to-3D Generation
by: Mercier, Antoine, et al.
Published: (2024)
by: Mercier, Antoine, et al.
Published: (2024)
Clockwork Diffusion: Efficient Generation With Model-Step Distillation
by: Habibian, Amirhossein, et al.
Published: (2023)
by: Habibian, Amirhossein, et al.
Published: (2023)
Learning Optical Flow Field via Neural Ordinary Differential Equation
by: Mirvakhabova, Leyla, et al.
Published: (2025)
by: Mirvakhabova, Leyla, et al.
Published: (2025)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
FouRA: Fourier Low Rank Adaptation
by: Borse, Shubhankar, et al.
Published: (2024)
by: Borse, Shubhankar, et al.
Published: (2024)
PADRe: A Unifying Polynomial Attention Drop-in Replacement for Efficient Vision Transformer
by: Letourneau, Pierre-David, et al.
Published: (2024)
by: Letourneau, Pierre-David, et al.
Published: (2024)
Leveraging Near-Field Lighting for Monocular Depth Estimation from Endoscopy Videos
by: Paruchuri, Akshay, et al.
Published: (2024)
by: Paruchuri, Akshay, et al.
Published: (2024)
StarryGazer: Leveraging Monocular Depth Estimation Models for Domain-Agnostic Single Depth Image Completion
by: Hong, Sangmin, et al.
Published: (2025)
by: Hong, Sangmin, et al.
Published: (2025)
TanDepth: Leveraging Global DEMs for Metric Monocular Depth Estimation in UAVs
by: Florea, Horatiu, et al.
Published: (2024)
by: Florea, Horatiu, et al.
Published: (2024)
Self-supervised Monocular Depth Estimation with Large Kernel Attention
by: Xiang, Xuezhi, et al.
Published: (2024)
by: Xiang, Xuezhi, et al.
Published: (2024)
StableDPT: Temporal Stable Monocular Video Depth Estimation
by: Sobko, Ivan, et al.
Published: (2026)
by: Sobko, Ivan, et al.
Published: (2026)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Planar Gaussian Splatting
by: Zanjani, Farhad G., et al.
Published: (2024)
by: Zanjani, Farhad G., et al.
Published: (2024)
Neural Mesh Fusion: Unsupervised 3D Planar Surface Understanding
by: Zanjani, Farhad G., et al.
Published: (2024)
by: Zanjani, Farhad G., et al.
Published: (2024)
CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers
by: Li, Zhuojin, et al.
Published: (2026)
by: Li, Zhuojin, et al.
Published: (2026)
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos
by: Lu, Jiahao, et al.
Published: (2024)
by: Lu, Jiahao, et al.
Published: (2024)
Adaptive Depth-converted-Scale Convolution for Self-supervised Monocular Depth Estimation
by: Gao, Yanbo, et al.
Published: (2026)
by: Gao, Yanbo, et al.
Published: (2026)
UniDepth: Universal Monocular Metric Depth Estimation
by: Piccinelli, Luigi, et al.
Published: (2024)
by: Piccinelli, Luigi, et al.
Published: (2024)
Survey on Monocular Metric Depth Estimation
by: Zhang, Jiuling
Published: (2025)
by: Zhang, Jiuling
Published: (2025)
The Third Monocular Depth Estimation Challenge
by: Spencer, Jaime, et al.
Published: (2024)
by: Spencer, Jaime, et al.
Published: (2024)
The Fourth Monocular Depth Estimation Challenge
by: Obukhov, Anton, et al.
Published: (2025)
by: Obukhov, Anton, et al.
Published: (2025)
Scalable Autoregressive Monocular Depth Estimation
by: Wang, Jinhong, et al.
Published: (2024)
by: Wang, Jinhong, et al.
Published: (2024)
STATIC : Surface Temporal Affine for TIme Consistency in Video Monocular Depth Estimation
by: Yang, Sunghun, et al.
Published: (2024)
by: Yang, Sunghun, et al.
Published: (2024)
Similar Items
-
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2024) -
OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
by: Jeong, Jisoo, et al.
Published: (2024) -
SciFlow: Empowering Lightweight Optical Flow Models with Self-Cleaning Iterations
by: Lin, Jamie Menjay, et al.
Published: (2024) -
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2025) -
Improving Optical Flow and Stereo Depth Estimation by Leveraging Uncertainty-Based Learning Difficulties
by: Jeong, Jisoo, et al.
Published: (2025)