OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
Fuente:
arXiv
Saved in:
| Main Authors: | Jeong, Jisoo, Cai, Hong, Garrepalli, Risheek, Lin, Jamie Menjay, Hayat, Munawar, Porikli, Fatih |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SciFlow: Empowering Lightweight Optical Flow Models with Self-Cleaning Iterations
by: Lin, Jamie Menjay, et al.
Published: (2024)
by: Lin, Jamie Menjay, et al.
Published: (2024)
Improving Optical Flow and Stereo Depth Estimation by Leveraging Uncertainty-Based Learning Difficulties
by: Jeong, Jisoo, et al.
Published: (2025)
by: Jeong, Jisoo, et al.
Published: (2025)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
by: Garrepalli, Risheek, et al.
Published: (2024)
by: Garrepalli, Risheek, et al.
Published: (2024)
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2023)
by: Yasarla, Rajeev, et al.
Published: (2023)
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2024)
by: Yasarla, Rajeev, et al.
Published: (2024)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
by: Kadambi, Shreya, et al.
Published: (2025)
by: Kadambi, Shreya, et al.
Published: (2025)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Learning Optical Flow Field via Neural Ordinary Differential Equation
by: Mirvakhabova, Leyla, et al.
Published: (2025)
by: Mirvakhabova, Leyla, et al.
Published: (2025)
FouRA: Fourier Low Rank Adaptation
by: Borse, Shubhankar, et al.
Published: (2024)
by: Borse, Shubhankar, et al.
Published: (2024)
Resolving the Identity Crisis in Text-to-Image Generation
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Clockwork Diffusion: Efficient Generation With Model-Step Distillation
by: Habibian, Amirhossein, et al.
Published: (2023)
by: Habibian, Amirhossein, et al.
Published: (2023)
RoCA: Robust Cross-Domain End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2025)
by: Yasarla, Rajeev, et al.
Published: (2025)
CustomKD: Customizing Large Vision Foundation for Edge Model Improvement via Knowledge Distillation
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Distilling Multi-modal Large Language Models for Autonomous Driving
by: Hegde, Deepti, et al.
Published: (2025)
by: Hegde, Deepti, et al.
Published: (2025)
BePo: Dual Representation for 3D Occupancy Prediction
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
ODG: Occupancy Prediction Using Dual Gaussians
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
Attention Guided Alignment in Efficient Vision-Language Models
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
FLoC: Facility Location-Based Efficient Visual Token Compression for Long Video Understanding
by: Cho, Janghoon, et al.
Published: (2025)
by: Cho, Janghoon, et al.
Published: (2025)
ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints
by: Das, Debasmit, et al.
Published: (2025)
by: Das, Debasmit, et al.
Published: (2025)
Personalized OVSS: Understanding Personal Concept in Open-Vocabulary Semantic Segmentation
by: Park, Sunghyun, et al.
Published: (2025)
by: Park, Sunghyun, et al.
Published: (2025)
PosSAM: Panoptic Open-vocabulary Segment Anything
by: VS, Vibashan, et al.
Published: (2024)
by: VS, Vibashan, et al.
Published: (2024)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
Generalized Contrastive Learning for Universal Multimodal Retrieval
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Generative Scenario Rollouts for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos
by: Yasarla, Rajeev, et al.
Published: (2025)
by: Yasarla, Rajeev, et al.
Published: (2025)
H3O: Hyper-Efficient 3D Occupancy Prediction with Heterogeneous Supervision
by: Shi, Yunxiao, et al.
Published: (2025)
by: Shi, Yunxiao, et al.
Published: (2025)
DeCoTR: Enhancing Depth Completion with 2D and 3D Attentions
by: Shi, Yunxiao, et al.
Published: (2024)
by: Shi, Yunxiao, et al.
Published: (2024)
Ar2Can: An Architect and an Artist Leveraging a Canvas for Multi-Human Generation
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
by: Park, Sunghyun, et al.
Published: (2026)
by: Park, Sunghyun, et al.
Published: (2026)
ForeSea: AI Forensic Search with Multi-modal Queries for Video Surveillance
by: Park, Hyojin, et al.
Published: (2026)
by: Park, Hyojin, et al.
Published: (2026)
ToSA: Token Selective Attention for Efficient Vision Transformers
by: Singh, Manish Kumar, et al.
Published: (2024)
by: Singh, Manish Kumar, et al.
Published: (2024)
Planar Gaussian Splatting
by: Zanjani, Farhad G., et al.
Published: (2024)
by: Zanjani, Farhad G., et al.
Published: (2024)
Neural Mesh Fusion: Unsupervised 3D Planar Surface Understanding
by: Zanjani, Farhad G., et al.
Published: (2024)
by: Zanjani, Farhad G., et al.
Published: (2024)
CoReDiT: Spatial Coherence-Guided Token Pruning and Reconstruction for Efficient Diffusion Transformers
by: Li, Zhuojin, et al.
Published: (2026)
by: Li, Zhuojin, et al.
Published: (2026)
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
U$^{2}$Flow: Uncertainty-Aware Unsupervised Optical Flow Estimation
by: Sun, Xunpei, et al.
Published: (2026)
by: Sun, Xunpei, et al.
Published: (2026)
FALO: Fast and Accurate LiDAR 3D Object Detection on Resource-Constrained Devices
by: Han, Shizhong, et al.
Published: (2025)
by: Han, Shizhong, et al.
Published: (2025)
Tripartite Weight-Space Ensemble for Few-Shot Class-Incremental Learning
by: Lee, Juntae, et al.
Published: (2025)
by: Lee, Juntae, et al.
Published: (2025)
Erasing Undesirable Influence in Diffusion Models
by: Wu, Jing, et al.
Published: (2024)
by: Wu, Jing, et al.
Published: (2024)
HexaGen3D: StableDiffusion is just one step away from Fast and Diverse Text-to-3D Generation
by: Mercier, Antoine, et al.
Published: (2024)
by: Mercier, Antoine, et al.
Published: (2024)
Similar Items
-
SciFlow: Empowering Lightweight Optical Flow Models with Self-Cleaning Iterations
by: Lin, Jamie Menjay, et al.
Published: (2024) -
Improving Optical Flow and Stereo Depth Estimation by Leveraging Uncertainty-Based Learning Difficulties
by: Jeong, Jisoo, et al.
Published: (2025) -
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
by: Garrepalli, Risheek, et al.
Published: (2024) -
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2023) -
FutureDepth: Learning to Predict the Future Improves Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2024)