Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Shi, Ge, Yang, Zhili
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909105930633216
author Shi, Ge
Yang, Zhili
author_facet Shi, Ge
Yang, Zhili
contents Dynamic scene understanding is one of the most conspicuous field of interest among computer vision community. In order to enhance dynamic scene understanding, pixel-wise segmentation with neural networks is widely accepted. The latest researches on pixel-wise segmentation combined semantic and motion information and produced good performance. In this work, we propose a state of art architecture of neural networks to accurately and efficiently get the moving object proposals (MOP). We first train an unsupervised convolutional neural network (UnFlow) to generate optical flow estimation. Then we render the output of optical flow net to a fully convolutional SegNet model. The main contribution of our work is (1) Fine-tuning the pretrained optical flow model on the brand new DAVIS Dataset; (2) Leveraging fully convolutional neural networks with Encoder-Decoder architecture to segment objects. We developed the codes with TensorFlow, and executed the training and evaluation processes on an AWS EC2 instance.
format Preprint
id arxiv_https___arxiv_org_abs_2402_08882
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation
Shi, Ge
Yang, Zhili
Computer Vision and Pattern Recognition
Machine Learning
68Txx
Dynamic scene understanding is one of the most conspicuous field of interest among computer vision community. In order to enhance dynamic scene understanding, pixel-wise segmentation with neural networks is widely accepted. The latest researches on pixel-wise segmentation combined semantic and motion information and produced good performance. In this work, we propose a state of art architecture of neural networks to accurately and efficiently get the moving object proposals (MOP). We first train an unsupervised convolutional neural network (UnFlow) to generate optical flow estimation. Then we render the output of optical flow net to a fully convolutional SegNet model. The main contribution of our work is (1) Fine-tuning the pretrained optical flow model on the brand new DAVIS Dataset; (2) Leveraging fully convolutional neural networks with Encoder-Decoder architecture to segment objects. We developed the codes with TensorFlow, and executed the training and evaluation processes on an AWS EC2 instance.
title Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation
topic Computer Vision and Pattern Recognition
Machine Learning
68Txx
url https://arxiv.org/abs/2402.08882