Guardado en:
| Autores principales: | , , , |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2412.03490 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
| _version_ | 1866909415994556416 |
|---|---|
| author | Yusuf, Md Abu Khan, Md Rezaul Karim Saha, Partha Pratim Rahaman, Mohammed Mahbubur |
| author_facet | Yusuf, Md Abu Khan, Md Rezaul Karim Saha, Partha Pratim Rahaman, Mohammed Mahbubur |
| contents | Considerable study has already been conducted regarding autonomous driving in modern era. An autonomous driving system must be extremely good at detecting objects surrounding the car to ensure safety. In this paper, classification, and estimation of an object's (pedestrian) position (concerning an ego 3D coordinate system) are studied and the distance between the ego vehicle and the object in the context of autonomous driving is measured. To classify the object, faster Region-based Convolution Neural Network (R-CNN) with inception v2 is utilized. First, a network is trained with customized dataset to estimate the reference position of objects as well as the distance from the vehicle. From camera calibration to computing the distance, cutting-edge technologies of computer vision algorithms in a series of processes are applied to generate a 3D reference point of the region of interest. The foremost step in this process is generating a disparity map using the concept of stereo vision. |
| format | Preprint |
| id |
arxiv_https___arxiv_org_abs_2412_03490 |
| institution | arXiv |
| publishDate | 2024 |
| record_format | arxiv |
| spellingShingle | Data Fusion of Semantic and Depth Information in the Context of Object Detection Yusuf, Md Abu Khan, Md Rezaul Karim Saha, Partha Pratim Rahaman, Mohammed Mahbubur Computer Vision and Pattern Recognition Considerable study has already been conducted regarding autonomous driving in modern era. An autonomous driving system must be extremely good at detecting objects surrounding the car to ensure safety. In this paper, classification, and estimation of an object's (pedestrian) position (concerning an ego 3D coordinate system) are studied and the distance between the ego vehicle and the object in the context of autonomous driving is measured. To classify the object, faster Region-based Convolution Neural Network (R-CNN) with inception v2 is utilized. First, a network is trained with customized dataset to estimate the reference position of objects as well as the distance from the vehicle. From camera calibration to computing the distance, cutting-edge technologies of computer vision algorithms in a series of processes are applied to generate a 3D reference point of the region of interest. The foremost step in this process is generating a disparity map using the concept of stereo vision. |
| title | Data Fusion of Semantic and Depth Information in the Context of Object Detection |
| topic | Computer Vision and Pattern Recognition |
| url | https://arxiv.org/abs/2412.03490 |