Staff View: :: Library Catalog

Saved in:

Bibliographic Details
Main Authors:	Zhou, Chaochao, Faruqui, Syed Hasib Akhter, Patel, Abhinav, Abdalla, Ramez N., Hurley, Michael C., Shaibani, Ali, Potts, Matthew B., Jahromi, Babak S., Ansari, Sameer A., Cantrell, Donald R.
Format:	Preprint
Published:	2023
Subjects:	Computer Vision and Pattern Recognition Machine Learning Image and Video Processing
Online Access:	https://arxiv.org/abs/2308.00214
Tags:	Add Tag No Tags, Be the first to tag this record!

_version_	1866916561690820608
author	Zhou, Chaochao Faruqui, Syed Hasib Akhter Patel, Abhinav Abdalla, Ramez N. Hurley, Michael C. Shaibani, Ali Potts, Matthew B. Jahromi, Babak S. Ansari, Sameer A. Cantrell, Donald R.
author_facet	Zhou, Chaochao Faruqui, Syed Hasib Akhter Patel, Abhinav Abdalla, Ramez N. Hurley, Michael C. Shaibani, Ali Potts, Matthew B. Jahromi, Babak S. Ansari, Sameer A. Cantrell, Donald R.
contents	Many tasks performed in image-guided procedures can be cast as pose estimation problems, where specific projections are chosen to reach a target in 3D space. In this study, we first develop a differentiable projection (DiffProj) rendering framework for the efficient computation of Digitally Reconstructed Radiographs (DRRs) with automatic differentiability from either Cone-Beam Computerized Tomography (CBCT) or neural scene representations, including two newly proposed methods, Neural Tuned Tomography (NeTT) and masked Neural Radiance Fields (mNeRF). We then perform pose estimation by iterative gradient descent using various candidate loss functions, that quantify the image discrepancy of the synthesized DRR with respect to the ground-truth fluoroscopic X-ray image. Compared to alternative loss functions, the Mutual Information loss function can significantly improve pose estimation accuracy, as it can effectively prevent entrapment in local optima. Using the Mutual Information loss, a comprehensive evaluation of pose estimation performed on a tomographic X-ray dataset of 50 patients$'$ skulls shows that utilizing either discretized (CBCT) or neural (NeTT/mNeRF) scene representations in DiffProj leads to comparable performance in DRR appearance and pose estimation (3D angle errors: mean $\leq$ 3.2° and 90% quantile $\leq$ 3.4°), despite the latter often incurring considerable training expenses and time. These findings could be instrumental for selecting appropriate approaches to improve the efficiency and effectiveness of fluoroscopic X-ray pose estimation in widespread image-guided interventions.
format	Preprint
id	arxiv_https___arxiv_org_abs_2308_00214
institution	arXiv
publishDate	2023
record_format	arxiv
spellingShingle	The Impact of Loss Functions and Scene Representations for 3D/2D Registration on Single-view Fluoroscopic X-ray Pose Estimation Zhou, Chaochao Faruqui, Syed Hasib Akhter Patel, Abhinav Abdalla, Ramez N. Hurley, Michael C. Shaibani, Ali Potts, Matthew B. Jahromi, Babak S. Ansari, Sameer A. Cantrell, Donald R. Computer Vision and Pattern Recognition Machine Learning Image and Video Processing Many tasks performed in image-guided procedures can be cast as pose estimation problems, where specific projections are chosen to reach a target in 3D space. In this study, we first develop a differentiable projection (DiffProj) rendering framework for the efficient computation of Digitally Reconstructed Radiographs (DRRs) with automatic differentiability from either Cone-Beam Computerized Tomography (CBCT) or neural scene representations, including two newly proposed methods, Neural Tuned Tomography (NeTT) and masked Neural Radiance Fields (mNeRF). We then perform pose estimation by iterative gradient descent using various candidate loss functions, that quantify the image discrepancy of the synthesized DRR with respect to the ground-truth fluoroscopic X-ray image. Compared to alternative loss functions, the Mutual Information loss function can significantly improve pose estimation accuracy, as it can effectively prevent entrapment in local optima. Using the Mutual Information loss, a comprehensive evaluation of pose estimation performed on a tomographic X-ray dataset of 50 patients$'$ skulls shows that utilizing either discretized (CBCT) or neural (NeTT/mNeRF) scene representations in DiffProj leads to comparable performance in DRR appearance and pose estimation (3D angle errors: mean $\leq$ 3.2° and 90% quantile $\leq$ 3.4°), despite the latter often incurring considerable training expenses and time. These findings could be instrumental for selecting appropriate approaches to improve the efficiency and effectiveness of fluoroscopic X-ray pose estimation in widespread image-guided interventions.
title	The Impact of Loss Functions and Scene Representations for 3D/2D Registration on Single-view Fluoroscopic X-ray Pose Estimation
topic	Computer Vision and Pattern Recognition Machine Learning Image and Video Processing
url	https://arxiv.org/abs/2308.00214

Similar Items