ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Ren, Huan, Chen, Yihan, Wang, Chuxin, Liu, Nailong, Yang, Wenfei, Zhang, Tianzhu
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866913161325576192
author Ren, Huan
Chen, Yihan
Wang, Chuxin
Liu, Nailong
Yang, Wenfei
Zhang, Tianzhu
author_facet Ren, Huan
Chen, Yihan
Wang, Chuxin
Liu, Nailong
Yang, Wenfei
Zhang, Tianzhu
contents Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with the inherent incompleteness of observed point clouds, which limits their ability to capture complete object shapes for robust pose reasoning. While point cloud completion offers a promising solution, naively treating it as a separate preprocessing step for partial observations introduces compounding errors and additional computational overhead, ultimately hindering both accuracy and efficiency. To address these challenges, we propose ComPose, a novel unified framework that tightly integrates shape completion to provide complete geometric cues for enhanced pose estimation. At the core of ComPose is a keypoint-based progressive completion module, which recovers full shape representations by progressively predicting a sparse set of keypoints and their surrounding dense point sets, empowering the keypoints to capture holistic object geometries. A geometric relation encoding module further enriches keypoint features with both local and global geometric context. In addition, we introduce a novel geometric relation consistency loss to enforce structural alignment between observed keypoints and their predicted NOCS coordinates, ensuring globally coherent coordinate transformations. Extensive experiments on standard benchmarks demonstrate that our method outperforms state-of-the-art approaches without relying on category-level shape priors.
format Preprint
id arxiv_https___arxiv_org_abs_2605_25553
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
Ren, Huan
Chen, Yihan
Wang, Chuxin
Liu, Nailong
Yang, Wenfei
Zhang, Tianzhu
Computer Vision and Pattern Recognition
Robotics
Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with the inherent incompleteness of observed point clouds, which limits their ability to capture complete object shapes for robust pose reasoning. While point cloud completion offers a promising solution, naively treating it as a separate preprocessing step for partial observations introduces compounding errors and additional computational overhead, ultimately hindering both accuracy and efficiency. To address these challenges, we propose ComPose, a novel unified framework that tightly integrates shape completion to provide complete geometric cues for enhanced pose estimation. At the core of ComPose is a keypoint-based progressive completion module, which recovers full shape representations by progressively predicting a sparse set of keypoints and their surrounding dense point sets, empowering the keypoints to capture holistic object geometries. A geometric relation encoding module further enriches keypoint features with both local and global geometric context. In addition, we introduce a novel geometric relation consistency loss to enforce structural alignment between observed keypoints and their predicted NOCS coordinates, ensuring globally coherent coordinate transformations. Extensive experiments on standard benchmarks demonstrate that our method outperforms state-of-the-art approaches without relying on category-level shape priors.
title ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
topic Computer Vision and Pattern Recognition
Robotics
url https://arxiv.org/abs/2605.25553