CAP-Net: A Unified Network for 6D Pose and Size Estimation of Categorical Articulated Parts from a Single RGB-D Image
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Jingshun, Lin, Haitao, Wang, Tianyu, Fu, Yanwei, Xue, Xiangyang, Zhu, Yi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
You Only Estimate Once: Unified, One-stage, Real-Time Category-level Articulated Object 6D Pose Estimation for Robotic Grasping
by: Huang, Jingshun, et al.
Published: (2025)
by: Huang, Jingshun, et al.
Published: (2025)
Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
by: Lin, Haitao, et al.
Published: (2026)
by: Lin, Haitao, et al.
Published: (2026)
SparseGrasp: Robotic Grasping via 3D Semantic Gaussian Splatting from Sparse Multi-View RGB Images
by: Yu, Junqiu, et al.
Published: (2024)
by: Yu, Junqiu, et al.
Published: (2024)
Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View
by: Zhang, Jinyu, et al.
Published: (2025)
by: Zhang, Jinyu, et al.
Published: (2025)
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting
by: Yang, Linqi, et al.
Published: (2025)
by: Yang, Linqi, et al.
Published: (2025)
Active 6D Pose Estimation for Textureless Objects using Multi-View RGB Frames
by: Yang, Jun, et al.
Published: (2025)
by: Yang, Jun, et al.
Published: (2025)
Single-Shot 6DoF Pose and 3D Size Estimation for Robotic Strawberry Harvesting
by: Li, Lun, et al.
Published: (2024)
by: Li, Lun, et al.
Published: (2024)
LAC-Net: Linear-Fusion Attention-Guided Convolutional Network for Accurate Robotic Grasping Under the Occlusion
by: Zhang, Jinyu, et al.
Published: (2024)
by: Zhang, Jinyu, et al.
Published: (2024)
6D Pose Estimation via Keypoint Heatmap Regression with RGB-D Residual Neural Networks
by: Aljosevic, Ismail, et al.
Published: (2026)
by: Aljosevic, Ismail, et al.
Published: (2026)
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
by: Liu, Zhenyang, et al.
Published: (2026)
by: Liu, Zhenyang, et al.
Published: (2026)
Sparse Color-Code Net: Real-Time RGB-Based 6D Object Pose Estimation on Edge Devices
by: Yang, Xingjian, et al.
Published: (2024)
by: Yang, Xingjian, et al.
Published: (2024)
FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects
by: Wen, Bowen, et al.
Published: (2023)
by: Wen, Bowen, et al.
Published: (2023)
Fusing Monocular RGB Images with AIS Data to Create a 6D Pose Estimation Dataset for Marine Vessels
by: Holst, Fabian, et al.
Published: (2025)
by: Holst, Fabian, et al.
Published: (2025)
TriVLA: A Triple-System-Based Unified Vision-Language-Action Model with Episodic World Modeling for General Robot Control
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
RAG-6DPose: Retrieval-Augmented 6D Pose Estimation via Leveraging CAD as Knowledge Base
by: Wang, Kuanning, et al.
Published: (2025)
by: Wang, Kuanning, et al.
Published: (2025)
OmniPose6D: Towards Short-Term Object Pose Tracking in Dynamic Scenes from Monocular RGB
by: Lin, Yunzhi, et al.
Published: (2024)
by: Lin, Yunzhi, et al.
Published: (2024)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
ContactArt: Learning 3D Interaction Priors for Category-level Articulated Object and Hand Poses Estimation
by: Zhu, Zehao, et al.
Published: (2023)
by: Zhu, Zehao, et al.
Published: (2023)
ActivePose: Active 6D Object Pose Estimation and Tracking for Robotic Manipulation
by: Liu, Sheng, et al.
Published: (2025)
by: Liu, Sheng, et al.
Published: (2025)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
ToolEENet: Tool Affordance 6D Pose Estimation
by: Wang, Yunlong, et al.
Published: (2024)
by: Wang, Yunlong, et al.
Published: (2024)
Towards Real-World Aerial Vision Guidance with Categorical 6D Pose Tracker
by: Sun, Jingtao, et al.
Published: (2024)
by: Sun, Jingtao, et al.
Published: (2024)
DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects
by: Chen, Jiahong, et al.
Published: (2025)
by: Chen, Jiahong, et al.
Published: (2025)
SCOOP'D: Learning Mixed-Liquid-Solid Scooping via Sim2Real Generative Policy
by: Wang, Kuanning, et al.
Published: (2025)
by: Wang, Kuanning, et al.
Published: (2025)
A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
by: Wu, Zijian, et al.
Published: (2025)
by: Wu, Zijian, et al.
Published: (2025)
Part-Guided 3D RL for Sim2Real Articulated Object Manipulation
by: Xie, Pengwei, et al.
Published: (2024)
by: Xie, Pengwei, et al.
Published: (2024)
MR6D: Benchmarking 6D Pose Estimation for Mobile Robots
by: Gouda, Anas, et al.
Published: (2025)
by: Gouda, Anas, et al.
Published: (2025)
Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Markerless Robot Detection and 6D Pose Estimation for Multi-Agent SLAM
by: Rueggeberg, Markus, et al.
Published: (2026)
by: Rueggeberg, Markus, et al.
Published: (2026)
OMNI-PoseX: A Fast Vision Model for 6D Object Pose Estimation in Embodied Tasks
by: Zhang, Michael, et al.
Published: (2026)
by: Zhang, Michael, et al.
Published: (2026)
UA-Pose: Uncertainty-Aware 6D Object Pose Estimation and Online Object Completion with Partial References
by: Li, Ming-Feng, et al.
Published: (2025)
by: Li, Ming-Feng, et al.
Published: (2025)
SplArt: Articulation Estimation and Part-Level Reconstruction with 3D Gaussian Splatting
by: Lin, Shengjie, et al.
Published: (2025)
by: Lin, Shengjie, et al.
Published: (2025)
HIPPo: Harnessing Image-to-3D Priors for Model-free Zero-shot 6D Pose Estimation
by: Liu, Yibo, et al.
Published: (2025)
by: Liu, Yibo, et al.
Published: (2025)
PS6D: Point Cloud Based Symmetry-Aware 6D Object Pose Estimation in Robot Bin-Picking
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
Realistic Data Generation for 6D Pose Estimation of Surgical Instruments
by: Barragan, Juan Antonio, et al.
Published: (2024)
by: Barragan, Juan Antonio, et al.
Published: (2024)
Triplane Grasping: Efficient 6-DoF Grasping with Single RGB Images
by: Li, Yiming, et al.
Published: (2024)
by: Li, Yiming, et al.
Published: (2024)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
by: Serpiva, Valerii, et al.
Published: (2024)
by: Serpiva, Valerii, et al.
Published: (2024)
OCRA: Object-Centric Learning with 3D and Tactile Priors for Human-to-Robot Action Transfer
by: Wang, Kuanning, et al.
Published: (2026)
by: Wang, Kuanning, et al.
Published: (2026)
You Only Pose Once: A Minimalist's Detection Transformer for Monocular RGB Category-level 9D Multi-Object Pose Estimation
by: Lee, Hakjin, et al.
Published: (2025)
by: Lee, Hakjin, et al.
Published: (2025)
Similar Items
-
You Only Estimate Once: Unified, One-stage, Real-Time Category-level Articulated Object 6D Pose Estimation for Robotic Grasping
by: Huang, Jingshun, et al.
Published: (2025) -
Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
by: Lin, Haitao, et al.
Published: (2026) -
SparseGrasp: Robotic Grasping via 3D Semantic Gaussian Splatting from Sparse Multi-View RGB Images
by: Yu, Junqiu, et al.
Published: (2024) -
Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View
by: Zhang, Jinyu, et al.
Published: (2025) -
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting
by: Yang, Linqi, et al.
Published: (2025)