Deep Learning-Based Object Pose Estimation: A Comprehensive Survey
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jian, Sun, Wei, Yang, Hui, Zeng, Zhiwen, Liu, Chongpei, Zheng, Jin, Liu, Xingyu, Rahmani, Hossein, Sebe, Nicu, Mian, Ajmal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
by: Yang, Hui, et al.
Published: (2026)
by: Yang, Hui, et al.
Published: (2026)
Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
by: Yang, Hui, et al.
Published: (2025)
by: Yang, Hui, et al.
Published: (2025)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2023)
by: Li, Wenhao, et al.
Published: (2023)
Deep Learning Based 3D Segmentation: A Survey
by: He, Yong, et al.
Published: (2021)
by: He, Yong, et al.
Published: (2021)
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation
by: Sun, Jingtao, et al.
Published: (2024)
by: Sun, Jingtao, et al.
Published: (2024)
GraphMLP: A Graph MLP-Like Architecture for 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2022)
by: Li, Wenhao, et al.
Published: (2022)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
by: Li, Congliang, et al.
Published: (2022)
by: Li, Congliang, et al.
Published: (2022)
UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Uncertainty-Aware Testing-Time Optimization for 3D Human Pose Estimation
by: Wang, Ti, et al.
Published: (2024)
by: Wang, Ti, et al.
Published: (2024)
Unifying Causal Reinforcement Learning: Survey, Taxonomy, Algorithms and Applications
by: Cunha, Cristiano da Costa, et al.
Published: (2025)
by: Cunha, Cristiano da Costa, et al.
Published: (2025)
H$_{2}$OT: Hierarchical Hourglass Tokenizer for Efficient Video Pose Transformers
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Avatar Concept Slider: Controllable Editing of Concepts in 3D Human Avatars
by: Foo, Lin Geng, et al.
Published: (2024)
by: Foo, Lin Geng, et al.
Published: (2024)
Implicit Neural Representation-Based Continuous Single Image Super-Resolution: An Empirical Benchmark
by: Nasir, Tayyab, et al.
Published: (2026)
by: Nasir, Tayyab, et al.
Published: (2026)
Visual Attention Methods in Deep Learning: An In-Depth Survey
by: Hassanin, Mohammed, et al.
Published: (2022)
by: Hassanin, Mohammed, et al.
Published: (2022)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
Vision+X: A Survey on Multimodal Learning in the Light of Data
by: Zhu, Ye, et al.
Published: (2022)
by: Zhu, Ye, et al.
Published: (2022)
An Image-like Diffusion Method for Human-Object Interaction Detection
by: Hui, Xiaofei, et al.
Published: (2025)
by: Hui, Xiaofei, et al.
Published: (2025)
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Asymmetric GANs for Image-to-Image Translation
by: Tang, Hao, et al.
Published: (2019)
by: Tang, Hao, et al.
Published: (2019)
Object Gaussian for Monocular 6D Pose Estimation from Sparse Views
by: Luo, Luqing, et al.
Published: (2024)
by: Luo, Luqing, et al.
Published: (2024)
NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
by: Nasir, Tayyab, et al.
Published: (2026)
by: Nasir, Tayyab, et al.
Published: (2026)
GDRNPP: A Geometry-guided and Fully Learning-based Object Pose Estimator
by: Liu, Xingyu, et al.
Published: (2021)
by: Liu, Xingyu, et al.
Published: (2021)
RankFeat&RankWeight: Rank-1 Feature/Weight Removal for Out-of-distribution Detection
by: Song, Yue, et al.
Published: (2023)
by: Song, Yue, et al.
Published: (2023)
PicoPose: Progressive Pixel-to-Pixel Correspondence Learning for Novel Object Pose Estimation
by: Liu, Lihua, et al.
Published: (2025)
by: Liu, Lihua, et al.
Published: (2025)
Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation
by: Geng, Zichen, et al.
Published: (2026)
by: Geng, Zichen, et al.
Published: (2026)
Risk-Aware World Model Predictive Control for Generalizable End-to-End Autonomous Driving
by: Sun, Jiangxin, et al.
Published: (2026)
by: Sun, Jiangxin, et al.
Published: (2026)
Multi-focal Conditioned Latent Diffusion for Person Image Synthesis
by: Liu, Jiaqi, et al.
Published: (2025)
by: Liu, Jiaqi, et al.
Published: (2025)
A Lie Group Approach to Riemannian Batch Normalization
by: Chen, Ziheng, et al.
Published: (2024)
by: Chen, Ziheng, et al.
Published: (2024)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
by: Guo, Zihao, et al.
Published: (2026)
by: Guo, Zihao, et al.
Published: (2026)
Causal Reinforcement Learning for Complex Card Games: A Magic The Gathering Benchmark
by: Cunha, Cristiano da Costa, et al.
Published: (2026)
by: Cunha, Cristiano da Costa, et al.
Published: (2026)
AI-Generated Content (AIGC) for Various Data Modalities: A Survey
by: Foo, Lin Geng, et al.
Published: (2023)
by: Foo, Lin Geng, et al.
Published: (2023)
ManiPose: A Comprehensive Benchmark for Pose-aware Object Manipulation in Robotics
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
UNOPose: Unseen Object Pose Estimation with an Unposed RGB-D Reference Image
by: Liu, Xingyu, et al.
Published: (2024)
by: Liu, Xingyu, et al.
Published: (2024)
Reverse Personalization
by: Kung, Han-Wei, et al.
Published: (2025)
by: Kung, Han-Wei, et al.
Published: (2025)
SM4Depth: Seamless Monocular Metric Depth Estimation across Multiple Cameras and Scenes by One Model
by: Liu, Yihao, et al.
Published: (2024)
by: Liu, Yihao, et al.
Published: (2024)
Occlusion-aware Text-Image-Point Cloud Pretraining for Open-World 3D Object Recognition
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
Similar Items
-
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation
by: Liu, Jian, et al.
Published: (2025) -
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
by: Yang, Hui, et al.
Published: (2026) -
Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration
by: Liu, Jian, et al.
Published: (2025) -
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025) -
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
by: Yang, Hui, et al.
Published: (2025)