CLIPose: Category-Level Object Pose Estimation with Pre-trained Vision-Language Knowledge
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Xiao, Zhu, Minghao, Dang, Ronghao, Zhou, Guangliang, Shu, Shaolong, Lin, Feng, Liu, Chengju, Chen, Qijun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CleanPose: Category-Level Object Pose Estimation via Causal Learning and Knowledge Distillation
by: Lin, Xiao, et al.
Published: (2025)
by: Lin, Xiao, et al.
Published: (2025)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023)
by: Lin, Xiao, et al.
Published: (2023)
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning
by: Zhu, Minghao, et al.
Published: (2023)
by: Zhu, Minghao, et al.
Published: (2023)
MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer
by: Zhu, Minghao, et al.
Published: (2024)
by: Zhu, Minghao, et al.
Published: (2024)
Vision-and-Language Navigation via Causal Learning
by: Wang, Liuyi, et al.
Published: (2024)
by: Wang, Liuyi, et al.
Published: (2024)
Causality-based Cross-Modal Representation Learning for Vision-and-Language Navigation
by: Wang, Liuyi, et al.
Published: (2024)
by: Wang, Liuyi, et al.
Published: (2024)
A Dual Semantic-Aware Recurrent Global-Adaptive Network For Vision-and-Language Navigation
by: Wang, Liuyi, et al.
Published: (2023)
by: Wang, Liuyi, et al.
Published: (2023)
InstructDET: Diversifying Referring Object Detection with Generalized Instructions
by: Dang, Ronghao, et al.
Published: (2023)
by: Dang, Ronghao, et al.
Published: (2023)
Instance-Adaptive and Geometric-Aware Keypoint Learning for Category-Level 6D Object Pose Estimation
by: Lin, Xiao, et al.
Published: (2024)
by: Lin, Xiao, et al.
Published: (2024)
MLANet: Multi-Level Attention Network with Sub-instruction for Continuous Vision-and-Language Navigation
by: He, Zongtao, et al.
Published: (2023)
by: He, Zongtao, et al.
Published: (2023)
Universal Features Guided Zero-Shot Category-Level Object Pose Estimation
by: Qu, Wentian, et al.
Published: (2025)
by: Qu, Wentian, et al.
Published: (2025)
Category-Level and Open-Set Object Pose Estimation for Robotics
by: Hönig, Peter, et al.
Published: (2025)
by: Hönig, Peter, et al.
Published: (2025)
Efficient Text-driven Motion Generation via Latent Consistency Training
by: Hu, Mengxian, et al.
Published: (2024)
by: Hu, Mengxian, et al.
Published: (2024)
DecomPose: Disentangling Cross-Category Optimization Contention for Category-Level 6D Object Pose Estimation
by: Gao, Yifan, et al.
Published: (2026)
by: Gao, Yifan, et al.
Published: (2026)
ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
by: Ren, Huan, et al.
Published: (2026)
by: Ren, Huan, et al.
Published: (2026)
TSM-Pose: Topology-Aware Learning with Semantic Mamba for Category-Level Object Pose Estimation
by: Liu, Jinshuo, et al.
Published: (2026)
by: Liu, Jinshuo, et al.
Published: (2026)
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
CapeNext: Rethinking and Refining Dynamic Support Information for Category-Agnostic Pose Estimation
by: Zhu, Yu, et al.
Published: (2025)
by: Zhu, Yu, et al.
Published: (2025)
PASTS: Progress-Aware Spatio-Temporal Transformer Speaker For Vision-and-Language Navigation
by: Wang, Liuyi, et al.
Published: (2023)
by: Wang, Liuyi, et al.
Published: (2023)
Category-Level Object Shape and Pose Estimation in Less Than a Millisecond
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
THE-Pose: Topological Prior with Hybrid Graph Fusion for Estimating Category-Level 6D Object Pose
by: Lee, Eunho, et al.
Published: (2025)
by: Lee, Eunho, et al.
Published: (2025)
LaPose: Laplacian Mixture Shape Modeling for RGB-Based Category-Level Object Pose Estimation
by: Zhang, Ruida, et al.
Published: (2024)
by: Zhang, Ruida, et al.
Published: (2024)
Instance-Adaptive Keypoint Learning with Local-to-Global Geometric Aggregation for Category-Level Object Pose Estimation
by: Zhang, Xiao, et al.
Published: (2025)
by: Zhang, Xiao, et al.
Published: (2025)
Omni6D: Large-Vocabulary 3D Object Dataset for Category-Level 6D Object Pose Estimation
by: Zhang, Mengchen, et al.
Published: (2024)
by: Zhang, Mengchen, et al.
Published: (2024)
NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization
by: He, Zongtao, et al.
Published: (2025)
by: He, Zongtao, et al.
Published: (2025)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
Learning Shape-Independent Transformation via Spherical Representations for Category-Level Object Pose Estimation
by: Ren, Huan, et al.
Published: (2025)
by: Ren, Huan, et al.
Published: (2025)
RCGNet: RGB-based Category-Level 6D Object Pose Estimation with Geometric Guidance
by: Yu, Sheng, et al.
Published: (2025)
by: Yu, Sheng, et al.
Published: (2025)
Learning a Category-level Object Pose Estimator without Pose Annotations
by: Tian, Fengrui, et al.
Published: (2024)
by: Tian, Fengrui, et al.
Published: (2024)
GCE-Pose: Global Context Enhancement for Category-level Object Pose Estimation
by: Li, Weihang, et al.
Published: (2025)
by: Li, Weihang, et al.
Published: (2025)
SCOPE: Semantic Conditioning for Sim2Real Category-Level Object Pose Estimation in Robotics
by: Hönig, Peter, et al.
Published: (2025)
by: Hönig, Peter, et al.
Published: (2025)
Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View
by: Zhang, Jinyu, et al.
Published: (2025)
by: Zhang, Jinyu, et al.
Published: (2025)
Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose Estimation
by: Kaushik, Prakhar, et al.
Published: (2024)
by: Kaushik, Prakhar, et al.
Published: (2024)
GIVEPose: Gradual Intra-class Variation Elimination for RGB-based Category-Level Object Pose Estimation
by: Huang, Zinqin, et al.
Published: (2025)
by: Huang, Zinqin, et al.
Published: (2025)
Learning Point Cloud Representations with Pose Continuity for Depth-Based Category-Level 6D Object Pose Estimation
by: Li, Zhujun, et al.
Published: (2025)
by: Li, Zhujun, et al.
Published: (2025)
MAGIC: Meta-Ability Guided Interactive Chain-of-Distillation for Effective-and-Efficient Vision-and-Language Navigation
by: Wang, Liuyi, et al.
Published: (2024)
by: Wang, Liuyi, et al.
Published: (2024)
Unified Category-Level Object Detection and Pose Estimation from RGB Images using 3D Prototypes
by: Fischer, Tom, et al.
Published: (2025)
by: Fischer, Tom, et al.
Published: (2025)
Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion
by: Bethell, Adam, et al.
Published: (2024)
by: Bethell, Adam, et al.
Published: (2024)
FastPoseCNN: Real-Time Monocular Category-Level Pose and Size Estimation Framework
by: Davalos, Eduardo, et al.
Published: (2024)
by: Davalos, Eduardo, et al.
Published: (2024)
Lightweight Model Pre-training via Language Guided Knowledge Distillation
by: Li, Mingsheng, et al.
Published: (2024)
by: Li, Mingsheng, et al.
Published: (2024)
Similar Items
-
CleanPose: Category-Level Object Pose Estimation via Causal Learning and Knowledge Distillation
by: Lin, Xiao, et al.
Published: (2025) -
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023) -
Fine-Grained Spatiotemporal Motion Alignment for Contrastive Video Representation Learning
by: Zhu, Minghao, et al.
Published: (2023) -
MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer
by: Zhu, Minghao, et al.
Published: (2024) -
Vision-and-Language Navigation via Causal Learning
by: Wang, Liuyi, et al.
Published: (2024)