Making Pose Representations More Expressive and Disentangled via Residual Vector Quantization
Fuente:
arXiv
Saved in:
| Main Authors: | Jeong, Sukhyun, Shin, Hong-Gi, Choi, Yong-Hoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pose-Guided Residual Refinement for Interpretable Text-to-Motion Generation and Editing
by: Jeong, Sukhyun, et al.
Published: (2025)
by: Jeong, Sukhyun, et al.
Published: (2025)
Residual Vector Quantization For Communication-Efficient Multi-Agent Perception
by: Shenkut, Dereje, et al.
Published: (2025)
by: Shenkut, Dereje, et al.
Published: (2025)
KGpose: Keypoint-Graph Driven End-to-End Multi-Object 6D Pose Estimation via Point-Wise Pose Voting
by: Jeong, Andrew
Published: (2024)
by: Jeong, Andrew
Published: (2024)
6D Pose Estimation via Keypoint Heatmap Regression with RGB-D Residual Neural Networks
by: Aljosevic, Ismail, et al.
Published: (2026)
by: Aljosevic, Ismail, et al.
Published: (2026)
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
by: Wang, Yating, et al.
Published: (2025)
by: Wang, Yating, et al.
Published: (2025)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Multimodal Graph Representation Learning for Robust Surgical Workflow Recognition with Adversarial Feature Disentanglement
by: Bai, Long, et al.
Published: (2025)
by: Bai, Long, et al.
Published: (2025)
MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
CA-Cut: Crop-Aligned Cutout for Data Augmentation to Learn More Robust Under-Canopy Navigation
by: Mamo, Robel, et al.
Published: (2025)
by: Mamo, Robel, et al.
Published: (2025)
ViTaSCOPE: Visuo-tactile Implicit Representation for In-hand Pose and Extrinsic Contact Estimation
by: Lee, Jayjun, et al.
Published: (2025)
by: Lee, Jayjun, et al.
Published: (2025)
Modular Quantization-Aware Training for 6D Object Pose Estimation
by: Javed, Saqib, et al.
Published: (2023)
by: Javed, Saqib, et al.
Published: (2023)
Probabilistic Rotation Representation With an Efficiently Computable Bingham Loss Function and Its Application to Pose Estimation
by: Sato, Hiroya, et al.
Published: (2022)
by: Sato, Hiroya, et al.
Published: (2022)
Learning Scene-Level Signed Directional Distance Function with Ellipsoidal Priors and Neural Residuals
by: Dai, Zhirui, et al.
Published: (2025)
by: Dai, Zhirui, et al.
Published: (2025)
Post-Training and Test-Time Scaling of Generative Agent Behavior Models for Interactive Autonomous Driving
by: Seong, Hyunki, et al.
Published: (2025)
by: Seong, Hyunki, et al.
Published: (2025)
VectorWorld: Efficient Streaming World Model via Diffusion Flow on Vector Graphs
by: Jiang, Chaokang, et al.
Published: (2026)
by: Jiang, Chaokang, et al.
Published: (2026)
SMPLest-X: Ultimate Scaling for Expressive Human Pose and Shape Estimation
by: Yin, Wanqi, et al.
Published: (2025)
by: Yin, Wanqi, et al.
Published: (2025)
Incremental Joint Learning of Depth, Pose and Implicit Scene Representation on Monocular Camera in Large-scale Scenes
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
MapGCLR: Geospatial Contrastive Learning of Representations for Online Vectorized HD Map Construction
by: Merkert, Jonas, et al.
Published: (2026)
by: Merkert, Jonas, et al.
Published: (2026)
ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation
by: Ren, Huan, et al.
Published: (2026)
by: Ren, Huan, et al.
Published: (2026)
Active Human Pose Estimation via an Autonomous UAV Agent
by: Chen, Jingxi, et al.
Published: (2024)
by: Chen, Jingxi, et al.
Published: (2024)
Perception with Guarantees: Certified Pose Estimation via Reachability Analysis
by: Ladner, Tobias, et al.
Published: (2026)
by: Ladner, Tobias, et al.
Published: (2026)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
by: Zhang, Wenyao, et al.
Published: (2026)
by: Zhang, Wenyao, et al.
Published: (2026)
FoundPose: Unseen Object Pose Estimation with Foundation Features
by: Örnek, Evin Pınar, et al.
Published: (2023)
by: Örnek, Evin Pınar, et al.
Published: (2023)
PhysPose: Refining 6D Object Poses with Physical Constraints
by: Malenický, Martin, et al.
Published: (2025)
by: Malenický, Martin, et al.
Published: (2025)
SpyroPose: SE(3) Pyramids for Object Pose Distribution Estimation
by: Haugaard, Rasmus Laurvig, et al.
Published: (2023)
by: Haugaard, Rasmus Laurvig, et al.
Published: (2023)
SplatPose: Geometry-Aware 6-DoF Pose Estimation from Single RGB Image via 3D Gaussian Splatting
by: Yang, Linqi, et al.
Published: (2025)
by: Yang, Linqi, et al.
Published: (2025)
Breathless: An 8-hour Performance Contrasting Human and Robot Expressiveness
by: Cuan, Catie, et al.
Published: (2024)
by: Cuan, Catie, et al.
Published: (2024)
Learning High-resolution Vector Representation from Multi-Camera Images for 3D Object Detection
by: Chen, Zhili, et al.
Published: (2024)
by: Chen, Zhili, et al.
Published: (2024)
UnPose: Uncertainty-Guided Diffusion Priors for Zero-Shot Pose Estimation
by: Jiang, Zhaodong, et al.
Published: (2025)
by: Jiang, Zhaodong, et al.
Published: (2025)
FlyPose: Towards Robust Human Pose Estimation From Aerial Views
by: Farooq, Hassaan, et al.
Published: (2026)
by: Farooq, Hassaan, et al.
Published: (2026)
LocPoseNet: Robust Location Prior for Unseen Object Pose Estimation
by: Zhao, Chen, et al.
Published: (2022)
by: Zhao, Chen, et al.
Published: (2022)
W-PoseNet: Dense Correspondence Regularized Pixel Pair Pose Regression
by: Xu, Zelin, et al.
Published: (2019)
by: Xu, Zelin, et al.
Published: (2019)
ManiPose: A Comprehensive Benchmark for Pose-aware Object Manipulation in Robotics
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
PROFusion: Robust and Accurate Dense Reconstruction via Camera Pose Regression and Optimization
by: Dong, Siyan, et al.
Published: (2025)
by: Dong, Siyan, et al.
Published: (2025)
Efficient Long-Horizon Vision-Language-Action Models via Static-Dynamic Disentanglement
by: Qiu, Weikang, et al.
Published: (2026)
by: Qiu, Weikang, et al.
Published: (2026)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
by: Wu, Zijian, et al.
Published: (2025)
by: Wu, Zijian, et al.
Published: (2025)
PoseINN: Realtime Visual-based Pose Regression and Localization with Invertible Neural Networks
by: Zang, Zirui, et al.
Published: (2024)
by: Zang, Zirui, et al.
Published: (2024)
Multi-view Pose Fusion for Occlusion-Aware 3D Human Pose Estimation
by: Bragagnolo, Laura, et al.
Published: (2024)
by: Bragagnolo, Laura, et al.
Published: (2024)
Similar Items
-
Pose-Guided Residual Refinement for Interpretable Text-to-Motion Generation and Editing
by: Jeong, Sukhyun, et al.
Published: (2025) -
Residual Vector Quantization For Communication-Efficient Multi-Agent Perception
by: Shenkut, Dereje, et al.
Published: (2025) -
KGpose: Keypoint-Graph Driven End-to-End Multi-Object 6D Pose Estimation via Point-Wise Pose Voting
by: Jeong, Andrew
Published: (2024) -
6D Pose Estimation via Keypoint Heatmap Regression with RGB-D Residual Neural Networks
by: Aljosevic, Ismail, et al.
Published: (2026) -
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers
by: Wang, Yating, et al.
Published: (2025)