Towards Balanced Multi-Modal Learning in 3D Human Pose Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Mengshi, Peng, Jiaxuan, Zhang, Xianlin, Ma, Huadong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Action Quality Assessment via Hierarchical Pose-guided Multi-stage Contrastive Regression
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Chain-of-Evidence Multimodal Reasoning for Few-shot Temporal Action Localization
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
A New Teacher-Reviewer-Student Framework for Semi-supervised 2D Human Pose Estimation
von: Yun, Wulian, et al.
Veröffentlicht: (2025)
von: Yun, Wulian, et al.
Veröffentlicht: (2025)
Towards Robust Unsupervised Attention Prediction in Autonomous Driving
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Multi-Stage Contrastive Regression for Action Quality Assessment
von: An, Qi, et al.
Veröffentlicht: (2024)
von: An, Qi, et al.
Veröffentlicht: (2024)
DreamPose3D: Hallucinative Diffusion with Prompt Learning for 3D Human Pose Estimation
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
von: Bright, Jerrin, et al.
Veröffentlicht: (2025)
Learning Group Interactions and Semantic Intentions for Multi-Object Trajectory Prediction
von: Qi, Mengshi, et al.
Veröffentlicht: (2024)
von: Qi, Mengshi, et al.
Veröffentlicht: (2024)
JRDB-Pose3D: A Multi-person 3D Human Pose and Shape Estimation Dataset for Robotics
von: Biswas, Sandika, et al.
Veröffentlicht: (2026)
von: Biswas, Sandika, et al.
Veröffentlicht: (2026)
PoseGRAF: Geometric-Reinforced Adaptive Fusion for Monocular 3D Human Pose Estimation
von: Xu, Ming, et al.
Veröffentlicht: (2025)
von: Xu, Ming, et al.
Veröffentlicht: (2025)
Decomposed Vector-Quantized Variational Autoencoder for Human Grasp Generation
von: Zhao, Zhe, et al.
Veröffentlicht: (2024)
von: Zhao, Zhe, et al.
Veröffentlicht: (2024)
PoseMoE: Mixture-of-Experts Network for Monocular 3D Human Pose Estimation
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
Global-Local Tree Search in VLMs for 3D Indoor Scene Generation
von: Deng, Wei, et al.
Veröffentlicht: (2025)
von: Deng, Wei, et al.
Veröffentlicht: (2025)
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
Unconstrained Multi-view Human Pose Estimation with Algebraic Priors
von: Qin, Xiaolin, et al.
Veröffentlicht: (2026)
von: Qin, Xiaolin, et al.
Veröffentlicht: (2026)
Lifelong Domain Adaptive 3D Human Pose Estimation
von: Peng, Qucheng, et al.
Veröffentlicht: (2025)
von: Peng, Qucheng, et al.
Veröffentlicht: (2025)
ManiPose: Manifold-Constrained Multi-Hypothesis 3D Human Pose Estimation
von: Rommel, Cédric, et al.
Veröffentlicht: (2023)
von: Rommel, Cédric, et al.
Veröffentlicht: (2023)
Toward Efficient Generalization in 3D Human Pose Estimation via a Canonical Domain Approach
von: Lee, Hoosang, et al.
Veröffentlicht: (2025)
von: Lee, Hoosang, et al.
Veröffentlicht: (2025)
Robust Disentangled Counterfactual Learning for Physical Audiovisual Commonsense Reasoning
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Toward a Real-Time Framework for Accurate Monocular 3D Human Pose Estimation with Geometric Priors
von: Adjel, Mohamed
Veröffentlicht: (2025)
von: Adjel, Mohamed
Veröffentlicht: (2025)
Cooperative Inference for Real-Time 3D Human Pose Estimation in Multi-Device Edge Networks
von: Choi, Hyun-Ho, et al.
Veröffentlicht: (2025)
von: Choi, Hyun-Ho, et al.
Veröffentlicht: (2025)
Active Exploring like a Pigeon: Reinforcing Spatial Reasoning via Agentic Vision-Language Models
von: Deng, Wei, et al.
Veröffentlicht: (2026)
von: Deng, Wei, et al.
Veröffentlicht: (2026)
Multi-Modal Parameter-Efficient Fine-tuning via Graph Neural Network
von: Cheng, Bin, et al.
Veröffentlicht: (2024)
von: Cheng, Bin, et al.
Veröffentlicht: (2024)
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation
von: Ye, Mingrui, et al.
Veröffentlicht: (2025)
von: Ye, Mingrui, et al.
Veröffentlicht: (2025)
SEAL-pose: Enhancing 3D Human Pose Estimation via a Learned Loss for Structural Consistency
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
von: Kim, Yeonsung, et al.
Veröffentlicht: (2026)
HYPERPOSE: Hyperbolic Kinematic Phase-Space Attention for 3D Human Pose Estimation
von: Thekkath, Vinduja, et al.
Veröffentlicht: (2026)
von: Thekkath, Vinduja, et al.
Veröffentlicht: (2026)
Pre-training a Density-Aware Pose Transformer for Robust LiDAR-based 3D Human Pose Estimation
von: An, Xiaoqi, et al.
Veröffentlicht: (2024)
von: An, Xiaoqi, et al.
Veröffentlicht: (2024)
Towards Efficient Object Re-Identification with A Novel Cloud-Edge Collaborative Framework
von: Wang, Chuanming, et al.
Veröffentlicht: (2024)
von: Wang, Chuanming, et al.
Veröffentlicht: (2024)
Question-Aware Evidence Ledgers for Video Relational Reasoning
von: Ou, Yilin, et al.
Veröffentlicht: (2026)
von: Ou, Yilin, et al.
Veröffentlicht: (2026)
Survey of 3D Human Body Pose and Shape Estimation Methods for Contemporary Dance Applications
von: Venkatrayappa, Darshan, et al.
Veröffentlicht: (2024)
von: Venkatrayappa, Darshan, et al.
Veröffentlicht: (2024)
PoseBench: Benchmarking the Robustness of Pose Estimation Models under Corruptions
von: Ma, Sihan, et al.
Veröffentlicht: (2024)
von: Ma, Sihan, et al.
Veröffentlicht: (2024)
MoViD: View-Invariant 3D Human Pose Estimation via Motion-View Disentanglement
von: Liu, Yejia, et al.
Veröffentlicht: (2026)
von: Liu, Yejia, et al.
Veröffentlicht: (2026)
Semi-Supervised Teacher-Reference-Student Architecture for Action Quality Assessment
von: Yun, Wulian, et al.
Veröffentlicht: (2024)
von: Yun, Wulian, et al.
Veröffentlicht: (2024)
Robo-SGG: Exploiting Layout-Oriented Normalization and Restitution Can Improve Robust Scene Graph Generation
von: Lv, Changsheng, et al.
Veröffentlicht: (2025)
von: Lv, Changsheng, et al.
Veröffentlicht: (2025)
SpaceSense-Bench: A Large-Scale Multi-Modal Benchmark for Spacecraft Perception and Pose Estimation
von: Wu, Aodi, et al.
Veröffentlicht: (2026)
von: Wu, Aodi, et al.
Veröffentlicht: (2026)
Towards Multi-Modal Animal Pose Estimation: A Survey and In-Depth Analysis
von: Deng, Qianyi, et al.
Veröffentlicht: (2024)
von: Deng, Qianyi, et al.
Veröffentlicht: (2024)
RT-Pose: A 4D Radar Tensor-based 3D Human Pose Estimation and Localization Benchmark
von: Ho, Yuan-Hao, et al.
Veröffentlicht: (2024)
von: Ho, Yuan-Hao, et al.
Veröffentlicht: (2024)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
MAEPose: Self-Supervised Spatiotemporal Learning for Human Pose Estimation on mmWave Video
von: Wei, Xijia, et al.
Veröffentlicht: (2026)
von: Wei, Xijia, et al.
Veröffentlicht: (2026)
E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control
von: Gu, Qiao, et al.
Veröffentlicht: (2026)
von: Gu, Qiao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Action Quality Assessment via Hierarchical Pose-guided Multi-stage Contrastive Regression
von: Qi, Mengshi, et al.
Veröffentlicht: (2025) -
Chain-of-Evidence Multimodal Reasoning for Few-shot Temporal Action Localization
von: Qi, Mengshi, et al.
Veröffentlicht: (2025) -
A New Teacher-Reviewer-Student Framework for Semi-supervised 2D Human Pose Estimation
von: Yun, Wulian, et al.
Veröffentlicht: (2025) -
Towards Robust Unsupervised Attention Prediction in Autonomous Driving
von: Qi, Mengshi, et al.
Veröffentlicht: (2025) -
Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)