On the Use of Hierarchical Vision Foundation Models for Low-Cost Human Mesh Recovery and Pose Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Tarashima, Shuhei, Wang, Yushan, Tagawa, Norio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ViLAaD: Enhancing "Attracting and Dispersing'' Source-Free Domain Adaptation with Vision-and-Language Model
by: Tarashima, Shuhei, et al.
Published: (2025)
by: Tarashima, Shuhei, et al.
Published: (2025)
Plug-and-Play Acceleration of Occupancy Grid-based NeRF Rendering using VDB Grid and Hierarchical Ray Traversal
by: Kato, Yoshio, et al.
Published: (2024)
by: Kato, Yoshio, et al.
Published: (2024)
Validation of Human Pose Estimation and Human Mesh Recovery for Extracting Clinically Relevant Motion Data from Videos
by: Armstrong, Kai, et al.
Published: (2025)
by: Armstrong, Kai, et al.
Published: (2025)
Adapting Human Mesh Recovery with Vision-Language Feedback
by: Xu, Chongyang, et al.
Published: (2025)
by: Xu, Chongyang, et al.
Published: (2025)
PostoMETRO: Pose Token Enhanced Mesh Transformer for Robust 3D Human Mesh Recovery
by: Yang, Wendi, et al.
Published: (2024)
by: Yang, Wendi, et al.
Published: (2024)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
Multi-Human Mesh Recovery with Transformers
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
Utilizing Uncertainty in 2D Pose Detectors for Probabilistic 3D Human Mesh Recovery
by: Wehrbein, Tom, et al.
Published: (2024)
by: Wehrbein, Tom, et al.
Published: (2024)
Diffusion Models are Efficient Data Generators for Human Mesh Recovery
by: Ge, Yongtao, et al.
Published: (2024)
by: Ge, Yongtao, et al.
Published: (2024)
PromptHMR: Promptable Human Mesh Recovery
by: Wang, Yufu, et al.
Published: (2025)
by: Wang, Yufu, et al.
Published: (2025)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
by: Jeong, David C., et al.
Published: (2025)
by: Jeong, David C., et al.
Published: (2025)
DeforHMR: Vision Transformer with Deformable Cross-Attention for 3D Human Mesh Recovery
by: Heo, Jaewoo, et al.
Published: (2024)
by: Heo, Jaewoo, et al.
Published: (2024)
Monocular Models are Strong Learners for Multi-View Human Mesh Recovery
by: Xie, Haoyu, et al.
Published: (2026)
by: Xie, Haoyu, et al.
Published: (2026)
InterMesh: Explicit Interaction-Aware End-to-End Multi-Person Human Mesh Recovery
by: Zheng, Kaili, et al.
Published: (2026)
by: Zheng, Kaili, et al.
Published: (2026)
Anny-Fit: All-Age Human Mesh Recovery
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
MEGA: Masked Generative Autoencoder for Human Mesh Recovery
by: Fiche, Guénolé, et al.
Published: (2024)
by: Fiche, Guénolé, et al.
Published: (2024)
Efficient Diffusion-Based 3D Human Pose Estimation with Hierarchical Temporal Pruning
by: Bi, Yuquan, et al.
Published: (2025)
by: Bi, Yuquan, et al.
Published: (2025)
Latent-Info and Low-Dimensional Learning for Human Mesh Recovery and Parallel Optimization
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
by: Fang, Hongwei, et al.
Published: (2026)
by: Fang, Hongwei, et al.
Published: (2026)
Human Mesh Recovery from Arbitrary Multi-view Images
by: Li, Xiaoben, et al.
Published: (2024)
by: Li, Xiaoben, et al.
Published: (2024)
DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh Recovery
by: Zhu, Yixuan, et al.
Published: (2024)
by: Zhu, Yixuan, et al.
Published: (2024)
Cross-Domain Knowledge Distillation for Low-Resolution Human Pose Estimation
by: Gu, Zejun, et al.
Published: (2024)
by: Gu, Zejun, et al.
Published: (2024)
Focus on Low-Resolution Information: Multi-Granular Information-Lossless Model for Low-Resolution Human Pose Estimation
by: Gu, Zejun, et al.
Published: (2024)
by: Gu, Zejun, et al.
Published: (2024)
Implicit Modeling for Transferability Estimation of Vision Foundation Models
by: Zheng, Yaoyan, et al.
Published: (2025)
by: Zheng, Yaoyan, et al.
Published: (2025)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
by: Zhao, Yiwen, et al.
Published: (2026)
by: Zhao, Yiwen, et al.
Published: (2026)
VLM-Guided Group Preference Alignment for Diffusion-based Human Mesh Recovery
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)
by: Kim, Donghwan, et al.
Published: (2025)
HeatFormer: A Neural Optimizer for Multiview Human Mesh Recovery
by: Matsubara, Yuto, et al.
Published: (2024)
by: Matsubara, Yuto, et al.
Published: (2024)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
by: Chen, Hongjun, et al.
Published: (2026)
by: Chen, Hongjun, et al.
Published: (2026)
Egocentric Whole-Body Human Mesh Recovery with Prior-Guided Learning
by: Na, Soyeon, et al.
Published: (2026)
by: Na, Soyeon, et al.
Published: (2026)
Humans as Checkerboards: Calibrating Camera Motion Scale for World-Coordinate Human Mesh Recovery
by: Yang, Fengyuan, et al.
Published: (2024)
by: Yang, Fengyuan, et al.
Published: (2024)
$\text{Di}^2\text{Pose}$: Discrete Diffusion Model for Occluded 3D Human Pose Estimation
by: Wang, Weiquan, et al.
Published: (2024)
by: Wang, Weiquan, et al.
Published: (2024)
SMART: SMPLest-X Mesh Adaptation and RAFT Tracking for Soccer Pose Estimation
by: Rawat, Parthsarthi
Published: (2026)
by: Rawat, Parthsarthi
Published: (2026)
Robust Low-Light Human Pose Estimation through Illumination-Texture Modulation
by: Zhang, Feng, et al.
Published: (2025)
by: Zhang, Feng, et al.
Published: (2025)
Disentangled Diffusion-Based 3D Human Pose Estimation with Hierarchical Spatial and Temporal Denoiser
by: Cai, Qingyuan, et al.
Published: (2024)
by: Cai, Qingyuan, et al.
Published: (2024)
Sapiens: Foundation for Human Vision Models
by: Khirodkar, Rawal, et al.
Published: (2024)
by: Khirodkar, Rawal, et al.
Published: (2024)
FoundPose: Unseen Object Pose Estimation with Foundation Features
by: Örnek, Evin Pınar, et al.
Published: (2023)
by: Örnek, Evin Pınar, et al.
Published: (2023)
3D Human Mesh Estimation from Virtual Markers
by: Ma, Xiaoxuan, et al.
Published: (2023)
by: Ma, Xiaoxuan, et al.
Published: (2023)
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026)
by: Cen, Jiaxin, et al.
Published: (2026)
Similar Items
-
ViLAaD: Enhancing "Attracting and Dispersing'' Source-Free Domain Adaptation with Vision-and-Language Model
by: Tarashima, Shuhei, et al.
Published: (2025) -
Plug-and-Play Acceleration of Occupancy Grid-based NeRF Rendering using VDB Grid and Hierarchical Ray Traversal
by: Kato, Yoshio, et al.
Published: (2024) -
Validation of Human Pose Estimation and Human Mesh Recovery for Extracting Clinically Relevant Motion Data from Videos
by: Armstrong, Kai, et al.
Published: (2025) -
Adapting Human Mesh Recovery with Vision-Language Feedback
by: Xu, Chongyang, et al.
Published: (2025) -
PostoMETRO: Pose Token Enhanced Mesh Transformer for Robust 3D Human Mesh Recovery
by: Yang, Wendi, et al.
Published: (2024)