Adapting Human Mesh Recovery with Vision-Language Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Chongyang, Huang, Buzhen, Zhang, Chengfang, Feng, Ziliang, Wang, Yangang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
by: Huang, Buzhen, et al.
Published: (2024)
by: Huang, Buzhen, et al.
Published: (2024)
Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption
by: Huang, Buzhen, et al.
Published: (2024)
by: Huang, Buzhen, et al.
Published: (2024)
Reconstructing Close Human Interaction with Appearance and Proxemics Reasoning
by: Huang, Buzhen, et al.
Published: (2025)
by: Huang, Buzhen, et al.
Published: (2025)
Contrastive Multi-Modal Hypergraph Reasoning for 3D Crowd Mesh Recovery
by: Sun, Minghao, et al.
Published: (2026)
by: Sun, Minghao, et al.
Published: (2026)
E-React: Towards Emotionally Controlled Synthesis of Human Reactions
by: Zhu, Chen, et al.
Published: (2025)
by: Zhu, Chen, et al.
Published: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
by: Xu, Jiahao, et al.
Published: (2026)
by: Xu, Jiahao, et al.
Published: (2026)
Synthesizing Physically Plausible Human Motions in 3D Scenes
by: Pan, Liang, et al.
Published: (2023)
by: Pan, Liang, et al.
Published: (2023)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
Multi-Human Mesh Recovery with Transformers
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
Ada3Drift: Adaptive Training-Time Drifting for One-Step 3D Visuomotor Robotic Manipulation
by: Xu, Chongyang, et al.
Published: (2026)
by: Xu, Chongyang, et al.
Published: (2026)
PromptHMR: Promptable Human Mesh Recovery
by: Wang, Yufu, et al.
Published: (2025)
by: Wang, Yufu, et al.
Published: (2025)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
by: Jeong, David C., et al.
Published: (2025)
by: Jeong, David C., et al.
Published: (2025)
On the Use of Hierarchical Vision Foundation Models for Low-Cost Human Mesh Recovery and Pose Estimation
by: Tarashima, Shuhei, et al.
Published: (2025)
by: Tarashima, Shuhei, et al.
Published: (2025)
DeforHMR: Vision Transformer with Deformable Cross-Attention for 3D Human Mesh Recovery
by: Heo, Jaewoo, et al.
Published: (2024)
by: Heo, Jaewoo, et al.
Published: (2024)
Diffusion Models are Efficient Data Generators for Human Mesh Recovery
by: Ge, Yongtao, et al.
Published: (2024)
by: Ge, Yongtao, et al.
Published: (2024)
Incorporating Test-Time Optimization into Training with Dual Networks for Human Mesh Recovery
by: Nie, Yongwei, et al.
Published: (2024)
by: Nie, Yongwei, et al.
Published: (2024)
Monocular Models are Strong Learners for Multi-View Human Mesh Recovery
by: Xie, Haoyu, et al.
Published: (2026)
by: Xie, Haoyu, et al.
Published: (2026)
InterMesh: Explicit Interaction-Aware End-to-End Multi-Person Human Mesh Recovery
by: Zheng, Kaili, et al.
Published: (2026)
by: Zheng, Kaili, et al.
Published: (2026)
HeRO: Hierarchical 3D Semantic Representation for Pose-aware Object Manipulation
by: Xu, Chongyang, et al.
Published: (2026)
by: Xu, Chongyang, et al.
Published: (2026)
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
Anny-Fit: All-Age Human Mesh Recovery
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
MEGA: Masked Generative Autoencoder for Human Mesh Recovery
by: Fiche, Guénolé, et al.
Published: (2024)
by: Fiche, Guénolé, et al.
Published: (2024)
LiDAR-HMR: 3D Human Mesh Recovery from LiDAR
by: Fan, Bohao, et al.
Published: (2023)
by: Fan, Bohao, et al.
Published: (2023)
DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
MetricHMSR:Metric Human Mesh and Scene Recovery from Monocular Images
by: Song, Chentao, et al.
Published: (2025)
by: Song, Chentao, et al.
Published: (2025)
Human Mesh Recovery from Arbitrary Multi-view Images
by: Li, Xiaoben, et al.
Published: (2024)
by: Li, Xiaoben, et al.
Published: (2024)
DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh Recovery
by: Zhu, Yixuan, et al.
Published: (2024)
by: Zhu, Yixuan, et al.
Published: (2024)
DiffProxy: Multi-View Human Mesh Recovery via Diffusion-Generated Dense Proxies
by: Wang, Renke, et al.
Published: (2026)
by: Wang, Renke, et al.
Published: (2026)
Multi-RoI Human Mesh Recovery with Camera Consistency and Contrastive Losses
by: Nie, Yongwei, et al.
Published: (2024)
by: Nie, Yongwei, et al.
Published: (2024)
PostoMETRO: Pose Token Enhanced Mesh Transformer for Robust 3D Human Mesh Recovery
by: Yang, Wendi, et al.
Published: (2024)
by: Yang, Wendi, et al.
Published: (2024)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
by: Zhao, Yiwen, et al.
Published: (2026)
by: Zhao, Yiwen, et al.
Published: (2026)
MoPO: Incorporating Motion Prior for Occluded Human Mesh Recovery
by: Tang, Tao, et al.
Published: (2026)
by: Tang, Tao, et al.
Published: (2026)
VLM-Guided Group Preference Alignment for Diffusion-based Human Mesh Recovery
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
Generalizable Human Gaussians from Single-View Image
by: Chen, Jinnan, et al.
Published: (2024)
by: Chen, Jinnan, et al.
Published: (2024)
Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models
by: Huang, Yuhang, et al.
Published: (2024)
by: Huang, Yuhang, et al.
Published: (2024)
Action-Geometry Prediction with 3D Geometric Prior for Bimanual Manipulation
by: Xu, Chongyang, et al.
Published: (2026)
by: Xu, Chongyang, et al.
Published: (2026)
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)
by: Kim, Donghwan, et al.
Published: (2025)
HeatFormer: A Neural Optimizer for Multiview Human Mesh Recovery
by: Matsubara, Yuto, et al.
Published: (2024)
by: Matsubara, Yuto, et al.
Published: (2024)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
by: Chen, Hongjun, et al.
Published: (2026)
by: Chen, Hongjun, et al.
Published: (2026)
Egocentric Whole-Body Human Mesh Recovery with Prior-Guided Learning
by: Na, Soyeon, et al.
Published: (2026)
by: Na, Soyeon, et al.
Published: (2026)
Similar Items
-
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
by: Huang, Buzhen, et al.
Published: (2024) -
Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption
by: Huang, Buzhen, et al.
Published: (2024) -
Reconstructing Close Human Interaction with Appearance and Proxemics Reasoning
by: Huang, Buzhen, et al.
Published: (2025) -
Contrastive Multi-Modal Hypergraph Reasoning for 3D Crowd Mesh Recovery
by: Sun, Minghao, et al.
Published: (2026) -
E-React: Towards Emotionally Controlled Synthesis of Human Reactions
by: Zhu, Chen, et al.
Published: (2025)