DeforHMR: Vision Transformer with Deformable Cross-Attention for 3D Human Mesh Recovery
Fuente:
arXiv
Saved in:
| Main Authors: | Heo, Jaewoo, Hu, George, Wang, Zeyu, Yeung-Levy, Serena |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Human Mesh Recovery with Transformers
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
Motion Diffusion-Guided 3D Global HMR from a Dynamic Camera
by: Heo, Jaewoo, et al.
Published: (2024)
by: Heo, Jaewoo, et al.
Published: (2024)
PromptHMR: Promptable Human Mesh Recovery
by: Wang, Yufu, et al.
Published: (2025)
by: Wang, Yufu, et al.
Published: (2025)
Diffusion-HPC: Synthetic Data Generation for Human Mesh Recovery in Challenging Domains
by: Weng, Zhenzhen, et al.
Published: (2023)
by: Weng, Zhenzhen, et al.
Published: (2023)
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)
by: Kim, Donghwan, et al.
Published: (2025)
LiDAR-HMR: 3D Human Mesh Recovery from LiDAR
by: Fan, Bohao, et al.
Published: (2023)
by: Fan, Bohao, et al.
Published: (2023)
GenHMR: Generative Human Mesh Recovery
by: Saleem, Muhammad Usama, et al.
Published: (2024)
by: Saleem, Muhammad Usama, et al.
Published: (2024)
Anny-Fit: All-Age Human Mesh Recovery
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
by: Bravo-Sánchez, Laura, et al.
Published: (2026)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
by: Zhao, Yiwen, et al.
Published: (2026)
by: Zhao, Yiwen, et al.
Published: (2026)
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
ClothHMR: 3D Mesh Recovery of Humans in Diverse Clothing from Single Image
by: Gao, Yunqi, et al.
Published: (2025)
by: Gao, Yunqi, et al.
Published: (2025)
FactorizedHMR: A Hybrid Framework for Video Human Mesh Recovery
by: Kwon, Patrick, et al.
Published: (2026)
by: Kwon, Patrick, et al.
Published: (2026)
HumanSplatHMR: Closing the Loop Between Human Mesh Recovery and Gaussian Splatting Avatar
by: Zong, Yeheng, et al.
Published: (2026)
by: Zong, Yeheng, et al.
Published: (2026)
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
by: Yao, Wei, et al.
Published: (2023)
by: Yao, Wei, et al.
Published: (2023)
ResiHMR: Residual-Limb Aware Single-Image 3D Human Mesh Recovery for Individuals with Limb Loss
by: Ying, Jiaying, et al.
Published: (2026)
by: Ying, Jiaying, et al.
Published: (2026)
Ask, Pose, Unite: Scaling Data Acquisition for Close Interactions with Vision Language Models
by: Bravo-Sánchez, Laura, et al.
Published: (2024)
by: Bravo-Sánchez, Laura, et al.
Published: (2024)
Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot
by: Baradel, Fabien, et al.
Published: (2024)
by: Baradel, Fabien, et al.
Published: (2024)
FastHMR: Accelerating Human Mesh Recovery via Token and Layer Merging with Diffusion Decoding
by: Mehraban, Soroush, et al.
Published: (2025)
by: Mehraban, Soroush, et al.
Published: (2025)
DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
Continuous Perception Matters: Diagnosing Temporal Integration Failures in Multimodal Models
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
by: Jeong, David C., et al.
Published: (2025)
by: Jeong, David C., et al.
Published: (2025)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
PressTrack-HMR: Pressure-Based Top-Down Multi-Person Global Human Mesh Recovery
by: Yuan, Jiayue, et al.
Published: (2025)
by: Yuan, Jiayue, et al.
Published: (2025)
Zero-shot Action Localization via the Confidence of Large Vision-Language Models
by: Aklilu, Josiah, et al.
Published: (2024)
by: Aklilu, Josiah, et al.
Published: (2024)
Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration
by: Endo, Mark, et al.
Published: (2024)
by: Endo, Mark, et al.
Published: (2024)
PostoMETRO: Pose Token Enhanced Mesh Transformer for Robust 3D Human Mesh Recovery
by: Yang, Wendi, et al.
Published: (2024)
by: Yang, Wendi, et al.
Published: (2024)
Systematic Evaluation of Large Vision-Language Models for Surgical Artificial Intelligence
by: Rau, Anita, et al.
Published: (2025)
by: Rau, Anita, et al.
Published: (2025)
Adapting Human Mesh Recovery with Vision-Language Feedback
by: Xu, Chongyang, et al.
Published: (2025)
by: Xu, Chongyang, et al.
Published: (2025)
SAT-HMR: Real-Time Multi-Person 3D Mesh Estimation via Scale-Adaptive Tokens
by: Su, Chi, et al.
Published: (2024)
by: Su, Chi, et al.
Published: (2024)
Distribution and Depth-Aware Transformers for 3D Human Mesh Recovery
by: Bright, Jerrin, et al.
Published: (2024)
by: Bright, Jerrin, et al.
Published: (2024)
Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models
by: Endo, Mark, et al.
Published: (2025)
by: Endo, Mark, et al.
Published: (2025)
Just Shift It: Test-Time Prototype Shifting for Zero-Shot Generalization with Vision-Language Models
by: Sui, Elaine, et al.
Published: (2024)
by: Sui, Elaine, et al.
Published: (2024)
Viewpoint Textual Inversion: Discovering Scene Representations and 3D View Control in 2D Diffusion Models
by: Burgess, James, et al.
Published: (2023)
by: Burgess, James, et al.
Published: (2023)
Template-Free Single-View 3D Human Digitalization with Diffusion-Guided LRM
by: Weng, Zhenzhen, et al.
Published: (2024)
by: Weng, Zhenzhen, et al.
Published: (2024)
HMR3D: Hierarchical Multimodal Representation for 3D Scene Understanding with Large Vision-Language Model
by: Li, Chen, et al.
Published: (2025)
by: Li, Chen, et al.
Published: (2025)
M3DHMR: Monocular 3D Hand Mesh Recovery
by: Lin, Yihong, et al.
Published: (2025)
by: Lin, Yihong, et al.
Published: (2025)
Foundation Models Secretly Understand Neural Network Weights: Enhancing Hypernetwork Architectures with Foundation Models
by: Gu, Jeffrey, et al.
Published: (2025)
by: Gu, Jeffrey, et al.
Published: (2025)
Towards Automated Initial Probe Placement in Transthoracic Teleultrasound Using Human Mesh and Skeleton Recovery
by: Lee, Yu Chung, et al.
Published: (2026)
by: Lee, Yu Chung, et al.
Published: (2026)
PhysHMR: Learning Humanoid Control Policies from Vision for Physically Plausible Human Motion Reconstruction
by: Feng, Qiao, et al.
Published: (2025)
by: Feng, Qiao, et al.
Published: (2025)
SAM 3D Body: Robust Full-Body Human Mesh Recovery
by: Yang, Xitong, et al.
Published: (2026)
by: Yang, Xitong, et al.
Published: (2026)
Similar Items
-
Multi-Human Mesh Recovery with Transformers
by: Wang, Zeyu, et al.
Published: (2024) -
Motion Diffusion-Guided 3D Global HMR from a Dynamic Camera
by: Heo, Jaewoo, et al.
Published: (2024) -
PromptHMR: Promptable Human Mesh Recovery
by: Wang, Yufu, et al.
Published: (2025) -
Diffusion-HPC: Synthetic Data Generation for Human Mesh Recovery in Challenging Domains
by: Weng, Zhenzhen, et al.
Published: (2023) -
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)