PyMAF-X: Towards Well-aligned Full-body Model Regression from Monocular Images
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Hongwen, Tian, Yating, Zhang, Yuxiang, Li, Mengcheng, An, Liang, Sun, Zhenan, Liu, Yebin |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
by: Hu, Junxing, et al.
Published: (2023)
by: Hu, Junxing, et al.
Published: (2023)
Recovering 3D Human Mesh from Monocular Images: A Survey
by: Tian, Yating, et al.
Published: (2022)
by: Tian, Yating, et al.
Published: (2022)
HHMR: Holistic Hand Mesh Recovery by Enhancing the Multimodal Controllability of Graph Diffusion Models
by: Li, Mengcheng, et al.
Published: (2024)
by: Li, Mengcheng, et al.
Published: (2024)
ManiDext: Hand-Object Manipulation Synthesis via Continuous Correspondence Embeddings and Residual-Guided Diffusion
by: Zhang, Jiajun, et al.
Published: (2024)
by: Zhang, Jiajun, et al.
Published: (2024)
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
by: Lin, Dixuan, et al.
Published: (2024)
by: Lin, Dixuan, et al.
Published: (2024)
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
by: Yao, Wei, et al.
Published: (2023)
by: Yao, Wei, et al.
Published: (2023)
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
SpeechAct: Towards Generating Whole-body Motion from Speech
by: Zhang, Jinsong, et al.
Published: (2023)
by: Zhang, Jinsong, et al.
Published: (2023)
MoReMouse: Monocular Reconstruction of Laboratory Mouse
by: Zhong, Yuan, et al.
Published: (2025)
by: Zhong, Yuan, et al.
Published: (2025)
DevilSight: Augmenting Monocular Human Avatar Reconstruction through a Virtual Perspective
by: Chen, Yushuo, et al.
Published: (2025)
by: Chen, Yushuo, et al.
Published: (2025)
SemanticSplat: Feed-Forward 3D Scene Understanding with Language-Aware Gaussian Fields
by: Li, Qijing, et al.
Published: (2025)
by: Li, Qijing, et al.
Published: (2025)
CloSET: Modeling Clothed Humans on Continuous Surface with Explicit Template Decomposition
by: Zhang, Hongwen, et al.
Published: (2023)
by: Zhang, Hongwen, et al.
Published: (2023)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
by: Zhang, Jiajun, et al.
Published: (2023)
by: Zhang, Jiajun, et al.
Published: (2023)
MetricHMSR:Metric Human Mesh and Scene Recovery from Monocular Images
by: Song, Chentao, et al.
Published: (2025)
by: Song, Chentao, et al.
Published: (2025)
Mix3R: Mixing Feed-forward Reconstruction and Generative 3D Priors for Joint Multi-view Aligned 3D Reconstruction and Pose Estimation
by: Lin, Siyou, et al.
Published: (2026)
by: Lin, Siyou, et al.
Published: (2026)
SyncMV4D: Synchronized Multi-view Joint Diffusion of Appearance and Motion for Hand-Object Interaction Synthesis
by: Dang, Lingwei, et al.
Published: (2025)
by: Dang, Lingwei, et al.
Published: (2025)
GaussianAvatar: Towards Realistic Human Avatar Modeling from a Single Video via Animatable 3D Gaussians
by: Hu, Liangxiao, et al.
Published: (2023)
by: Hu, Liangxiao, et al.
Published: (2023)
GPHM: Gaussian Parametric Head Model for Monocular Head Avatar Reconstruction
by: Xu, Yuelang, et al.
Published: (2024)
by: Xu, Yuelang, et al.
Published: (2024)
Lodge++: High-quality and Long Dance Generation with Vivid Choreography Patterns
by: Li, Ronghui, et al.
Published: (2024)
by: Li, Ronghui, et al.
Published: (2024)
Monocular Mesh Recovery and Body Measurement of Female Saanen Goats
by: Jin, Bo, et al.
Published: (2026)
by: Jin, Bo, et al.
Published: (2026)
4DEquine: Disentangling Motion and Appearance for 4D Equine Reconstruction from Monocular Video
by: Lyu, Jin, et al.
Published: (2026)
by: Lyu, Jin, et al.
Published: (2026)
SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios
by: Dang, Lingwei, et al.
Published: (2025)
by: Dang, Lingwei, et al.
Published: (2025)
LayGA: Layered Gaussian Avatars for Animatable Clothing Transfer
by: Lin, Siyou, et al.
Published: (2024)
by: Lin, Siyou, et al.
Published: (2024)
Graph and Skipped Transformer: Exploiting Spatial and Temporal Modeling Capacities for Efficient 3D Human Pose Estimation
by: Cui, Mengmeng, et al.
Published: (2024)
by: Cui, Mengmeng, et al.
Published: (2024)
SharpTimeGS: Sharp and Stable Dynamic Gaussian Splatting via Lifespan Modulation
by: Liao, Zhanfeng, et al.
Published: (2026)
by: Liao, Zhanfeng, et al.
Published: (2026)
ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping
by: Pang, Youxin, et al.
Published: (2024)
by: Pang, Youxin, et al.
Published: (2024)
FOF-X: Towards Real-time Detailed Human Reconstruction from a Single Image
by: Feng, Qiao, et al.
Published: (2024)
by: Feng, Qiao, et al.
Published: (2024)
Towards aligned body representations in vision models
by: Gizdov, Andrey, et al.
Published: (2025)
by: Gizdov, Andrey, et al.
Published: (2025)
FATE: Full-head Gaussian Avatar with Textural Editing from Monocular Video
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Tessellation GS: Neural Mesh Gaussians for Robust Monocular Reconstruction of Dynamic Objects
by: Tao, Shuohan, et al.
Published: (2025)
by: Tao, Shuohan, et al.
Published: (2025)
GAF: Gaussian Action Field as a 4D Representation for Dynamic World Modeling in Robotic Manipulation
by: Chai, Ying, et al.
Published: (2025)
by: Chai, Ying, et al.
Published: (2025)
KBody: Towards general, robust, and aligned monocular whole-body estimation
by: Zioulis, Nikolaos, et al.
Published: (2023)
by: Zioulis, Nikolaos, et al.
Published: (2023)
FullAnno: A Data Engine for Enhancing Image Comprehension of MLLMs
by: Hao, Jing, et al.
Published: (2024)
by: Hao, Jing, et al.
Published: (2024)
CFCPalsy: Facial Image Synthesis with Cross-Fusion Cycle Diffusion Model for Facial Paralysis Individuals
by: Gao, Weixiang, et al.
Published: (2024)
by: Gao, Weixiang, et al.
Published: (2024)
How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing
by: Zhang, Huanyu, et al.
Published: (2026)
by: Zhang, Huanyu, et al.
Published: (2026)
FMGS-Avatar: Mesh-Guided 2D Gaussian Splatting with Foundation Model Priors for 3D Monocular Avatar Reconstruction
by: Fan, Jinlong, et al.
Published: (2025)
by: Fan, Jinlong, et al.
Published: (2025)
FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions
by: Shi, Xu, et al.
Published: (2023)
by: Shi, Xu, et al.
Published: (2023)
SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
Human as Points: Explicit Point-based 3D Human Reconstruction from Single-view RGB Images
by: Tang, Yingzhi, et al.
Published: (2023)
by: Tang, Yingzhi, et al.
Published: (2023)
SpectralX: Parameter-efficient Domain Generalization for Spectral Remote Sensing Foundation Models
by: Zhang, Yuxiang, et al.
Published: (2025)
by: Zhang, Yuxiang, et al.
Published: (2025)
Similar Items
-
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
by: Hu, Junxing, et al.
Published: (2023) -
Recovering 3D Human Mesh from Monocular Images: A Survey
by: Tian, Yating, et al.
Published: (2022) -
HHMR: Holistic Hand Mesh Recovery by Enhancing the Multimodal Controllability of Graph Diffusion Models
by: Li, Mengcheng, et al.
Published: (2024) -
ManiDext: Hand-Object Manipulation Synthesis via Continuous Correspondence Embeddings and Residual-Guided Diffusion
by: Zhang, Jiajun, et al.
Published: (2024) -
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
by: Lin, Dixuan, et al.
Published: (2024)