STAF: 3D Human Mesh Recovery from Video with Spatio-Temporal Alignment Fusion
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Wei, Zhang, Hongwen, Sun, Yunlian, Tang, Jinhui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
by: Yao, Wei, et al.
Published: (2023)
by: Yao, Wei, et al.
Published: (2023)
PhysiInter: Integrating Physical Mapping for High-Fidelity Human Interaction Generation
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions
by: Shi, Xu, et al.
Published: (2023)
by: Shi, Xu, et al.
Published: (2023)
SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
MetricHMSR:Metric Human Mesh and Scene Recovery from Monocular Images
by: Song, Chentao, et al.
Published: (2025)
by: Song, Chentao, et al.
Published: (2025)
Recovering 3D Human Mesh from Monocular Images: A Survey
by: Tian, Yating, et al.
Published: (2022)
by: Tian, Yating, et al.
Published: (2022)
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026)
by: Cen, Jiaxin, et al.
Published: (2026)
VideoFusion: A Spatio-Temporal Collaborative Network for Multi-modal Video Fusion
by: Tang, Linfeng, et al.
Published: (2025)
by: Tang, Linfeng, et al.
Published: (2025)
Video-Language Alignment via Spatio-Temporal Graph Transformer
by: Zhang, Shi-Xue, et al.
Published: (2024)
by: Zhang, Shi-Xue, et al.
Published: (2024)
STGFormer: Spatio-Temporal GraphFormer for 3D Human Pose Estimation in Video
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
HHMR: Holistic Hand Mesh Recovery by Enhancing the Multimodal Controllability of Graph Diffusion Models
by: Li, Mengcheng, et al.
Published: (2024)
by: Li, Mengcheng, et al.
Published: (2024)
Rethinking the Spatio-Temporal Alignment of End-to-End 3D Perception
by: Li, Xiaoyu, et al.
Published: (2025)
by: Li, Xiaoyu, et al.
Published: (2025)
Enhancing Video-Language Representations with Structural Spatio-Temporal Alignment
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
ARTS: Semi-Analytical Regressor using Disentangled Skeletal Representations for Human Mesh Recovery from Videos
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
Fast 4D Mesh Generation by Spatio-Temporal Attention Chains
by: Samuel, Dvir, et al.
Published: (2026)
by: Samuel, Dvir, et al.
Published: (2026)
Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding
by: Tu, Xuezhen, et al.
Published: (2026)
by: Tu, Xuezhen, et al.
Published: (2026)
Fish2Mesh Transformer: 3D Human Mesh Recovery from Egocentric Vision
by: Jeong, David C., et al.
Published: (2025)
by: Jeong, David C., et al.
Published: (2025)
SAM-Body4D: Training-Free 4D Human Body Mesh Recovery from Videos
by: Gao, Mingqi, et al.
Published: (2025)
by: Gao, Mingqi, et al.
Published: (2025)
VLM-Guided Group Preference Alignment for Diffusion-based Human Mesh Recovery
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
Patient4D: Temporally Consistent Patient Body Mesh Recovery from Monocular Operating Room Video
by: Tu, Mingxiao, et al.
Published: (2026)
by: Tu, Mingxiao, et al.
Published: (2026)
PromptHMR: Promptable Human Mesh Recovery
by: Wang, Yufu, et al.
Published: (2025)
by: Wang, Yufu, et al.
Published: (2025)
PostoMETRO: Pose Token Enhanced Mesh Transformer for Robust 3D Human Mesh Recovery
by: Yang, Wendi, et al.
Published: (2024)
by: Yang, Wendi, et al.
Published: (2024)
Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Towards Geometry-Aware and Motion-Guided Video Human Mesh Recovery
by: Chen, Hongjun, et al.
Published: (2026)
by: Chen, Hongjun, et al.
Published: (2026)
TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
by: Dwivedi, Sai Kumar, et al.
Published: (2024)
STAF: Sinusoidal Trainable Activation Functions for Implicit Neural Representation
by: Morsali, Alireza, et al.
Published: (2025)
by: Morsali, Alireza, et al.
Published: (2025)
LiDAR-HMR: 3D Human Mesh Recovery from LiDAR
by: Fan, Bohao, et al.
Published: (2023)
by: Fan, Bohao, et al.
Published: (2023)
OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer
by: Lin, Dixuan, et al.
Published: (2024)
by: Lin, Dixuan, et al.
Published: (2024)
Validation of Human Pose Estimation and Human Mesh Recovery for Extracting Clinically Relevant Motion Data from Videos
by: Armstrong, Kai, et al.
Published: (2025)
by: Armstrong, Kai, et al.
Published: (2025)
DanceHMR: Hand-Aware Whole-Body Human Mesh Recovery from Monocular Videos
by: Shen, Wenhao, et al.
Published: (2026)
by: Shen, Wenhao, et al.
Published: (2026)
Dual-Branch Graph Transformer Network for 3D Human Mesh Reconstruction from Video
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
SAM 3D Body: Robust Full-Body Human Mesh Recovery
by: Yang, Xitong, et al.
Published: (2026)
by: Yang, Xitong, et al.
Published: (2026)
STF: Spatio-Temporal Fusion Module for Improving Video Object Detection
by: Anwar, Noreen, et al.
Published: (2024)
by: Anwar, Noreen, et al.
Published: (2024)
Towards Long-Form Spatio-Temporal Video Grounding
by: Gu, Xin, et al.
Published: (2026)
by: Gu, Xin, et al.
Published: (2026)
Multi-Human Mesh Recovery with Transformers
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
OnlineHMR: Video-based Online World-Grounded Human Mesh Recovery
by: Zhao, Yiwen, et al.
Published: (2026)
by: Zhao, Yiwen, et al.
Published: (2026)
DPMesh: Exploiting Diffusion Prior for Occluded Human Mesh Recovery
by: Zhu, Yixuan, et al.
Published: (2024)
by: Zhu, Yixuan, et al.
Published: (2024)
MF2Summ: Multimodal Fusion for Video Summarization with Temporal Alignment
by: wang, Shuo, et al.
Published: (2025)
by: wang, Shuo, et al.
Published: (2025)
Similar Items
-
W-HMR: Monocular Human Mesh Recovery in World Space with Weak-Supervised Calibration
by: Yao, Wei, et al.
Published: (2023) -
PhysiInter: Integrating Physical Mapping for High-Fidelity Human Interaction Generation
by: Yao, Wei, et al.
Published: (2025) -
HOSIG: Full-Body Human-Object-Scene Interaction Generation with Hierarchical Scene Perception
by: Yao, Wei, et al.
Published: (2025) -
FG-MDM: Towards Zero-Shot Human Motion Generation via ChatGPT-Refined Descriptions
by: Shi, Xu, et al.
Published: (2023) -
SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy
by: Yao, Wei, et al.
Published: (2026)