End-to-End Multi-Person Pose Estimation with Pose-Aware Video Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Yonghui, Cai, Jiahang, Wang, Xun, Yang, Wenwu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
by: Fang, Hongwei, et al.
Published: (2026)
by: Fang, Hongwei, et al.
Published: (2026)
An End-to-End Framework for Video Multi-Person Pose Estimation
by: Wei, Zhihong
Published: (2025)
by: Wei, Zhihong
Published: (2025)
LPSNet: End-to-End Human Pose and Shape Estimation with Lensless Imaging
by: Ge, Haoyang, et al.
Published: (2024)
by: Ge, Haoyang, et al.
Published: (2024)
SEMPose: A Single End-to-end Network for Multi-object Pose Estimation
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
KGpose: Keypoint-Graph Driven End-to-End Multi-Object 6D Pose Estimation via Point-Wise Pose Voting
by: Jeong, Andrew
Published: (2024)
by: Jeong, Andrew
Published: (2024)
InterMesh: Explicit Interaction-Aware End-to-End Multi-Person Human Mesh Recovery
by: Zheng, Kaili, et al.
Published: (2026)
by: Zheng, Kaili, et al.
Published: (2026)
End-to-End Probabilistic Geometry-Guided Regression for 6DoF Object Pose Estimation
by: Pöllabauer, Thomas, et al.
Published: (2024)
by: Pöllabauer, Thomas, et al.
Published: (2024)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023)
by: Lin, Xiao, et al.
Published: (2023)
VividAnimator: An End-to-End Audio and Pose-driven Half-Body Human Animation Framework
by: Huang, Donglin, et al.
Published: (2025)
by: Huang, Donglin, et al.
Published: (2025)
SDT-6D: Fully Sparse Depth-Transformer for Staged End-to-End 6D Pose Estimation in Industrial Multi-View Bin Picking
by: Leuze, Nico, et al.
Published: (2025)
by: Leuze, Nico, et al.
Published: (2025)
Leveraging Image Matching Toward End-to-End Relative Camera Pose Regression
by: Khatib, Fadi, et al.
Published: (2022)
by: Khatib, Fadi, et al.
Published: (2022)
PoseCrafter: Extreme Pose Estimation with Hybrid Video Synthesis
by: Mao, Qing, et al.
Published: (2025)
by: Mao, Qing, et al.
Published: (2025)
Multi-Person Pose Estimation Evaluation Using Optimal Transportation and Improved Pose Matching
by: Moriki, Takato, et al.
Published: (2026)
by: Moriki, Takato, et al.
Published: (2026)
Pose-Based Sign Language Spotting via an End-to-End Encoder Architecture
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
Egocentric Visibility-Aware Human Pose Estimation
by: Dai, Peng, et al.
Published: (2026)
by: Dai, Peng, et al.
Published: (2026)
Waterfall Transformer for Multi-person Pose Estimation
by: Ranjan, Navin, et al.
Published: (2024)
by: Ranjan, Navin, et al.
Published: (2024)
PoseTraj: Pose-Aware Trajectory Control in Video Diffusion
by: Ji, Longbin, et al.
Published: (2025)
by: Ji, Longbin, et al.
Published: (2025)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
by: Shim, Gyumin, et al.
Published: (2025)
by: Shim, Gyumin, et al.
Published: (2025)
Video-Based Human Pose Regression via Decoupled Space-Time Aggregation
by: He, Jijie, et al.
Published: (2024)
by: He, Jijie, et al.
Published: (2024)
SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose Estimation
by: Srivastav, Vinkle, et al.
Published: (2024)
by: Srivastav, Vinkle, et al.
Published: (2024)
Object Pose Transformer: Unifying Unseen Object Pose Estimation
by: Li, Weihang, et al.
Published: (2026)
by: Li, Weihang, et al.
Published: (2026)
Unsupervised Multi-Person 3D Human Pose Estimation From 2D Poses Alone
by: Hardy, Peter, et al.
Published: (2023)
by: Hardy, Peter, et al.
Published: (2023)
Beyond Imperfections: A Conditional Inpainting Approach for End-to-End Artifact Removal in VTON and Pose Transfer
by: Tabatabaei, Aref, et al.
Published: (2024)
by: Tabatabaei, Aref, et al.
Published: (2024)
SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation
by: Liu, Yujian, et al.
Published: (2025)
by: Liu, Yujian, et al.
Published: (2025)
Multi-view Pose Fusion for Occlusion-Aware 3D Human Pose Estimation
by: Bragagnolo, Laura, et al.
Published: (2024)
by: Bragagnolo, Laura, et al.
Published: (2024)
Can Generative Video Models Help Pose Estimation?
by: Cai, Ruojin, et al.
Published: (2024)
by: Cai, Ruojin, et al.
Published: (2024)
Linear Relative Pose Estimation Founded on Pose-only Imaging Geometry
by: Cai, Qi, et al.
Published: (2024)
by: Cai, Qi, et al.
Published: (2024)
TSM-Pose: Topology-Aware Learning with Semantic Mamba for Category-Level Object Pose Estimation
by: Liu, Jinshuo, et al.
Published: (2026)
by: Liu, Jinshuo, et al.
Published: (2026)
PoseGAM: Robust Unseen Object Pose Estimation via Geometry-Aware Multi-View Reasoning
by: Chen, Jianqi, et al.
Published: (2025)
by: Chen, Jianqi, et al.
Published: (2025)
Spectral Compression Transformer with Line Pose Graph for Monocular 3D Human Pose Estimation
by: Zheng, Zenghao, et al.
Published: (2025)
by: Zheng, Zenghao, et al.
Published: (2025)
Multi-Grained Feature Pruning for Video-Based Human Pose Estimation
by: Wang, Zhigang, et al.
Published: (2025)
by: Wang, Zhigang, et al.
Published: (2025)
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
by: Ma, Yue, et al.
Published: (2023)
by: Ma, Yue, et al.
Published: (2023)
Multi-Person 3D Pose Estimation from Multi-View Uncalibrated Depth Cameras
by: Li, Yu-Jhe, et al.
Published: (2024)
by: Li, Yu-Jhe, et al.
Published: (2024)
MVTOP: Multi-View Transformer-based Object Pose-Estimation
by: Ranftl, Lukas, et al.
Published: (2025)
by: Ranftl, Lukas, et al.
Published: (2025)
DRSI-Net: Dual-Residual Spatial Interaction Network for Multi-Person Pose Estimation
by: Wu, Shang, et al.
Published: (2024)
by: Wu, Shang, et al.
Published: (2024)
PoseCrafter: One-Shot Personalized Video Synthesis Following Flexible Pose Control
by: Zhong, Yong, et al.
Published: (2024)
by: Zhong, Yong, et al.
Published: (2024)
Structure-Aware Correspondence Learning for Relative Pose Estimation
by: Chen, Yihan, et al.
Published: (2025)
by: Chen, Yihan, et al.
Published: (2025)
From Category to Scenery: An End-to-End Framework for Multi-Person Human-Object Interaction Recognition in Videos
by: Qiao, Tanqiu, et al.
Published: (2024)
by: Qiao, Tanqiu, et al.
Published: (2024)
Joint-Motion Mutual Learning for Pose Estimation in Videos
by: Wu, Sifan, et al.
Published: (2024)
by: Wu, Sifan, et al.
Published: (2024)
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
by: Ancey, Pierre, et al.
Published: (2025)
by: Ancey, Pierre, et al.
Published: (2025)
Similar Items
-
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
by: Fang, Hongwei, et al.
Published: (2026) -
An End-to-End Framework for Video Multi-Person Pose Estimation
by: Wei, Zhihong
Published: (2025) -
LPSNet: End-to-End Human Pose and Shape Estimation with Lensless Imaging
by: Ge, Haoyang, et al.
Published: (2024) -
SEMPose: A Single End-to-end Network for Multi-object Pose Estimation
by: Liu, Xin, et al.
Published: (2024) -
KGpose: Keypoint-Graph Driven End-to-End Multi-Object 6D Pose Estimation via Point-Wise Pose Voting
by: Jeong, Andrew
Published: (2024)