Transformers with Joint Tokens and Local-Global Attention for Efficient Human Pose Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Kinfu, Kaleab A., Vidal, René |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
di: Li, Wenhao, et al.
Pubblicazione: (2023)
di: Li, Wenhao, et al.
Pubblicazione: (2023)
EPOCH: Jointly Estimating the 3D Pose of Cameras and Humans
di: Garau, Nicola, et al.
Pubblicazione: (2024)
di: Garau, Nicola, et al.
Pubblicazione: (2024)
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation
di: Zheng, Hongwei, et al.
Pubblicazione: (2025)
di: Zheng, Hongwei, et al.
Pubblicazione: (2025)
Computerized Assessment of Motor Imitation for Distinguishing Autism in Video (CAMI-2DNet)
di: Kinfu, Kaleab A., et al.
Pubblicazione: (2025)
di: Kinfu, Kaleab A., et al.
Pubblicazione: (2025)
Human Pose Estimation from Ambiguous Pressure Recordings with Spatio-temporal Masked Transformers
di: Davoodnia, Vandad, et al.
Pubblicazione: (2023)
di: Davoodnia, Vandad, et al.
Pubblicazione: (2023)
Enhancing 3D Human Pose Estimation Amidst Severe Occlusion with Dual Transformer Fusion
di: Ghafoor, Mehwish, et al.
Pubblicazione: (2024)
di: Ghafoor, Mehwish, et al.
Pubblicazione: (2024)
H$_{2}$OT: Hierarchical Hourglass Tokenizer for Efficient Video Pose Transformers
di: Li, Wenhao, et al.
Pubblicazione: (2025)
di: Li, Wenhao, et al.
Pubblicazione: (2025)
Occlusion Resilient 3D Human Pose Estimation
di: Roy, Soumava Kumar, et al.
Pubblicazione: (2024)
di: Roy, Soumava Kumar, et al.
Pubblicazione: (2024)
MovePose: A High-performance Human Pose Estimation Algorithm on Mobile and Edge Devices
di: Yu, Dongyang, et al.
Pubblicazione: (2023)
di: Yu, Dongyang, et al.
Pubblicazione: (2023)
The Influence of Faulty Labels in Data Sets on Human Pose Estimation
di: Schwarz, Arnold, et al.
Pubblicazione: (2024)
di: Schwarz, Arnold, et al.
Pubblicazione: (2024)
Uncertainty-Aware Token Importance Estimation in Spiking Transformers
di: Liu, Wenxuan, et al.
Pubblicazione: (2026)
di: Liu, Wenxuan, et al.
Pubblicazione: (2026)
Efficient Visual Transformer by Learnable Token Merging
di: Wang, Yancheng, et al.
Pubblicazione: (2024)
di: Wang, Yancheng, et al.
Pubblicazione: (2024)
Disentangling Safe and Unsafe Corruptions via Anisotropy and Locality
di: Muthukumar, Ramchandran, et al.
Pubblicazione: (2025)
di: Muthukumar, Ramchandran, et al.
Pubblicazione: (2025)
Recurrent Attention-based Token Selection for Efficient Streaming Video-LLMs
di: Dorovatas, Vaggelis, et al.
Pubblicazione: (2025)
di: Dorovatas, Vaggelis, et al.
Pubblicazione: (2025)
GTPT: Group-based Token Pruning Transformer for Efficient Human Pose Estimation
di: Wang, Haonan, et al.
Pubblicazione: (2024)
di: Wang, Haonan, et al.
Pubblicazione: (2024)
Joint Coordinate Regression and Association For Multi-Person Pose Estimation, A Pure Neural Network Approach
di: Yu, Dongyang, et al.
Pubblicazione: (2023)
di: Yu, Dongyang, et al.
Pubblicazione: (2023)
A Novel Convolution and Attention Mechanism-based Model for 6D Object Pose Estimation
di: Du, Alexander, et al.
Pubblicazione: (2024)
di: Du, Alexander, et al.
Pubblicazione: (2024)
In-Bed Human Pose Estimation from Unseen and Privacy-Preserving Image Domains
di: Cao, Ting, et al.
Pubblicazione: (2021)
di: Cao, Ting, et al.
Pubblicazione: (2021)
EfficientPose 6D: Scalable and Efficient 6D Object Pose Estimation
di: Fang, Zixuan, et al.
Pubblicazione: (2025)
di: Fang, Zixuan, et al.
Pubblicazione: (2025)
WiCompass: Oracle-driven Data Scaling for mmWave Human Pose Estimation
di: Liang, Bo, et al.
Pubblicazione: (2026)
di: Liang, Bo, et al.
Pubblicazione: (2026)
Neural Human Pose Prior
di: Heker, Michal, et al.
Pubblicazione: (2025)
di: Heker, Michal, et al.
Pubblicazione: (2025)
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
di: Becker, Philipp, et al.
Pubblicazione: (2025)
di: Becker, Philipp, et al.
Pubblicazione: (2025)
ManiPose: Manifold-Constrained Multi-Hypothesis 3D Human Pose Estimation
di: Rommel, Cédric, et al.
Pubblicazione: (2023)
di: Rommel, Cédric, et al.
Pubblicazione: (2023)
ToFe: Lagged Token Freezing and Reusing for Efficient Vision Transformer Inference
di: Zhang, Haoyue, et al.
Pubblicazione: (2025)
di: Zhang, Haoyue, et al.
Pubblicazione: (2025)
Efficient Onboard Spacecraft Pose Estimation with Event Cameras and Neuromorphic Hardware
di: Rathinam, Arunkumar, et al.
Pubblicazione: (2026)
di: Rathinam, Arunkumar, et al.
Pubblicazione: (2026)
Clebsch-Gordan Transformer: Fast and Global Equivariant Attention
di: Howell, Owen Lewis, et al.
Pubblicazione: (2025)
di: Howell, Owen Lewis, et al.
Pubblicazione: (2025)
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
di: Gabetni, Firas, et al.
Pubblicazione: (2025)
di: Gabetni, Firas, et al.
Pubblicazione: (2025)
AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer
di: Shan, Jiquan, et al.
Pubblicazione: (2025)
di: Shan, Jiquan, et al.
Pubblicazione: (2025)
STAR: Stage-Wise Attention-Guided Token Reduction for Efficient Large Vision-Language Models Inference
di: Guo, Yichen, et al.
Pubblicazione: (2025)
di: Guo, Yichen, et al.
Pubblicazione: (2025)
Lifelong Domain Adaptive 3D Human Pose Estimation
di: Peng, Qucheng, et al.
Pubblicazione: (2025)
di: Peng, Qucheng, et al.
Pubblicazione: (2025)
Adaptive Deep Learning for Efficient Visual Pose Estimation aboard Ultra-low-power Nano-drones
di: Motetti, Beatrice Alessandra, et al.
Pubblicazione: (2024)
di: Motetti, Beatrice Alessandra, et al.
Pubblicazione: (2024)
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
Design and Analysis of Efficient Attention in Transformers for Social Group Activity Recognition
di: Tamura, Masato
Pubblicazione: (2024)
di: Tamura, Masato
Pubblicazione: (2024)
BiGain: Unified Token Compression for Joint Generation and Classification
di: Liu, Jiacheng, et al.
Pubblicazione: (2026)
di: Liu, Jiacheng, et al.
Pubblicazione: (2026)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
Map-Relative Pose Regression for Visual Re-Localization
di: Chen, Shuai, et al.
Pubblicazione: (2024)
di: Chen, Shuai, et al.
Pubblicazione: (2024)
Token Caching for Diffusion Transformer Acceleration
di: Lou, Jinming, et al.
Pubblicazione: (2024)
di: Lou, Jinming, et al.
Pubblicazione: (2024)
Cameras as Rays: Pose Estimation via Ray Diffusion
di: Zhang, Jason Y., et al.
Pubblicazione: (2024)
di: Zhang, Jason Y., et al.
Pubblicazione: (2024)
Mixture of Distributions Matters: Dynamic Sparse Attention for Efficient Video Diffusion Transformers
di: Liu, Yuxi, et al.
Pubblicazione: (2026)
di: Liu, Yuxi, et al.
Pubblicazione: (2026)
Hierarchical Concept Embedding & Pursuit for Interpretable Image Classification
di: Nguyen, Nghia, et al.
Pubblicazione: (2026)
di: Nguyen, Nghia, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
di: Li, Wenhao, et al.
Pubblicazione: (2023) -
EPOCH: Jointly Estimating the 3D Pose of Cameras and Humans
di: Garau, Nicola, et al.
Pubblicazione: (2024) -
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation
di: Zheng, Hongwei, et al.
Pubblicazione: (2025) -
Computerized Assessment of Motor Imitation for Distinguishing Autism in Video (CAMI-2DNet)
di: Kinfu, Kaleab A., et al.
Pubblicazione: (2025) -
Human Pose Estimation from Ambiguous Pressure Recordings with Spatio-temporal Masked Transformers
di: Davoodnia, Vandad, et al.
Pubblicazione: (2023)