High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Runyang, Chang, Hyung Jin, Tse, Tze Ho Elden, Kim, Boeun, Chang, Yi, Gao, Yixing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
by: Tse, Tze Ho Elden, et al.
Published: (2025)
by: Tse, Tze Ho Elden, et al.
Published: (2025)
Bidirectional Regression for Monocular 6DoF Head Pose Estimation and Reference System Alignment
by: Chun, Sungho, et al.
Published: (2024)
by: Chun, Sungho, et al.
Published: (2024)
Improving Human Motion Plausibility with Body Momentum
by: Nguyen, Ha Linh, et al.
Published: (2025)
by: Nguyen, Ha Linh, et al.
Published: (2025)
Leveraging RGB Images for Pre-Training of Event-Based Hand Pose Estimation
by: Liu, Ruicong, et al.
Published: (2025)
by: Liu, Ruicong, et al.
Published: (2025)
MoST: Motion Style Transformer between Diverse Action Contents
by: Kim, Boeun, et al.
Published: (2024)
by: Kim, Boeun, et al.
Published: (2024)
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
by: Jeong, Uyoung, et al.
Published: (2025)
by: Jeong, Uyoung, et al.
Published: (2025)
Humans as Checkerboards: Calibrating Camera Motion Scale for World-Coordinate Human Mesh Recovery
by: Yang, Fengyuan, et al.
Published: (2024)
by: Yang, Fengyuan, et al.
Published: (2024)
DAS3R: Dynamics-Aware Gaussian Splatting for Static Scene Reconstruction
by: Xu, Kai, et al.
Published: (2024)
by: Xu, Kai, et al.
Published: (2024)
A Constrained Optimization Approach for Gaussian Splatting from Coarsely-posed Images and Noisy Lidar Point Clouds
by: Peng, Jizong, et al.
Published: (2025)
by: Peng, Jizong, et al.
Published: (2025)
GeoReF: Geometric Alignment Across Shape Variation for Category-level Object Pose Refinement
by: Zheng, Linfang, et al.
Published: (2024)
by: Zheng, Linfang, et al.
Published: (2024)
PoseMamba: Monocular 3D Human Pose Estimation with Bidirectional Global-Local Spatio-Temporal State Space Model
by: Huang, Yunlong, et al.
Published: (2024)
by: Huang, Yunlong, et al.
Published: (2024)
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
by: Jeong, Uyoung, et al.
Published: (2023)
by: Jeong, Uyoung, et al.
Published: (2023)
Efficient High-Resolution Visual Representation Learning with State Space Model for Human Pose Estimation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
Joint-Motion Mutual Learning for Pose Estimation in Videos
by: Wu, Sifan, et al.
Published: (2024)
by: Wu, Sifan, et al.
Published: (2024)
Visual Intention Grounding for Egocentric Assistants
by: Sun, Pengzhan, et al.
Published: (2025)
by: Sun, Pengzhan, et al.
Published: (2025)
TIGeR: Text-Instructed Generation and Refinement for Template-Free Hand-Object Interaction
by: Huang, Yiyao, et al.
Published: (2025)
by: Huang, Yiyao, et al.
Published: (2025)
SA-GS: Semantic-Aware Gaussian Splatting for Large Scene Reconstruction with Geometry Constrain
by: Xiong, Butian, et al.
Published: (2024)
by: Xiong, Butian, et al.
Published: (2024)
STAR-Pose: Efficient Low-Resolution Video Human Pose Estimation via Spatial-Temporal Adaptive Super-Resolution
by: Jin, Yucheng, et al.
Published: (2025)
by: Jin, Yucheng, et al.
Published: (2025)
Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models
by: Jin, Hyundong, et al.
Published: (2025)
by: Jin, Hyundong, et al.
Published: (2025)
LGM-Pose: A Lightweight Global Modeling Network for Real-time Human Pose Estimation
by: Guo, Biao, et al.
Published: (2025)
by: Guo, Biao, et al.
Published: (2025)
MAEPose: Self-Supervised Spatiotemporal Learning for Human Pose Estimation on mmWave Video
by: Wei, Xijia, et al.
Published: (2026)
by: Wei, Xijia, et al.
Published: (2026)
Kinematics Modeling Network for Video-based Human Pose Estimation
by: Dang, Yonghao, et al.
Published: (2022)
by: Dang, Yonghao, et al.
Published: (2022)
MV-SSM: Multi-View State Space Modeling for 3D Human Pose Estimation
by: Chharia, Aviral, et al.
Published: (2025)
by: Chharia, Aviral, et al.
Published: (2025)
Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball Rotation
by: Choi, YoungChan, et al.
Published: (2025)
by: Choi, YoungChan, et al.
Published: (2025)
Foundation Model-Driven Framework for Human-Object Interaction Prediction with Segmentation Mask Integration
by: Park, Juhan, et al.
Published: (2025)
by: Park, Juhan, et al.
Published: (2025)
Mental Workload Estimation with Electroencephalogram Signals by Combining Multi-Space Deep Models
by: Nguyen, Hong-Hai, et al.
Published: (2023)
by: Nguyen, Hong-Hai, et al.
Published: (2023)
RadMamba: Efficient Human Activity Recognition through Radar-based Micro-Doppler-Oriented Mamba State-Space Model
by: Wu, Yizhuo, et al.
Published: (2025)
by: Wu, Yizhuo, et al.
Published: (2025)
GraspALL: Adaptive Structural Compensation from Illumination Variation for Robotic Garment Grasping in Any Low-Light Conditions
by: Zhong, Haifeng, et al.
Published: (2026)
by: Zhong, Haifeng, et al.
Published: (2026)
JointLoc: A Real-time Visual Localization Framework for Planetary UAVs Based on Joint Relative and Absolute Pose Estimation
by: Luo, Xubo, et al.
Published: (2024)
by: Luo, Xubo, et al.
Published: (2024)
Optimizing Local-Global Dependencies for Accurate 3D Human Pose Estimation
by: Xu, Guangsheng, et al.
Published: (2024)
by: Xu, Guangsheng, et al.
Published: (2024)
Transformers with Joint Tokens and Local-Global Attention for Efficient Human Pose Estimation
by: Kinfu, Kaleab A., et al.
Published: (2025)
by: Kinfu, Kaleab A., et al.
Published: (2025)
Exploiting Spatiotemporal Properties for Efficient Event-Driven Human Pose Estimation
by: Zhou, Haoxian, et al.
Published: (2025)
by: Zhou, Haoxian, et al.
Published: (2025)
Local-Global Temporal Difference Learning for Satellite Video Super-Resolution
by: Xiao, Yi, et al.
Published: (2023)
by: Xiao, Yi, et al.
Published: (2023)
Dite-HRNet: Dynamic Lightweight High-Resolution Network for Human Pose Estimation
by: Li, Qun, et al.
Published: (2022)
by: Li, Qun, et al.
Published: (2022)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
by: Hyung, Junha, et al.
Published: (2024)
by: Hyung, Junha, et al.
Published: (2024)
Focus on Low-Resolution Information: Multi-Granular Information-Lossless Model for Low-Resolution Human Pose Estimation
by: Gu, Zejun, et al.
Published: (2024)
by: Gu, Zejun, et al.
Published: (2024)
Human Modelling and Pose Estimation Overview
by: Knap, Pawel
Published: (2024)
by: Knap, Pawel
Published: (2024)
Localization Through Particle Filter Powered Neural Network Estimated Monocular Camera Poses
by: Shen, Yi, et al.
Published: (2024)
by: Shen, Yi, et al.
Published: (2024)
A novel thiourea‐based fluorescent turn‐on sensor for rapidly detecting hypochlorite through a desulfurization reaction
by: Boeun Choi, et al.
Published: (2024)
by: Boeun Choi, et al.
Published: (2024)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
Similar Items
-
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
by: Tse, Tze Ho Elden, et al.
Published: (2025) -
Bidirectional Regression for Monocular 6DoF Head Pose Estimation and Reference System Alignment
by: Chun, Sungho, et al.
Published: (2024) -
Improving Human Motion Plausibility with Body Momentum
by: Nguyen, Ha Linh, et al.
Published: (2025) -
Leveraging RGB Images for Pre-Training of Event-Based Hand Pose Estimation
by: Liu, Ruicong, et al.
Published: (2025) -
MoST: Motion Style Transformer between Diverse Action Contents
by: Kim, Boeun, et al.
Published: (2024)