Highly Efficient 3D Human Pose Tracking from Events with Spiking Spatiotemporal Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zou, Shihao, Mu, Yuxuan, Ji, Wei, Wang, Zi-An, Zuo, Xinxin, Wang, Sen, Si, Weixin, Cheng, Li |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toward Real-Time Surgical Scene Segmentation via a Spike-Driven Video Transformer with Spike-Informed Pretraining
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
Sketch2PoseNet: Efficient and Generalized Sketch to 3D Human Pose Prediction
von: Wang, Li, et al.
Veröffentlicht: (2025)
von: Wang, Li, et al.
Veröffentlicht: (2025)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
von: Zou, Shihao, et al.
Veröffentlicht: (2025)
Exploiting Spatiotemporal Properties for Efficient Event-Driven Human Pose Estimation
von: Zhou, Haoxian, et al.
Veröffentlicht: (2025)
von: Zhou, Haoxian, et al.
Veröffentlicht: (2025)
TempDiffReg: Temporal Diffusion Model for Non-Rigid 2D-3D Vascular Registration
von: Liu, Zehua, et al.
Veröffentlicht: (2026)
von: Liu, Zehua, et al.
Veröffentlicht: (2026)
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space
von: Yu, Shiyao, et al.
Veröffentlicht: (2025)
von: Yu, Shiyao, et al.
Veröffentlicht: (2025)
Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs
von: Huang, Jincai, et al.
Veröffentlicht: (2026)
von: Huang, Jincai, et al.
Veröffentlicht: (2026)
Exploring Event-based Human Pose Estimation with 3D Event Representations
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
Generative Human Motion Stylization in Latent Space
von: Guo, Chuan, et al.
Veröffentlicht: (2024)
von: Guo, Chuan, et al.
Veröffentlicht: (2024)
DS2TA: Denoising Spiking Transformer with Attenuated Spatiotemporal Attention
von: Xu, Boxun, et al.
Veröffentlicht: (2024)
von: Xu, Boxun, et al.
Veröffentlicht: (2024)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
PICS: Pairwise Image Compositing with Spatial Interactions
von: Zhou, Hang, et al.
Veröffentlicht: (2026)
von: Zhou, Hang, et al.
Veröffentlicht: (2026)
SEDformer: Event-Synchronous Spiking Transformers for Irregular Telemetry Time Series Forecasting
von: Zhou, Ziyu, et al.
Veröffentlicht: (2026)
von: Zhou, Ziyu, et al.
Veröffentlicht: (2026)
SpikeTrack: A Spike-driven Framework for Efficient Visual Tracking
von: Zhang, Qiuyang, et al.
Veröffentlicht: (2026)
von: Zhang, Qiuyang, et al.
Veröffentlicht: (2026)
HiPART: Hierarchical Pose AutoRegressive Transformer for Occluded 3D Human Pose Estimation
von: Zheng, Hongwei, et al.
Veröffentlicht: (2025)
von: Zheng, Hongwei, et al.
Veröffentlicht: (2025)
GSD: View-Guided Gaussian Splatting Diffusion for 3D Reconstruction
von: Mu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Mu, Yuxuan, et al.
Veröffentlicht: (2024)
Semantics-Aware Human Motion Generation from Audio Instructions
von: Wang, Zi-An, et al.
Veröffentlicht: (2025)
von: Wang, Zi-An, et al.
Veröffentlicht: (2025)
Event6D: Event-based Novel Object 6D Pose Tracking
von: Kang, Jae-Young, et al.
Veröffentlicht: (2026)
von: Kang, Jae-Young, et al.
Veröffentlicht: (2026)
Lifelong Learning with Task-Specific Adaptation: Addressing the Stability-Plasticity Dilemma
von: Wang, Ruiyu, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyu, et al.
Veröffentlicht: (2025)
RACon: Retrieval-Augmented Simulated Character Locomotion Control
von: Mu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Mu, Yuxuan, et al.
Veröffentlicht: (2024)
Binary Event-Driven Spiking Transformer
von: Cao, Honglin, et al.
Veröffentlicht: (2025)
von: Cao, Honglin, et al.
Veröffentlicht: (2025)
CoordAR: One-Reference 6D Pose Estimation of Novel Objects via Autoregressive Coordinate Map Generation
von: Zuo, Dexin, et al.
Veröffentlicht: (2025)
von: Zuo, Dexin, et al.
Veröffentlicht: (2025)
Efficient 3D Recognition with Event-driven Spike Sparse Convolution
von: Qiu, Xuerui, et al.
Veröffentlicht: (2024)
von: Qiu, Xuerui, et al.
Veröffentlicht: (2024)
$\text{Di}^2\text{Pose}$: Discrete Diffusion Model for Occluded 3D Human Pose Estimation
von: Wang, Weiquan, et al.
Veröffentlicht: (2024)
von: Wang, Weiquan, et al.
Veröffentlicht: (2024)
SpikeMOT: Event-based Multi-Object Tracking with Sparse Motion Features
von: Wang, Song, et al.
Veröffentlicht: (2023)
von: Wang, Song, et al.
Veröffentlicht: (2023)
SDTrack: A Baseline for Event-based Tracking via Spiking Neural Networks
von: Shan, Yimeng, et al.
Veröffentlicht: (2025)
von: Shan, Yimeng, et al.
Veröffentlicht: (2025)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
von: Qiu, Lingteng, et al.
Veröffentlicht: (2025)
GTPT: Group-based Token Pruning Transformer for Efficient Human Pose Estimation
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
BOOTPLACE: Bootstrapped Object Placement with Detection Transformers
von: Zhou, Hang, et al.
Veröffentlicht: (2025)
von: Zhou, Hang, et al.
Veröffentlicht: (2025)
SpikePool: Event-driven Spiking Transformer with Pooling Attention
von: Lee, Donghyun, et al.
Veröffentlicht: (2025)
von: Lee, Donghyun, et al.
Veröffentlicht: (2025)
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
von: Liu, Hanbing, et al.
Veröffentlicht: (2023)
Spectral Compression Transformer with Line Pose Graph for Monocular 3D Human Pose Estimation
von: Zheng, Zenghao, et al.
Veröffentlicht: (2025)
von: Zheng, Zenghao, et al.
Veröffentlicht: (2025)
S3CE-Net: Spike-guided Spatiotemporal Semantic Coupling and Expansion Network for Long Sequence Event Re-Identification
von: Ma, Xianheng, et al.
Veröffentlicht: (2025)
von: Ma, Xianheng, et al.
Veröffentlicht: (2025)
CREST: An Efficient Conjointly-trained Spike-driven Framework for Event-based Object Detection Exploiting Spatiotemporal Dynamics
von: Mao, Ruixin, et al.
Veröffentlicht: (2024)
von: Mao, Ruixin, et al.
Veröffentlicht: (2024)
Event-based Motion & Appearance Fusion for 6D Object Pose Tracking
von: Li, Zhichao, et al.
Veröffentlicht: (2026)
von: Li, Zhichao, et al.
Veröffentlicht: (2026)
Learning Human-Object Interaction for 3D Human Pose Estimation from LiDAR Point Clouds
von: Jung, Daniel Sungho, et al.
Veröffentlicht: (2026)
von: Jung, Daniel Sungho, et al.
Veröffentlicht: (2026)
Spike2Former: Efficient Spiking Transformer for High-performance Image Segmentation
von: Lei, Zhenxin, et al.
Veröffentlicht: (2024)
von: Lei, Zhenxin, et al.
Veröffentlicht: (2024)
A Novel Benchmark and Dataset for Efficient 3D Gaussian Splatting with Gaussian Point Cloud Compression
von: Wang, Kangli, et al.
Veröffentlicht: (2025)
von: Wang, Kangli, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Toward Real-Time Surgical Scene Segmentation via a Spike-Driven Video Transformer with Spike-Informed Pretraining
von: Zou, Shihao, et al.
Veröffentlicht: (2025) -
Sketch2PoseNet: Efficient and Generalized Sketch to 3D Human Pose Prediction
von: Wang, Li, et al.
Veröffentlicht: (2025) -
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
von: Zou, Shihao, et al.
Veröffentlicht: (2025) -
Exploiting Spatiotemporal Properties for Efficient Event-Driven Human Pose Estimation
von: Zhou, Haoxian, et al.
Veröffentlicht: (2025) -
TempDiffReg: Temporal Diffusion Model for Non-Rigid 2D-3D Vascular Registration
von: Liu, Zehua, et al.
Veröffentlicht: (2026)