Gespeichert in:
| Hauptverfasser: | Yan, Sheng, Wang, Yong, Du, Xin, Yuan, Junsong, Liu, Mengyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.08337 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MLP: Motion Label Prior for Temporal Sentence Localization in Untrimmed 3D Human Motions
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
MoSa: Motion Generation with Scalable Autoregressive Modeling
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
PoseMoE: Mixture-of-Experts Network for Monocular 3D Human Pose Estimation
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
von: Liu, Mengyuan, et al.
Veröffentlicht: (2026)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2026)
LaxMotion: Rethinking Supervision Granularity for 3D Human Motion Generation
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation
von: Wang, Xinshun, et al.
Veröffentlicht: (2026)
von: Wang, Xinshun, et al.
Veröffentlicht: (2026)
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
Eye Motion Matters for 3D Face Reconstruction
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
TokenMotion: Motion-Guided Vision Transformer for Video Camouflaged Object Detection Via Learnable Token Selection
von: Yu, Zifan, et al.
Veröffentlicht: (2023)
von: Yu, Zifan, et al.
Veröffentlicht: (2023)
Diffusion-based Pose Refinement and Muti-hypothesis Generation for 3D Human Pose Estimaiton
von: Kang, Hongbo, et al.
Veröffentlicht: (2024)
von: Kang, Hongbo, et al.
Veröffentlicht: (2024)
Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
von: Li, Wenhao, et al.
Veröffentlicht: (2023)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
Expressive Forecasting of 3D Whole-body Human Motions
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
von: Ding, Pengxiang, et al.
Veröffentlicht: (2023)
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
von: Li, Peiming, et al.
Veröffentlicht: (2025)
von: Li, Peiming, et al.
Veröffentlicht: (2025)
3D Skeleton-Based Action Recognition: A Review
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
Prompt When the Animal is: Temporal Animal Behavior Grounding with Positional Recovery Training
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
CrossGET: Cross-Guided Ensemble of Tokens for Accelerating Vision-Language Transformers
von: Shi, Dachuan, et al.
Veröffentlicht: (2023)
von: Shi, Dachuan, et al.
Veröffentlicht: (2023)
Motion Guided Token Compression for Efficient Masked Video Modeling
von: Feng, Yukun, et al.
Veröffentlicht: (2024)
von: Feng, Yukun, et al.
Veröffentlicht: (2024)
GTPT: Group-based Token Pruning Transformer for Efficient Human Pose Estimation
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
von: Wang, Yabiao, et al.
Veröffentlicht: (2024)
von: Wang, Yabiao, et al.
Veröffentlicht: (2024)
SRAM: Shape-Realism Alignment Metric for No Reference 3D Shape Evaluation
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
TokenMotion: Decoupled Motion Control via Token Disentanglement for Human-centric Video Generation
von: Li, Ruineng, et al.
Veröffentlicht: (2025)
von: Li, Ruineng, et al.
Veröffentlicht: (2025)
Forecasting Future Videos from Novel Views via Disentangled 3D Scene Representation
von: Yarram, Sudhir, et al.
Veröffentlicht: (2024)
von: Yarram, Sudhir, et al.
Veröffentlicht: (2024)
MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
Biomechanics-Guided Residual Approach to Generalizable Human Motion Generation and Estimation
von: Kang, Zixi, et al.
Veröffentlicht: (2025)
von: Kang, Zixi, et al.
Veröffentlicht: (2025)
Uni-Inter: Unifying 3D Human Motion Synthesis Across Diverse Interaction Contexts
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
FrankenMotion: Part-level Human Motion Generation and Composition
von: Li, Chuqiao, et al.
Veröffentlicht: (2026)
von: Li, Chuqiao, et al.
Veröffentlicht: (2026)
LumosFlow: Motion-Guided Long Video Generation
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
Morph: A Motion-free Physics Optimization Framework for Human Motion Generation
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
von: Jin, Peng, et al.
Veröffentlicht: (2024)
von: Jin, Peng, et al.
Veröffentlicht: (2024)
Human-in-Context: Unified Cross-Domain 3D Human Motion Modeling via In-Context Learning
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches
von: Yu, Qing, et al.
Veröffentlicht: (2024)
von: Yu, Qing, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MLP: Motion Label Prior for Temporal Sentence Localization in Untrimmed 3D Human Motions
von: Yan, Sheng, et al.
Veröffentlicht: (2024) -
MoSa: Motion Generation with Scalable Autoregressive Modeling
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025) -
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023) -
PoseMoE: Mixture-of-Experts Network for Monocular 3D Human Pose Estimation
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025) -
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
von: Liu, Mengyuan, et al.
Veröffentlicht: (2026)