Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Qing, Tanaka, Mikihiro, Fujiwara, Kent |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chronologically Accurate Retrieval for Temporal Grounding of Motion-Language Models
von: Fujiwara, Kent, et al.
Veröffentlicht: (2024)
von: Fujiwara, Kent, et al.
Veröffentlicht: (2024)
Causal Motion Diffusion Models for Autoregressive Motion Generation
von: Yu, Qing, et al.
Veröffentlicht: (2026)
von: Yu, Qing, et al.
Veröffentlicht: (2026)
ProjFlow: Projection Sampling with Flow Matching for Zero-Shot Exact Spatial Motion Control
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2026)
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2026)
PINO: Person-Interaction Noise Optimization for Long-Duration and Customizable Motion Generation of Arbitrary-Sized Groups
von: Ota, Sakuya, et al.
Veröffentlicht: (2025)
von: Ota, Sakuya, et al.
Veröffentlicht: (2025)
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
von: Chen, Changan, et al.
Veröffentlicht: (2024)
von: Chen, Changan, et al.
Veröffentlicht: (2024)
Learning Priors of Human Motion With Vision Transformers
von: Falqueto, Placido, et al.
Veröffentlicht: (2025)
von: Falqueto, Placido, et al.
Veröffentlicht: (2025)
DrawMotion: Generating 3D Human Motions by Freehand Drawing
von: Wang, Tao, et al.
Veröffentlicht: (2026)
von: Wang, Tao, et al.
Veröffentlicht: (2026)
Language-Guided Transformer Tokenizer for Human Motion Generation
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
ReVision: Refining Video Diffusion with Explicit 3D Motion Modeling
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
von: Liu, Qihao, et al.
Veröffentlicht: (2025)
Motion-2-To-3: Leveraging 2D Motion Data for 3D Motion Generations
von: Guo, Ruoxi, et al.
Veröffentlicht: (2024)
von: Guo, Ruoxi, et al.
Veröffentlicht: (2024)
Exploring Motion-Language Alignment for Text-driven Motion Generation
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
von: Gu, Ruxi, et al.
Veröffentlicht: (2026)
LaxMotion: Rethinking Supervision Granularity for 3D Human Motion Generation
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
von: Liu, Sheng, et al.
Veröffentlicht: (2025)
QuaMo: Quaternion Motions for Vision-based 3D Human Kinematics Capture
von: Le, Cuong, et al.
Veröffentlicht: (2026)
von: Le, Cuong, et al.
Veröffentlicht: (2026)
MotionFix: Text-Driven 3D Human Motion Editing
von: Athanasiou, Nikos, et al.
Veröffentlicht: (2024)
von: Athanasiou, Nikos, et al.
Veröffentlicht: (2024)
Improving 3D Foot Motion Reconstruction in Markerless Monocular Human Motion Capture
von: Wehrbein, Tom, et al.
Veröffentlicht: (2026)
von: Wehrbein, Tom, et al.
Veröffentlicht: (2026)
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
AdvMT: Adversarial Motion Transformer for Long-term Human Motion Prediction
von: Idrees, Sarmad, et al.
Veröffentlicht: (2024)
von: Idrees, Sarmad, et al.
Veröffentlicht: (2024)
MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
von: Hong, Wenyi, et al.
Veröffentlicht: (2025)
Robust Human Motion Forecasting using Transformer-based Model
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2023)
von: Mascaro, Esteve Valls, et al.
Veröffentlicht: (2023)
StickMotion: Generating 3D Human Motions by Drawing a Stickman
von: Wang, Tao, et al.
Veröffentlicht: (2025)
von: Wang, Tao, et al.
Veröffentlicht: (2025)
TokenMotion: Motion-Guided Vision Transformer for Video Camouflaged Object Detection Via Learnable Token Selection
von: Yu, Zifan, et al.
Veröffentlicht: (2023)
von: Yu, Zifan, et al.
Veröffentlicht: (2023)
LingoMotion: An Interpretable and Unambiguous Symbolic Representation for Human Motion
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models
von: Zhang, Zhikai, et al.
Veröffentlicht: (2024)
von: Zhang, Zhikai, et al.
Veröffentlicht: (2024)
Modelling the Distribution of Human Motion for Sign Language Assessment
von: Cory, Oliver, et al.
Veröffentlicht: (2024)
von: Cory, Oliver, et al.
Veröffentlicht: (2024)
MotionScript: Natural Language Descriptions for Expressive 3D Human Motions
von: Yazdian, Payam Jome, et al.
Veröffentlicht: (2023)
von: Yazdian, Payam Jome, et al.
Veröffentlicht: (2023)
EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation
von: Hou, Ruibing, et al.
Veröffentlicht: (2026)
von: Hou, Ruibing, et al.
Veröffentlicht: (2026)
Mogo: RQ Hierarchical Causal Transformer for High-Quality 3D Human Motion Generation
von: Fu, Dongjie
Veröffentlicht: (2024)
von: Fu, Dongjie
Veröffentlicht: (2024)
LangPose: Language-Aligned Motion for Robust 3D Human Pose Estimation
von: Liao, Longyun, et al.
Veröffentlicht: (2024)
von: Liao, Longyun, et al.
Veröffentlicht: (2024)
Motion-X++: A Large-Scale Multimodal 3D Whole-body Human Motion Dataset
von: Zhang, Yuhong, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhong, et al.
Veröffentlicht: (2025)
Motion-X: A Large-scale 3D Expressive Whole-body Human Motion Dataset
von: Lin, Jing, et al.
Veröffentlicht: (2023)
von: Lin, Jing, et al.
Veröffentlicht: (2023)
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
von: Yang, Yuxiao, et al.
Veröffentlicht: (2025)
MotionPRO: Exploring the Role of Pressure in Human MoCap and Beyond
von: Ren, Shenghao, et al.
Veröffentlicht: (2025)
von: Ren, Shenghao, et al.
Veröffentlicht: (2025)
Semantic Belief-State World Model for 3D Human Motion Prediction
von: Chaudhry, Sarim
Veröffentlicht: (2026)
von: Chaudhry, Sarim
Veröffentlicht: (2026)
Encoder-Free Human Motion Understanding via Structured Motion Descriptions
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
Semantics-aware Motion Retargeting with Vision-Language Models
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
von: Zhang, Haodong, et al.
Veröffentlicht: (2023)
Text-driven Human Motion Generation with Motion Masked Diffusion Model
von: Chen, Xingyu
Veröffentlicht: (2024)
von: Chen, Xingyu
Veröffentlicht: (2024)
Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
von: Li, Chuqiao, et al.
Veröffentlicht: (2024)
von: Li, Chuqiao, et al.
Veröffentlicht: (2024)
Multimodal Sense-Informed Prediction of 3D Human Motions
von: Lou, Zhenyu, et al.
Veröffentlicht: (2024)
von: Lou, Zhenyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Chronologically Accurate Retrieval for Temporal Grounding of Motion-Language Models
von: Fujiwara, Kent, et al.
Veröffentlicht: (2024) -
Causal Motion Diffusion Models for Autoregressive Motion Generation
von: Yu, Qing, et al.
Veröffentlicht: (2026) -
ProjFlow: Projection Sampling with Flow Matching for Zero-Shot Exact Spatial Motion Control
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2026) -
PINO: Person-Interaction Noise Optimization for Long-Duration and Customizable Motion Generation of Arbitrary-Sized Groups
von: Ota, Sakuya, et al.
Veröffentlicht: (2025) -
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
von: Chen, Changan, et al.
Veröffentlicht: (2024)