Gespeichert in:
| Hauptverfasser: | Liu, Mengyuan, Yan, Sheng, Wang, Yong, Li, Yingjie, Bian, Gui-Bin, Liu, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2511.01200 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language-Guided Transformer Tokenizer for Human Motion Generation
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
MLP: Motion Label Prior for Temporal Sentence Localization in Untrimmed 3D Human Motions
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
von: Yan, Sheng, et al.
Veröffentlicht: (2024)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
Coordinate-Based Dual-Constrained Autoregressive Motion Generation
von: Ding, Kang, et al.
Veröffentlicht: (2026)
von: Ding, Kang, et al.
Veröffentlicht: (2026)
HINT: Hierarchical Interaction Modeling for Autoregressive Multi-Human Motion Generation
von: Liu, Mengge, et al.
Veröffentlicht: (2026)
von: Liu, Mengge, et al.
Veröffentlicht: (2026)
MoReact: Generating Reactive Motion from Textual Descriptions
von: Xu, Xiyan, et al.
Veröffentlicht: (2025)
von: Xu, Xiyan, et al.
Veröffentlicht: (2025)
Eye Motion Matters for 3D Face Reconstruction
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
ScaMo: Exploring the Scaling Law in Autoregressive Motion Generation Model
von: Lu, Shunlin, et al.
Veröffentlicht: (2024)
von: Lu, Shunlin, et al.
Veröffentlicht: (2024)
Next-Scale Autoregressive Models for Text-to-Motion Generation
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
von: Zheng, Zhiwei, et al.
Veröffentlicht: (2026)
UniMo: Unifying 2D Video and 3D Human Motion with an Autoregressive Framework
von: Pang, Youxin, et al.
Veröffentlicht: (2025)
von: Pang, Youxin, et al.
Veröffentlicht: (2025)
ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation
von: Hwang, Inwoo, et al.
Veröffentlicht: (2026)
von: Hwang, Inwoo, et al.
Veröffentlicht: (2026)
Scalable Autoregressive Image Generation with Mamba
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
MeMoSORT: Memory-Assisted Filtering and Motion-Adaptive Association Metric for Multi-Person Tracking
von: Wang, Yingjie, et al.
Veröffentlicht: (2025)
von: Wang, Yingjie, et al.
Veröffentlicht: (2025)
ClickDiff: Click to Induce Semantic Contact Map for Controllable Grasp Generation with Diffusion Models
von: Li, Peiming, et al.
Veröffentlicht: (2024)
von: Li, Peiming, et al.
Veröffentlicht: (2024)
LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens
von: Li, Zekun, et al.
Veröffentlicht: (2026)
von: Li, Zekun, et al.
Veröffentlicht: (2026)
Causal Motion Diffusion Models for Autoregressive Motion Generation
von: Yu, Qing, et al.
Veröffentlicht: (2026)
von: Yu, Qing, et al.
Veröffentlicht: (2026)
Motion-Aware Caching for Efficient Autoregressive Video Generation
von: Xu, Jing, et al.
Veröffentlicht: (2026)
von: Xu, Jing, et al.
Veröffentlicht: (2026)
CAR: Controllable Autoregressive Modeling for Visual Generation
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
Scalable Autoregressive Monocular Depth Estimation
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
von: Wang, Jinhong, et al.
Veröffentlicht: (2024)
A Survey on 3D Skeleton-Based Action Recognition Using Learning Method
von: Ren, Bin, et al.
Veröffentlicht: (2020)
von: Ren, Bin, et al.
Veröffentlicht: (2020)
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
von: Wang, Ziyi, et al.
Veröffentlicht: (2026)
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
von: Liu, Mengyuan, et al.
Veröffentlicht: (2023)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2023)
Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation
von: Wang, Xinshun, et al.
Veröffentlicht: (2026)
von: Wang, Xinshun, et al.
Veröffentlicht: (2026)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
von: Li, Zhe, et al.
Veröffentlicht: (2025)
von: Li, Zhe, et al.
Veröffentlicht: (2025)
SurMo: Surface-based 4D Motion Modeling for Dynamic Human Rendering
von: Hu, Tao, et al.
Veröffentlicht: (2024)
von: Hu, Tao, et al.
Veröffentlicht: (2024)
2nd of the 5th PVUW MeViS-Audio Track: ASR-SaSaSa2VA
von: Wang, Zhiyu, et al.
Veröffentlicht: (2026)
von: Wang, Zhiyu, et al.
Veröffentlicht: (2026)
BAMM: Bidirectional Autoregressive Motion Model
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
StarGen: A Spatiotemporal Autoregression Framework with Video Diffusion Model for Scalable and Controllable Scene Generation
von: Zhai, Shangjin, et al.
Veröffentlicht: (2025)
von: Zhai, Shangjin, et al.
Veröffentlicht: (2025)
DreamForge: Motion-Aware Autoregressive Video Generation for Multi-View Driving Scenes
von: Mei, Jianbiao, et al.
Veröffentlicht: (2024)
von: Mei, Jianbiao, et al.
Veröffentlicht: (2024)
Customize Your Visual Autoregressive Recipe with Set Autoregressive Modeling
von: Liu, Wenze, et al.
Veröffentlicht: (2024)
von: Liu, Wenze, et al.
Veröffentlicht: (2024)
MoGIC: Boosting Motion Generation via Intention Understanding and Visual Context
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
TrackSSM: A General Motion Predictor by State-Space Model
von: Hu, Bin, et al.
Veröffentlicht: (2024)
von: Hu, Bin, et al.
Veröffentlicht: (2024)
TCPFormer: Learning Temporal Correlation with Implicit Pose Proxy for 3D Human Pose Estimation
von: Liu, Jiajie, et al.
Veröffentlicht: (2025)
von: Liu, Jiajie, et al.
Veröffentlicht: (2025)
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
von: Sun, Peize, et al.
Veröffentlicht: (2024)
von: Sun, Peize, et al.
Veröffentlicht: (2024)
Identity-aware Dual-constraint Network for Cloth-Changing Person Re-identification
von: Guo, Peini, et al.
Veröffentlicht: (2024)
von: Guo, Peini, et al.
Veröffentlicht: (2024)
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
von: Wen, Boran, et al.
Veröffentlicht: (2025)
von: Wen, Boran, et al.
Veröffentlicht: (2025)
MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation
von: Li, Zhiqi, et al.
Veröffentlicht: (2025)
von: Li, Zhiqi, et al.
Veröffentlicht: (2025)
VideoMAP: Toward Scalable Mamba-based Video Autoregressive Pretraining
von: Liu, Yunze, et al.
Veröffentlicht: (2025)
von: Liu, Yunze, et al.
Veröffentlicht: (2025)
DiMo: Discrete Diffusion Modeling for Motion Generation and Understanding
von: Zhang, Ning, et al.
Veröffentlicht: (2026)
von: Zhang, Ning, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Language-Guided Transformer Tokenizer for Human Motion Generation
von: Yan, Sheng, et al.
Veröffentlicht: (2026) -
MLP: Motion Label Prior for Temporal Sentence Localization in Untrimmed 3D Human Motions
von: Yan, Sheng, et al.
Veröffentlicht: (2024) -
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023) -
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
von: Li, Yiheng, et al.
Veröffentlicht: (2026) -
Coordinate-Based Dual-Constrained Autoregressive Motion Generation
von: Ding, Kang, et al.
Veröffentlicht: (2026)