MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
Fuente:
arXiv
Salvato in:
| Autori principali: | Yuan, Weihao, Shen, Weichao, He, Yisheng, Dong, Yuan, Gu, Xiaodong, Dong, Zilong, Bo, Liefeng, Huang, Qixing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
di: He, Yisheng, et al.
Pubblicazione: (2024)
di: He, Yisheng, et al.
Pubblicazione: (2024)
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
di: Li, Zhe, et al.
Pubblicazione: (2024)
di: Li, Zhe, et al.
Pubblicazione: (2024)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
di: He, Yisheng, et al.
Pubblicazione: (2025)
di: He, Yisheng, et al.
Pubblicazione: (2025)
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes
di: Zhao, Zhengyi, et al.
Pubblicazione: (2024)
di: Zhao, Zhengyi, et al.
Pubblicazione: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
di: Zuo, Qi, et al.
Pubblicazione: (2024)
di: Zuo, Qi, et al.
Pubblicazione: (2024)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
di: Li, Zhe, et al.
Pubblicazione: (2025)
di: Li, Zhe, et al.
Pubblicazione: (2025)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
di: Cai, Junhao, et al.
Pubblicazione: (2024)
di: Cai, Junhao, et al.
Pubblicazione: (2024)
MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
di: Li, Zhe, et al.
Pubblicazione: (2024)
di: Li, Zhe, et al.
Pubblicazione: (2024)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
di: Chen, Minglin, et al.
Pubblicazione: (2024)
di: Chen, Minglin, et al.
Pubblicazione: (2024)
MoSAM: Motion-Guided Segment Anything Model with Spatial-Temporal Memory Selection
di: Yang, Qiushi, et al.
Pubblicazione: (2025)
di: Yang, Qiushi, et al.
Pubblicazione: (2025)
GIC: Gaussian-Informed Continuum for Physical Property Identification and Simulation
di: Cai, Junhao, et al.
Pubblicazione: (2024)
di: Cai, Junhao, et al.
Pubblicazione: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
di: He, Chao, et al.
Pubblicazione: (2025)
di: He, Chao, et al.
Pubblicazione: (2025)
HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction
di: Gu, Xiaodong, et al.
Pubblicazione: (2024)
di: Gu, Xiaodong, et al.
Pubblicazione: (2024)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
di: Qiu, Lingteng, et al.
Pubblicazione: (2024)
di: Qiu, Lingteng, et al.
Pubblicazione: (2024)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
di: Li, Peng, et al.
Pubblicazione: (2025)
di: Li, Peng, et al.
Pubblicazione: (2025)
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
di: Wu, Yushuang, et al.
Pubblicazione: (2024)
di: Wu, Yushuang, et al.
Pubblicazione: (2024)
MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling
di: Men, Yifang, et al.
Pubblicazione: (2024)
di: Men, Yifang, et al.
Pubblicazione: (2024)
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
di: Hu, Yingdong, et al.
Pubblicazione: (2025)
di: Hu, Yingdong, et al.
Pubblicazione: (2025)
GenAnalysis: Joint Shape Analysis by Learning Man-Made Shape Generators with Deformation Regularizations
di: Yang, Yuezhi, et al.
Pubblicazione: (2025)
di: Yang, Yuezhi, et al.
Pubblicazione: (2025)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
di: Ye, Chongjie, et al.
Pubblicazione: (2024)
di: Ye, Chongjie, et al.
Pubblicazione: (2024)
GenCorres: Consistent Shape Matching via Coupled Implicit-Explicit Shape Generative Models
di: Yang, Haitao, et al.
Pubblicazione: (2023)
di: Yang, Haitao, et al.
Pubblicazione: (2023)
Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation
di: Chen, Yingjie, et al.
Pubblicazione: (2025)
di: Chen, Yingjie, et al.
Pubblicazione: (2025)
Polyp-Gen: Realistic and Diverse Polyp Image Generation for Endoscopic Dataset Expansion
di: Liu, Shengyuan, et al.
Pubblicazione: (2025)
di: Liu, Shengyuan, et al.
Pubblicazione: (2025)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
di: Qiu, Lingteng, et al.
Pubblicazione: (2025)
CoGenAV: Versatile Audio-Visual Representation Learning via Contrastive-Generative Synchronization
di: Bai, Detao, et al.
Pubblicazione: (2025)
di: Bai, Detao, et al.
Pubblicazione: (2025)
CrowdMoGen: Zero-Shot Text-Driven Collective Motion Generation
di: Cao, Yukang, et al.
Pubblicazione: (2024)
di: Cao, Yukang, et al.
Pubblicazione: (2024)
Atlas Gaussians Diffusion for 3D Generation
di: Yang, Haitao, et al.
Pubblicazione: (2024)
di: Yang, Haitao, et al.
Pubblicazione: (2024)
UniMoGen: Universal Motion Generation
di: Khani, Aliasghar, et al.
Pubblicazione: (2025)
di: Khani, Aliasghar, et al.
Pubblicazione: (2025)
TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation
di: Zhang, Hongyu, et al.
Pubblicazione: (2026)
di: Zhang, Hongyu, et al.
Pubblicazione: (2026)
CoProSketch: Controllable and Progressive Sketch Generation with Diffusion Model
di: Zhan, Ruohao, et al.
Pubblicazione: (2025)
di: Zhan, Ruohao, et al.
Pubblicazione: (2025)
DicFace: Dirichlet-Constrained Variational Codebook Learning for Temporally Coherent Video Face Restoration
di: Chen, Yan, et al.
Pubblicazione: (2025)
di: Chen, Yan, et al.
Pubblicazione: (2025)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
di: Zhu, Shenhao, et al.
Pubblicazione: (2024)
di: Zhu, Shenhao, et al.
Pubblicazione: (2024)
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
di: Han, Xiaoguang, et al.
Pubblicazione: (2024)
di: Han, Xiaoguang, et al.
Pubblicazione: (2024)
GaussianIP: Identity-Preserving Realistic 3D Human Generation via Human-Centric Diffusion Prior
di: Tang, Zichen, et al.
Pubblicazione: (2025)
di: Tang, Zichen, et al.
Pubblicazione: (2025)
OmniMoGen: Unifying Human Motion Generation via Learning from Interleaved Text-Motion Instructions
di: Bu, Wendong, et al.
Pubblicazione: (2025)
di: Bu, Wendong, et al.
Pubblicazione: (2025)
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
di: Zhang, Xiangyue, et al.
Pubblicazione: (2025)
di: Zhang, Xiangyue, et al.
Pubblicazione: (2025)
PPLNs: Parametric Piecewise Linear Networks for Event-Based Temporal Modeling and Beyond
di: Song, Chen, et al.
Pubblicazione: (2024)
di: Song, Chen, et al.
Pubblicazione: (2024)
MoReGen: Multi-Agent Motion-Reasoning Engine for Code-based Text-to-Video Synthesis
di: Bai, Xiangyu, et al.
Pubblicazione: (2025)
di: Bai, Xiangyu, et al.
Pubblicazione: (2025)
Exploring Timeline Control for Facial Motion Generation
di: Ma, Yifeng, et al.
Pubblicazione: (2025)
di: Ma, Yifeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
di: He, Yisheng, et al.
Pubblicazione: (2024) -
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
di: Li, Zhe, et al.
Pubblicazione: (2024) -
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
di: He, Yisheng, et al.
Pubblicazione: (2025) -
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes
di: Zhao, Zhengyi, et al.
Pubblicazione: (2024) -
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
di: Zuo, Qi, et al.
Pubblicazione: (2024)