MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zhe, He, Yisheng, Zhong, Lei, Shen, Weichao, Zuo, Qi, Qiu, Lingteng, Dong, Zilong, Yang, Laurence Tianruo, Yuan, Weihao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
by: Yuan, Weihao, et al.
Published: (2024)
by: Yuan, Weihao, et al.
Published: (2024)
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
by: Qiu, Lingteng, et al.
Published: (2024)
by: Qiu, Lingteng, et al.
Published: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
by: Zuo, Qi, et al.
Published: (2024)
by: Zuo, Qi, et al.
Published: (2024)
HyPlaneHead: Rethinking Tri-plane-like Representations in Full-Head Image Synthesis
by: Li, Heyuan, et al.
Published: (2025)
by: Li, Heyuan, et al.
Published: (2025)
LHM++: An Efficient Large Human Reconstruction Model for Pose-free Images to 3D
by: Qiu, Lingteng, et al.
Published: (2025)
by: Qiu, Lingteng, et al.
Published: (2025)
Large Depth Completion Model from Sparse Observations
by: Yu, Zhu, et al.
Published: (2026)
by: Yu, Zhu, et al.
Published: (2026)
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
by: Wu, Yushuang, et al.
Published: (2024)
by: Wu, Yushuang, et al.
Published: (2024)
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
by: He, Yisheng, et al.
Published: (2024)
by: He, Yisheng, et al.
Published: (2024)
SMooDi: Stylized Motion Diffusion Model
by: Zhong, Lei, et al.
Published: (2024)
by: Zhong, Lei, et al.
Published: (2024)
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
by: Chen, Minglin, et al.
Published: (2024)
by: Chen, Minglin, et al.
Published: (2024)
StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal
by: Ye, Chongjie, et al.
Published: (2024)
by: Ye, Chongjie, et al.
Published: (2024)
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
by: Han, Xiaoguang, et al.
Published: (2024)
by: Han, Xiaoguang, et al.
Published: (2024)
SMooGPT: Stylized Motion Generation using Large Language Models
by: Zhong, Lei, et al.
Published: (2025)
by: Zhong, Lei, et al.
Published: (2025)
D-LORD for Motion Stylization
by: Gupta, Meenakshi, et al.
Published: (2024)
by: Gupta, Meenakshi, et al.
Published: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
Generative Human Motion Stylization in Latent Space
by: Guo, Chuan, et al.
Published: (2024)
by: Guo, Chuan, et al.
Published: (2024)
Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos
by: Hu, Yingdong, et al.
Published: (2025)
by: Hu, Yingdong, et al.
Published: (2025)
EarthMapper: Visual Autoregressive Models for Controllable Bidirectional Satellite-Map Translation
by: Dong, Zhe, et al.
Published: (2025)
by: Dong, Zhe, et al.
Published: (2025)
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
by: He, Yisheng, et al.
Published: (2025)
by: He, Yisheng, et al.
Published: (2025)
Harmonizing Multi-Objective LLM Unlearning via Unified Domain Representation and Bidirectional Logit Distillation
by: Zhong, Yisheng, et al.
Published: (2026)
by: Zhong, Yisheng, et al.
Published: (2026)
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
by: Zhu, Shenhao, et al.
Published: (2024)
by: Zhu, Shenhao, et al.
Published: (2024)
CB-SMoT+: UNA EXTENSIÓN AL ALGORITMO CB-SMoT
by: Francisco Moreno
Published: (2012)
by: Francisco Moreno
Published: (2012)
GIC: Gaussian-Informed Continuum for Physical Property Identification and Simulation
by: Cai, Junhao, et al.
Published: (2024)
by: Cai, Junhao, et al.
Published: (2024)
An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D Meshes
by: Zhao, Zhengyi, et al.
Published: (2024)
by: Zhao, Zhengyi, et al.
Published: (2024)
PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image
by: Li, Peng, et al.
Published: (2025)
by: Li, Peng, et al.
Published: (2025)
ViSA: 3D-Aware Video Shading for Real-Time Upper-Body Avatar Creation
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
StylizedGS: Controllable Stylization for 3D Gaussian Splatting
by: Zhang, Dingxi, et al.
Published: (2024)
by: Zhang, Dingxi, et al.
Published: (2024)
CRAVES: Controlling Robotic Arm with a Vision-based Economic System
by: Zuo, Yiming, et al.
Published: (2018)
by: Zuo, Yiming, et al.
Published: (2018)
Condition Matters in Full-head 3D GANs
by: Li, Heyuan, et al.
Published: (2026)
by: Li, Heyuan, et al.
Published: (2026)
Euclid's Gift: Enhancing Spatial Perception and Reasoning in Vision-Language Models via Geometric Surrogate Tasks
by: Lian, Shijie, et al.
Published: (2025)
by: Lian, Shijie, et al.
Published: (2025)
DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry
by: Wu, Changti, et al.
Published: (2025)
by: Wu, Changti, et al.
Published: (2025)
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation
by: Dong, Zhe, et al.
Published: (2024)
by: Dong, Zhe, et al.
Published: (2024)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
by: Arazi, Alan, et al.
Published: (2026)
by: Arazi, Alan, et al.
Published: (2026)
CoSMo: A Multimodal Transformer for Page Stream Segmentation in Comic Books
by: Ortega, Marc Serra, et al.
Published: (2025)
by: Ortega, Marc Serra, et al.
Published: (2025)
HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction
by: Gu, Xiaodong, et al.
Published: (2024)
by: Gu, Xiaodong, et al.
Published: (2024)
Bandwidth-Efficient Two-Server ORAMs with O(1) Client Storage
by: Wang, Wei, et al.
Published: (2025)
by: Wang, Wei, et al.
Published: (2025)
One-Way Quantum Repeater with Rare-Earth-Ions Doped in Solids
by: Lei, Yisheng
Published: (2024)
by: Lei, Yisheng
Published: (2024)
Similar Items
-
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning
by: Li, Zhe, et al.
Published: (2024) -
OmniMotion: Multimodal Motion Generation with Continuous Masked Autoregression
by: Li, Zhe, et al.
Published: (2025) -
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
by: Yuan, Weihao, et al.
Published: (2024) -
LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds
by: Qiu, Lingteng, et al.
Published: (2025) -
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
by: Qiu, Lingteng, et al.
Published: (2024)