Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Fan, Ke, Zhang, Jiangning, Yi, Ran, Gong, Jingyu, Wang, Yabiao, Wang, Yating, Tan, Xin, Wang, Chengjie, Ma, Lizhuang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
di: Fan, Ke, et al.
Pubblicazione: (2024)
di: Fan, Ke, et al.
Pubblicazione: (2024)
MotionMaster: Training-free Camera Motion Transfer For Video Generation
di: Hu, Teng, et al.
Pubblicazione: (2024)
di: Hu, Teng, et al.
Pubblicazione: (2024)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
di: Wang, Sen, et al.
Pubblicazione: (2024)
di: Wang, Sen, et al.
Pubblicazione: (2024)
SaRA: High-Efficient Diffusion Model Fine-tuning with Progressive Sparse Low-Rank Adaptation
di: Hu, Teng, et al.
Pubblicazione: (2024)
di: Hu, Teng, et al.
Pubblicazione: (2024)
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
di: Wang, Yuji, et al.
Pubblicazione: (2025)
di: Wang, Yuji, et al.
Pubblicazione: (2025)
Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
di: Wang, Yuji, et al.
Pubblicazione: (2025)
di: Wang, Yuji, et al.
Pubblicazione: (2025)
Semantic Frame Interpolation
di: Hong, Yijia, et al.
Pubblicazione: (2025)
di: Hong, Yijia, et al.
Pubblicazione: (2025)
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
di: Hu, Teng, et al.
Pubblicazione: (2025)
di: Hu, Teng, et al.
Pubblicazione: (2025)
ID-Sculpt: ID-aware 3D Head Generation from Single In-the-wild Portrait Image
di: Hao, Jinkun, et al.
Pubblicazione: (2024)
di: Hao, Jinkun, et al.
Pubblicazione: (2024)
LLaVA-VSD: Large Language-and-Vision Assistant for Visual Spatial Description
di: Jin, Yizhang, et al.
Pubblicazione: (2024)
di: Jin, Yizhang, et al.
Pubblicazione: (2024)
TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
di: Wang, Yabiao, et al.
Pubblicazione: (2024)
di: Wang, Yabiao, et al.
Pubblicazione: (2024)
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation
di: Chen, Yuheng, et al.
Pubblicazione: (2026)
di: Chen, Yuheng, et al.
Pubblicazione: (2026)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
di: Hu, Teng, et al.
Pubblicazione: (2023)
di: Hu, Teng, et al.
Pubblicazione: (2023)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
di: Wang, Yating, et al.
Pubblicazione: (2024)
di: Wang, Yating, et al.
Pubblicazione: (2024)
UniM-OV3D: Uni-Modality Open-Vocabulary 3D Scene Understanding with Fine-Grained Feature Representation
di: He, Qingdong, et al.
Pubblicazione: (2024)
di: He, Qingdong, et al.
Pubblicazione: (2024)
M3DM-NR: RGB-3D Noisy-Resistant Industrial Anomaly Detection via Multimodal Denoising
di: Wang, Chengjie, et al.
Pubblicazione: (2024)
di: Wang, Chengjie, et al.
Pubblicazione: (2024)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
di: Tai, Hanchen, et al.
Pubblicazione: (2024)
di: Tai, Hanchen, et al.
Pubblicazione: (2024)
Continuous Piecewise-Affine Based Motion Model for Image Animation
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
CtlGAN: Few-shot Artistic Portraits Generation with Contrastive Transfer Learning
di: Wang, Yue, et al.
Pubblicazione: (2022)
di: Wang, Yue, et al.
Pubblicazione: (2022)
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction
di: Yi, Ran, et al.
Pubblicazione: (2025)
di: Yi, Ran, et al.
Pubblicazione: (2025)
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
di: Wen, Boran, et al.
Pubblicazione: (2025)
di: Wen, Boran, et al.
Pubblicazione: (2025)
3D Gaussian Head Avatars with Expressive Dynamic Appearances by Compact Tensorial Representations
di: Wang, Yating, et al.
Pubblicazione: (2025)
di: Wang, Yating, et al.
Pubblicazione: (2025)
Exploring Real&Synthetic Dataset and Linear Attention in Image Restoration
di: Du, Yuzhen, et al.
Pubblicazione: (2024)
di: Du, Yuzhen, et al.
Pubblicazione: (2024)
InstanceV: Instance-Level Video Generation
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
UniForward: Unified 3D Scene and Semantic Field Reconstruction via Feed-Forward Gaussian Splatting from Only Sparse-View Images
di: Tian, Qijian, et al.
Pubblicazione: (2025)
di: Tian, Qijian, et al.
Pubblicazione: (2025)
AdR-Gaussian: Accelerating Gaussian Splatting with Adaptive Radius
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
PVG: Progressive Vision Graph for Vision Recognition
di: Wu, Jiafu, et al.
Pubblicazione: (2023)
di: Wu, Jiafu, et al.
Pubblicazione: (2023)
PiT: Progressive Diffusion Transformer
di: Wu, Jiafu, et al.
Pubblicazione: (2025)
di: Wu, Jiafu, et al.
Pubblicazione: (2025)
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm
di: Zhang, Jiangning, et al.
Pubblicazione: (2022)
di: Zhang, Jiangning, et al.
Pubblicazione: (2022)
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
di: He, Qingdong, et al.
Pubblicazione: (2024)
di: He, Qingdong, et al.
Pubblicazione: (2024)
Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection
di: Wang, Haoxuan, et al.
Pubblicazione: (2024)
di: Wang, Haoxuan, et al.
Pubblicazione: (2024)
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
di: Liu, Dongqi, et al.
Pubblicazione: (2026)
di: Liu, Dongqi, et al.
Pubblicazione: (2026)
DEMOS: Dynamic Environment Motion Synthesis in 3D Scenes via Local Spherical-BEV Perception
di: Gong, Jingyu, et al.
Pubblicazione: (2024)
di: Gong, Jingyu, et al.
Pubblicazione: (2024)
HeadLighter: Disentangling Illumination in Generative 3D Gaussian Heads via Lightstage Captures
di: Wang, Yating, et al.
Pubblicazione: (2026)
di: Wang, Yating, et al.
Pubblicazione: (2026)
Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory
di: Liu, Jinzhuo, et al.
Pubblicazione: (2026)
di: Liu, Jinzhuo, et al.
Pubblicazione: (2026)
GSCompleter: A Distillation-Free Plugin for Metric-Aware 3D Gaussian Splatting Completion in Seconds
di: Gao, Ao, et al.
Pubblicazione: (2026)
di: Gao, Ao, et al.
Pubblicazione: (2026)
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection
di: Chen, Xuhai, et al.
Pubblicazione: (2023)
di: Chen, Xuhai, et al.
Pubblicazione: (2023)
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection
di: Zhang, Jiangning, et al.
Pubblicazione: (2023)
di: Zhang, Jiangning, et al.
Pubblicazione: (2023)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
di: Mao, Xiaofeng, et al.
Pubblicazione: (2024)
di: Mao, Xiaofeng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
di: Fan, Ke, et al.
Pubblicazione: (2024) -
MotionMaster: Training-free Camera Motion Transfer For Video Generation
di: Hu, Teng, et al.
Pubblicazione: (2024) -
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
di: Wang, Sen, et al.
Pubblicazione: (2024) -
SaRA: High-Efficient Diffusion Model Fine-tuning with Progressive Sparse Low-Rank Adaptation
di: Hu, Teng, et al.
Pubblicazione: (2024) -
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
di: Wang, Yuji, et al.
Pubblicazione: (2025)