Mitigating Error Accumulation in Co-Speech Motion Generation via Global Rotation Diffusion and Multi-Level Constraints
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xiangyue, Li, Jianfang, Ren, Jianqiang, Zhang, Jiaxu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025)
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
Not All Frames Are Equal: Complexity-Aware Masked Motion Generation via Motion Spectral Descriptors
von: Zhou, Pengfei, et al.
Veröffentlicht: (2026)
von: Zhou, Pengfei, et al.
Veröffentlicht: (2026)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Mitigating Error Accumulation in Continuous Navigation via Memory-Augmented Kalman Filtering
von: Tang, Yin, et al.
Veröffentlicht: (2026)
von: Tang, Yin, et al.
Veröffentlicht: (2026)
Generative Motion Stylization of Cross-structure Characters within Canonical Motion Space
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
von: Liu, Lin, et al.
Veröffentlicht: (2025)
von: Liu, Lin, et al.
Veröffentlicht: (2025)
SpeechAct: Towards Generating Whole-body Motion from Speech
von: Zhang, Jinsong, et al.
Veröffentlicht: (2023)
von: Zhang, Jinsong, et al.
Veröffentlicht: (2023)
CoPESD: A Multi-Level Surgical Motion Dataset for Training Large Vision-Language Models to Co-Pilot Endoscopic Submucosal Dissection
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
Unison: Harmonizing Motion, Speech, and Sound for Human-Centric Audio-Video Generation
von: Cheng, Shihao, et al.
Veröffentlicht: (2026)
von: Cheng, Shihao, et al.
Veröffentlicht: (2026)
Iterative Ensemble Training with Anti-Gradient Control for Mitigating Memorization in Diffusion Models
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
von: Liu, Xiao, et al.
Veröffentlicht: (2024)
Realistic Human Motion Generation with Cross-Diffusion Models
von: Ren, Zeping, et al.
Veröffentlicht: (2023)
von: Ren, Zeping, et al.
Veröffentlicht: (2023)
Towards Variable and Coordinated Holistic Co-Speech Motion Generation
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
von: Liu, Yifei, et al.
Veröffentlicht: (2024)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
von: Song, Yafei, et al.
Veröffentlicht: (2025)
von: Song, Yafei, et al.
Veröffentlicht: (2025)
Redistribute Ensemble Training for Mitigating Memorization in Diffusion Models
von: Guan, Xiaoliu, et al.
Veröffentlicht: (2025)
von: Guan, Xiaoliu, et al.
Veröffentlicht: (2025)
TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
von: Liu, Haiyang, et al.
Veröffentlicht: (2024)
von: Liu, Haiyang, et al.
Veröffentlicht: (2024)
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
ReDiffuse: Rotation Equivariant Diffusion Model for Multi-focus Image Fusion
von: Li, Bo, et al.
Veröffentlicht: (2026)
von: Li, Bo, et al.
Veröffentlicht: (2026)
Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
von: Chen, Bohong, et al.
Veröffentlicht: (2024)
von: Chen, Bohong, et al.
Veröffentlicht: (2024)
InterGen: Diffusion-based Multi-human Motion Generation under Complex Interactions
von: Liang, Han, et al.
Veröffentlicht: (2023)
von: Liang, Han, et al.
Veröffentlicht: (2023)
CoMA: Compositional Human Motion Generation with Multi-modal Agents
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
OMG-Avatar: One-shot Multi-LOD Gaussian Head Avatar
von: Ren, Jianqiang, et al.
Veröffentlicht: (2026)
von: Ren, Jianqiang, et al.
Veröffentlicht: (2026)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
Photovoltaic Defect Image Generator with Boundary Alignment Smoothing Constraint for Domain Shift Mitigation
von: Li, Dongying, et al.
Veröffentlicht: (2025)
von: Li, Dongying, et al.
Veröffentlicht: (2025)
MACE-Dance: Motion-Appearance Cascaded Experts for Music-Driven Dance Video Generation
von: Yang, Kaixing, et al.
Veröffentlicht: (2025)
von: Yang, Kaixing, et al.
Veröffentlicht: (2025)
Frequency Decoupling for Motion Magnification via Multi-Level Isomorphic Architecture
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
Incomplete Multi-view Clustering via Diffusion Contrastive Generation
von: Zhang, Yuanyang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuanyang, et al.
Veröffentlicht: (2025)
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenghao, et al.
Veröffentlicht: (2025)
MotionGrounder: Grounded Multi-Object Motion Transfer via Diffusion Transformer
von: Teodoro, Samuel, et al.
Veröffentlicht: (2026)
von: Teodoro, Samuel, et al.
Veröffentlicht: (2026)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
von: Mughal, Muhammad Hamza, et al.
Veröffentlicht: (2024)
von: Mughal, Muhammad Hamza, et al.
Veröffentlicht: (2024)
FIRE: Robust Detection of Diffusion-Generated Images via Frequency-Guided Reconstruction Error
von: Chu, Beilin, et al.
Veröffentlicht: (2024)
von: Chu, Beilin, et al.
Veröffentlicht: (2024)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
von: Wang, Siyuan, et al.
Veröffentlicht: (2025)
DiMo: Discrete Diffusion Modeling for Motion Generation and Understanding
von: Zhang, Ning, et al.
Veröffentlicht: (2026)
von: Zhang, Ning, et al.
Veröffentlicht: (2026)
MikuDance: Animating Character Art with Mixed Motion Dynamics
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2024)
Large Motion Model for Unified Multi-Modal Motion Generation
von: Zhang, Mingyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Mingyuan, et al.
Veröffentlicht: (2024)
Expressive Keypoints for Skeleton-based Action Recognition via Skeleton Transformation
von: Yang, Yijie, et al.
Veröffentlicht: (2024)
von: Yang, Yijie, et al.
Veröffentlicht: (2024)
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
von: Jin, Peng, et al.
Veröffentlicht: (2024)
von: Jin, Peng, et al.
Veröffentlicht: (2024)
SimDiff: Simulator-constrained Diffusion Model for Physically Plausible Motion Generation
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2025)
von: Watanabe, Akihisa, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2025) -
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2024) -
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026) -
Not All Frames Are Equal: Complexity-Aware Masked Motion Generation via Motion Spectral Descriptors
von: Zhou, Pengfei, et al.
Veröffentlicht: (2026) -
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)