ExGes: Expressive Human Motion Retrieval and Modulation for Audio-Driven Gesture Synthesis
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Xukun, Li, Fengxin, Chen, Ming, Zhou, Yan, Wan, Pengfei, Zhang, Di, Jin, Yeying, Fan, Zhaoxin, Liu, Hongyan, He, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation
di: Zhou, Xukun, et al.
Pubblicazione: (2024)
di: Zhou, Xukun, et al.
Pubblicazione: (2024)
VarGes: Improving Variation in Co-Speech 3D Gesture Generation via StyleCLIPS
di: Meng, Ming, et al.
Pubblicazione: (2025)
di: Meng, Ming, et al.
Pubblicazione: (2025)
LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation
di: Li, Zhiying, et al.
Pubblicazione: (2025)
di: Li, Zhiying, et al.
Pubblicazione: (2025)
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
di: Gao, Nan, et al.
Pubblicazione: (2023)
di: Gao, Nan, et al.
Pubblicazione: (2023)
VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
Ges-QA: A Multidimensional Quality Assessment Dataset for Audio-to-3D Gesture Generation
di: Gao, Zhilin, et al.
Pubblicazione: (2025)
di: Gao, Zhilin, et al.
Pubblicazione: (2025)
GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations
di: Guo, Wenxuan, et al.
Pubblicazione: (2026)
di: Guo, Wenxuan, et al.
Pubblicazione: (2026)
Not All Frames Are Equal: Complexity-Aware Masked Motion Generation via Motion Spectral Descriptors
di: Zhou, Pengfei, et al.
Pubblicazione: (2026)
di: Zhou, Pengfei, et al.
Pubblicazione: (2026)
Text-driven 3D Human Generation via Contrastive Preference Optimization
di: Zhou, Pengfei, et al.
Pubblicazione: (2025)
di: Zhou, Pengfei, et al.
Pubblicazione: (2025)
Lipschitz-Driven Noise Robustness in VQ-AE for High-Frequency Texture Repair in ID-Specific Talking Heads
di: Yang, Jian, et al.
Pubblicazione: (2024)
di: Yang, Jian, et al.
Pubblicazione: (2024)
M3G: Multi-Granular Gesture Generator for Audio-Driven Full-Body Human Motion Synthesis
di: Yin, Zhizhuo, et al.
Pubblicazione: (2025)
di: Yin, Zhizhuo, et al.
Pubblicazione: (2025)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
di: Liu, Haiyang, et al.
Pubblicazione: (2023)
di: Liu, Haiyang, et al.
Pubblicazione: (2023)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
di: Wang, Siyuan, et al.
Pubblicazione: (2025)
di: Wang, Siyuan, et al.
Pubblicazione: (2025)
MIBURI: Towards Expressive Interactive Gesture Synthesis
di: Mughal, M. Hamza, et al.
Pubblicazione: (2026)
di: Mughal, M. Hamza, et al.
Pubblicazione: (2026)
GesPrompt: Leveraging Co-Speech Gestures to Augment LLM-Based Interaction in Virtual Reality
di: Hu, Xiyun, et al.
Pubblicazione: (2025)
di: Hu, Xiyun, et al.
Pubblicazione: (2025)
Conversational Gesture Model (CGM): Extending Speaker‐Centric Audio‐Driven Motion Generation to Full Conversation Gestures
di: T. Koren, et al.
Pubblicazione: (2026)
di: T. Koren, et al.
Pubblicazione: (2026)
DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech
di: Cheng, Yongkang, et al.
Pubblicazione: (2025)
di: Cheng, Yongkang, et al.
Pubblicazione: (2025)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
di: Liu, Lanmiao, et al.
Pubblicazione: (2026)
di: Liu, Lanmiao, et al.
Pubblicazione: (2026)
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
di: Liu, Lanmiao, et al.
Pubblicazione: (2025)
di: Liu, Lanmiao, et al.
Pubblicazione: (2025)
EmoDiffGes: Emotion‐Aware Co‐Speech Holistic Gesture Generation with Progressive Synergistic Diffusion
di: Xinru Li, et al.
Pubblicazione: (2025)
di: Xinru Li, et al.
Pubblicazione: (2025)
ELGAR: Expressive Cello Performance Motion Generation for Audio Rendition
di: Qiu, Zhiping, et al.
Pubblicazione: (2025)
di: Qiu, Zhiping, et al.
Pubblicazione: (2025)
SnapMoGen: Human Motion Generation from Expressive Texts
di: Guo, Chuan, et al.
Pubblicazione: (2025)
di: Guo, Chuan, et al.
Pubblicazione: (2025)
KAN v.s. MLP for Offline Reinforcement Learning
di: Guo, Haihong, et al.
Pubblicazione: (2024)
di: Guo, Haihong, et al.
Pubblicazione: (2024)
EMG-UP: Unsupervised Personalization in Cross-User EMG Gesture Recognition
di: Wang, Nana, et al.
Pubblicazione: (2025)
di: Wang, Nana, et al.
Pubblicazione: (2025)
SARGes: Semantically Aligned Reliable Gesture Generation via Intent Chain
di: Gao, Nan, et al.
Pubblicazione: (2025)
di: Gao, Nan, et al.
Pubblicazione: (2025)
Ges3ViG: Incorporating Pointing Gestures into Language-Based 3D Visual Grounding for Embodied Reference Understanding
di: Mane, Atharv Mahesh, et al.
Pubblicazione: (2025)
di: Mane, Atharv Mahesh, et al.
Pubblicazione: (2025)
A Priori Error Estimate for H1$$ {H}^1 $$ Galerkin Mixed Virtual Element Discretization of Parabolic Equations
di: Fengxin Chen, et al.
Pubblicazione: (2025)
di: Fengxin Chen, et al.
Pubblicazione: (2025)
MOSPA: Human Motion Generation Driven by Spatial Audio
di: Xu, Shuyang, et al.
Pubblicazione: (2025)
di: Xu, Shuyang, et al.
Pubblicazione: (2025)
Erased, But Not Forgotten: Erased Rectified Flow Transformers Still Remain Unsafe Under Concept Attack
di: Jiang, Nanxiang, et al.
Pubblicazione: (2025)
di: Jiang, Nanxiang, et al.
Pubblicazione: (2025)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
di: Cheng, Yongkang, et al.
Pubblicazione: (2025)
di: Cheng, Yongkang, et al.
Pubblicazione: (2025)
Unveiling Hidden Vulnerabilities in Digital Human Generation via Adversarial Attacks
di: Li, Zhiying, et al.
Pubblicazione: (2025)
di: Li, Zhiying, et al.
Pubblicazione: (2025)
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
di: Qi, Xingqun, et al.
Pubblicazione: (2023)
di: Qi, Xingqun, et al.
Pubblicazione: (2023)
Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions
di: Liao, Ting-Hsuan, et al.
Pubblicazione: (2025)
di: Liao, Ting-Hsuan, et al.
Pubblicazione: (2025)
SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
di: Peng, Ziqiao, et al.
Pubblicazione: (2023)
GestureHYDRA: Semantic Co-speech Gesture Synthesis via Hybrid Modality Diffusion Transformer and Cascaded-Synchronized Retrieval-Augmented Generation
di: Yang, Quanwei, et al.
Pubblicazione: (2025)
di: Yang, Quanwei, et al.
Pubblicazione: (2025)
Checkpoints for paper "Visualising Pianists' Touch: Transcribing Expressive Piano Performance from Audio to Piano Key Motion"
di: Tang, Jingjing, et al.
Pubblicazione: (2026)
di: Tang, Jingjing, et al.
Pubblicazione: (2026)
Knitted Pneumatic Fabrics for Dynamic Pressure Modulation in Personalized Healthcare Wearables
di: Xiaoyu Chen, et al.
Pubblicazione: (2026)
di: Xiaoyu Chen, et al.
Pubblicazione: (2026)
SHLE: Devices Tracking and Depth Filtering for Stereo-based Height Limit Estimation
di: Fan, Zhaoxin, et al.
Pubblicazione: (2022)
di: Fan, Zhaoxin, et al.
Pubblicazione: (2022)
SMamDiff: Spatial Mamba for Stochastic Human Motion Prediction
di: Fan, Junqiao, et al.
Pubblicazione: (2025)
di: Fan, Junqiao, et al.
Pubblicazione: (2025)
CoheDancers: Enhancing Interactive Group Dance Generation through Music-Driven Coherence Decomposition
di: Yang, Kaixing, et al.
Pubblicazione: (2024)
di: Yang, Kaixing, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Meta-Learning Empowered Meta-Face: Personalized Speaking Style Adaptation for Audio-Driven 3D Talking Face Animation
di: Zhou, Xukun, et al.
Pubblicazione: (2024) -
VarGes: Improving Variation in Co-Speech 3D Gesture Generation via StyleCLIPS
di: Meng, Ming, et al.
Pubblicazione: (2025) -
LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation
di: Li, Zhiying, et al.
Pubblicazione: (2025) -
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
di: Gao, Nan, et al.
Pubblicazione: (2023) -
VGG-Tex: A Vivid Geometry-Guided Facial Texture Estimation Model for High Fidelity Monocular 3D Face Reconstruction
di: Wu, Haoyu, et al.
Pubblicazione: (2024)