A Unified Editing Method for Co-Speech Gesture Generation via Diffusion Inversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Zeyu, Gao, Nan, Zeng, Zhi, Zhang, Guixuan, Liu, Jie, Zhang, Shuwu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
von: Gao, Nan, et al.
Veröffentlicht: (2023)
von: Gao, Nan, et al.
Veröffentlicht: (2023)
Conversational Co-Speech Gesture Generation via Modeling Dialog Intention, Emotion, and Context with Diffusion Models
von: Xue, Haiwei, et al.
Veröffentlicht: (2023)
von: Xue, Haiwei, et al.
Veröffentlicht: (2023)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
SIGGesture: Generalized Co-Speech Gesture Synthesis via Semantic Injection with Large-Scale Pre-Training Diffusion Models
von: Cheng, Qingrong, et al.
Veröffentlicht: (2024)
von: Cheng, Qingrong, et al.
Veröffentlicht: (2024)
GesPrompt: Leveraging Co-Speech Gestures to Augment LLM-Based Interaction in Virtual Reality
von: Hu, Xiyun, et al.
Veröffentlicht: (2025)
von: Hu, Xiyun, et al.
Veröffentlicht: (2025)
SARGes: Semantically Aligned Reliable Gesture Generation via Intent Chain
von: Gao, Nan, et al.
Veröffentlicht: (2025)
von: Gao, Nan, et al.
Veröffentlicht: (2025)
GestureGPT: Toward Zero-Shot Free-Form Hand Gesture Understanding with Large Language Model Agents
von: Zeng, Xin, et al.
Veröffentlicht: (2023)
von: Zeng, Xin, et al.
Veröffentlicht: (2023)
WiOpen: A Robust Wi-Fi-based Open-set Gesture Recognition Framework
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
von: Zhang, Xiang, et al.
Veröffentlicht: (2024)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
von: Voss, Hendric, et al.
Veröffentlicht: (2025)
Optimizing Gesture Recognition for Seamless UI Interaction Using Convolutional Neural Networks
von: Sun, Qi, et al.
Veröffentlicht: (2024)
von: Sun, Qi, et al.
Veröffentlicht: (2024)
Speech-Gesture Mapping and Engagement Evaluation in Human Robot Interaction
von: Ghosh, Bishal, et al.
Veröffentlicht: (2018)
von: Ghosh, Bishal, et al.
Veröffentlicht: (2018)
DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
Beyond Physical Labels: Redefining Domains for Robust WiFi-based Gesture Recognition
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
OpenWatch: A Multimodal Benchmark for Hand Gesture Recognition on Smartwatches
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2026)
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2026)
Text2Gestures: A Transformer-Based Network for Generating Emotive Body Gestures for Virtual Agents
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2021)
GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture Recommendations
von: Ram, Ashwin, et al.
Veröffentlicht: (2025)
von: Ram, Ashwin, et al.
Veröffentlicht: (2025)
TalkLess: Blending Extractive and Abstractive Speech Summarization for Editing Speech to Preserve Content and Style
von: Benharrak, Karim, et al.
Veröffentlicht: (2025)
von: Benharrak, Karim, et al.
Veröffentlicht: (2025)
As Content and Layout Co-Evolve: TangibleSite for Scaffolding Blind People's Webpage Design through Multimodal Interaction
von: Li, Jiasheng, et al.
Veröffentlicht: (2026)
von: Li, Jiasheng, et al.
Veröffentlicht: (2026)
Customized Mid-Air Gestures for Accessibility: A $B Recognizer for Multi-Dimensional Biosignal Gestures
von: Yamagami, Momona, et al.
Veröffentlicht: (2024)
von: Yamagami, Momona, et al.
Veröffentlicht: (2024)
StoryDiffusion: How to Support UX Storyboarding With Generative-AI
von: Liang, Zhaohui, et al.
Veröffentlicht: (2024)
von: Liang, Zhaohui, et al.
Veröffentlicht: (2024)
Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
von: Fei, Hao, et al.
Veröffentlicht: (2024)
von: Fei, Hao, et al.
Veröffentlicht: (2024)
A Human-Powered Public Display that Nudges Social Biking via Motion Gesturing
von: Nguyen, Binh Vinh Duc, et al.
Veröffentlicht: (2024)
von: Nguyen, Binh Vinh Duc, et al.
Veröffentlicht: (2024)
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
von: Ni, Minheng, et al.
Veröffentlicht: (2026)
von: Ni, Minheng, et al.
Veröffentlicht: (2026)
Eliciting Understandable Architectonic Gestures for Robotic Furniture through Co-Design Improvisation
von: Nguyen, Alex Binh Vinh Duc, et al.
Veröffentlicht: (2025)
von: Nguyen, Alex Binh Vinh Duc, et al.
Veröffentlicht: (2025)
The People's Gaze: Co-Designing and Refining Gaze Gestures with General Users and Gaze Interaction Experts
von: Lei, Yaxiong, et al.
Veröffentlicht: (2026)
von: Lei, Yaxiong, et al.
Veröffentlicht: (2026)
iBreath: Usage Of Breathing Gestures as Means of Interactions
von: Liu, Mengxi, et al.
Veröffentlicht: (2025)
von: Liu, Mengxi, et al.
Veröffentlicht: (2025)
Exploring Uni-manual Around Ear Off-Device Gestures for Earables
von: Shimon, Shaikh Shawon Arefin, et al.
Veröffentlicht: (2024)
von: Shimon, Shaikh Shawon Arefin, et al.
Veröffentlicht: (2024)
Novobo: Supporting Teachers' Peer Learning of Instructional Gestures by Teaching a Mentee AI-Agent Together
von: Jiang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Jiang, Jiaqi, et al.
Veröffentlicht: (2025)
Application of Artificial Intelligence in Hand Gesture Recognition with Virtual Reality: Survey and Analysis of Hand Gesture Hardware Selection
von: Wang, Jindi
Veröffentlicht: (2024)
von: Wang, Jindi
Veröffentlicht: (2024)
Automated UI Interface Generation via Diffusion Models: Enhancing Personalization and Efficiency
von: Duan, Yifei, et al.
Veröffentlicht: (2025)
von: Duan, Yifei, et al.
Veröffentlicht: (2025)
Boosting Architectural Generation via Prompts: Report
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Are Conversational AI Agents the Way Out? Co-Designing Reader-Oriented News Experiences with Immigrants and Journalists
von: Zhang, Yongle, et al.
Veröffentlicht: (2026)
von: Zhang, Yongle, et al.
Veröffentlicht: (2026)
Generative Modeling of Human-Computer Interfaces with Diffusion Processes and Conditional Control
von: Liu, Rui, et al.
Veröffentlicht: (2026)
von: Liu, Rui, et al.
Veröffentlicht: (2026)
A Contextual Bandits Approach for Personalization of Hand Gesture Recognition
von: Lin, Duke, et al.
Veröffentlicht: (2025)
von: Lin, Duke, et al.
Veröffentlicht: (2025)
Schema-Guided Response Generation using Multi-Frame Dialogue State for Motivational Interviewing Systems
von: Zeng, Jie, et al.
Veröffentlicht: (2025)
von: Zeng, Jie, et al.
Veröffentlicht: (2025)
Say It, See It: A Systematic Evaluation on Speech-Based 3D Content Generation Methods in Augmented Reality
von: Xiu, Yanming, et al.
Veröffentlicht: (2025)
von: Xiu, Yanming, et al.
Veröffentlicht: (2025)
VisConductor: Affect-Varying Widgets for Animated Data Storytelling in Gesture-Aware Augmented Video Presentation
von: Femi-Gege, Temiloluwa, et al.
Veröffentlicht: (2024)
von: Femi-Gege, Temiloluwa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
von: Gao, Nan, et al.
Veröffentlicht: (2023) -
Conversational Co-Speech Gesture Generation via Modeling Dialog Intention, Emotion, and Context with Diffusion Models
von: Xue, Haiwei, et al.
Veröffentlicht: (2023) -
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
von: Fu, Chencan, et al.
Veröffentlicht: (2024) -
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024) -
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
von: Vu, Evgeniia, et al.
Veröffentlicht: (2025)