Semantic Motion Anchors: Bridging Motion and Meaning in Co-Speech Gestures
Fuente:
arXiv
Saved in:
| Main Authors: | Suresh, Varsha, Abootorabi, Mohammad Mahdi, Salman, Mohamed, Mughal, M. Hamza, Theobalt, Christian, Ram, Ashwin, Steimle, Jürgen, Demberg, Vera |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Modeling Turn-Taking with Semantically Informed Gestures
by: Suresh, Varsha, et al.
Published: (2025)
by: Suresh, Varsha, et al.
Published: (2025)
Enhancing Spoken Discourse Modeling in Language Models Using Gestural Cues
by: Suresh, Varsha, et al.
Published: (2025)
by: Suresh, Varsha, et al.
Published: (2025)
GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture Recommendations
by: Ram, Ashwin, et al.
Published: (2025)
by: Ram, Ashwin, et al.
Published: (2025)
MIBURI: Towards Expressive Interactive Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2026)
by: Mughal, M. Hamza, et al.
Published: (2026)
Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2024)
by: Mughal, M. Hamza, et al.
Published: (2024)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
Generation-Step-Aware Framework for Cross-Modal Representation and Control in Multilingual Speech-Text Models
by: Nakai, Toshiki, et al.
Published: (2026)
by: Nakai, Toshiki, et al.
Published: (2026)
CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2024)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2024)
MUStReason: A Benchmark for Diagnosing Pragmatic Reasoning in Video-LMs for Multimodal Sarcasm Detection
by: Saha, Anisha, et al.
Published: (2025)
by: Saha, Anisha, et al.
Published: (2025)
MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization
by: Saha, Anisha, et al.
Published: (2026)
by: Saha, Anisha, et al.
Published: (2026)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
Synthetic Data Augmentation for Cross-domain Implicit Discourse Relation Recognition
by: Yung, Frances, et al.
Published: (2025)
by: Yung, Frances, et al.
Published: (2025)
PersonaGest: Personalized Co-Speech Gesture Generation with Semantic-Guided Hierarchical Motion Representation
by: Zhao, Junchuan, et al.
Published: (2026)
by: Zhao, Junchuan, et al.
Published: (2026)
Intent Lenses: Inferring Capture-Time Intent to Transform Opportunistic Photo Captures into Structured Visual Notes
by: Ram, Ashwin, et al.
Published: (2026)
by: Ram, Ashwin, et al.
Published: (2026)
System-Mediated Attention Imbalances Make Vision-Language Models Say Yes
by: Chan, Tsan Tsai, et al.
Published: (2026)
by: Chan, Tsan Tsai, et al.
Published: (2026)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
by: He, Xu, et al.
Published: (2024)
by: He, Xu, et al.
Published: (2024)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
by: Wang, Siyuan, et al.
Published: (2025)
by: Wang, Siyuan, et al.
Published: (2025)
TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
by: Liu, Haiyang, et al.
Published: (2024)
by: Liu, Haiyang, et al.
Published: (2024)
Semantic Gesticulator: Semantics-Aware Co-Speech Gesture Synthesis
by: Zhang, Zeyi, et al.
Published: (2024)
by: Zhang, Zeyi, et al.
Published: (2024)
Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture Understanding
by: Li, Zhuoming, et al.
Published: (2025)
by: Li, Zhuoming, et al.
Published: (2025)
ProtoTTA: Prototype-Guided Test-Time Adaptation
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
Bridging Fairness and Explainability: Can Input-Based Explanations Promote Fairness in Hate Speech Detection?
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
by: Song, Yafei, et al.
Published: (2025)
by: Song, Yafei, et al.
Published: (2025)
Human Speech Perception in Noise: Can Large Language Models Paraphrase to Improve It?
by: Chingacham, Anupama, et al.
Published: (2024)
by: Chingacham, Anupama, et al.
Published: (2024)
On Crowdsourcing Task Design for Discourse Relation Annotation
by: Yung, Frances, et al.
Published: (2024)
by: Yung, Frances, et al.
Published: (2024)
RSA-Control: A Pragmatics-Grounded Lightweight Controllable Text Generation Framework
by: Wang, Yifan, et al.
Published: (2024)
by: Wang, Yifan, et al.
Published: (2024)
ChatGPT vs Human-authored Text: Insights into Controllable Text Summarization and Sentence Style Transfer
by: Liu, Dongqi, et al.
Published: (2023)
by: Liu, Dongqi, et al.
Published: (2023)
RST-LoRA: A Discourse-Aware Low-Rank Adaptation for Long Document Abstractive Summarization
by: Liu, Dongqi, et al.
Published: (2024)
by: Liu, Dongqi, et al.
Published: (2024)
Grounded Gesture Generation: Language, Motion, and Space
by: Deichler, Anna, et al.
Published: (2025)
by: Deichler, Anna, et al.
Published: (2025)
LiveGesture Streamable Co-Speech Gesture Generation Model
by: Saleem, Muhammad Usama, et al.
Published: (2026)
by: Saleem, Muhammad Usama, et al.
Published: (2026)
Semantic Co-Speech Gesture Synthesis and Real-Time Control for Humanoid Robots
by: Zhang, Gang
Published: (2025)
by: Zhang, Gang
Published: (2025)
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
by: Ghosh, Anindita, et al.
Published: (2023)
by: Ghosh, Anindita, et al.
Published: (2023)
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
by: Chen, Bohong, et al.
Published: (2025)
by: Chen, Bohong, et al.
Published: (2025)
Recognizing Co-Speech Gestures in-the-Wild
by: Hegde, Sindhu B, et al.
Published: (2026)
by: Hegde, Sindhu B, et al.
Published: (2026)
Interaction Design with Generative AI: An Empirical Study of Emerging Strategies Across the Four Phases of Design
by: Muehlhaus, Marie, et al.
Published: (2024)
by: Muehlhaus, Marie, et al.
Published: (2024)
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
by: Liu, Lanmiao, et al.
Published: (2025)
by: Liu, Lanmiao, et al.
Published: (2025)
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
by: Zhang, Wanyue, et al.
Published: (2023)
by: Zhang, Wanyue, et al.
Published: (2023)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
by: Cheng, Yongkang, et al.
Published: (2025)
by: Cheng, Yongkang, et al.
Published: (2025)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
by: Liu, Pinxin, et al.
Published: (2025)
by: Liu, Pinxin, et al.
Published: (2025)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
Similar Items
-
Modeling Turn-Taking with Semantically Informed Gestures
by: Suresh, Varsha, et al.
Published: (2025) -
Enhancing Spoken Discourse Modeling in Language Models Using Gestural Cues
by: Suresh, Varsha, et al.
Published: (2025) -
GestureCoach: Rehearsing for Engaging Talks with LLM-Driven Gesture Recommendations
by: Ram, Ashwin, et al.
Published: (2025) -
MIBURI: Towards Expressive Interactive Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2026) -
Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2024)