ReactMotion: Generating Reactive Listener Motions from Speaker Utterance
Fuente:
arXiv
Guardado en:
| Autores principales: | Luo, Cheng, Wu, Bizhu, Li, Bing, Ren, Jianfeng, Bai, Ruibin, Qu, Rong, Shen, Linlin, Ghanem, Bernard |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Sound Clouds: Exploring ambient intelligence in public spaces to elicit deep human experience of awe, wonder, and beauty
por: Zhang, Chengzhi, et al.
Publicado: (2025)
por: Zhang, Chengzhi, et al.
Publicado: (2025)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
por: Shi, Weiyan, et al.
Publicado: (2025)
por: Shi, Weiyan, et al.
Publicado: (2025)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
por: Luo, Cheng, et al.
Publicado: (2023)
por: Luo, Cheng, et al.
Publicado: (2023)
Music of Changing Lines: Toward a Culturally Situated Approach to the I-Ching
por: Qi, Ling, et al.
Publicado: (2026)
por: Qi, Ling, et al.
Publicado: (2026)
G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition
por: Peng, Jing, et al.
Publicado: (2026)
por: Peng, Jing, et al.
Publicado: (2026)
Proceedings of The third international workshop on eXplainable AI for the Arts (XAIxArts)
por: Ford, Corey, et al.
Publicado: (2025)
por: Ford, Corey, et al.
Publicado: (2025)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
por: Yang, Sicheng, et al.
Publicado: (2024)
por: Yang, Sicheng, et al.
Publicado: (2024)
Real-Time Word-Level Temporal Segmentation in Streaming Speech Recognition
por: Nishida, Naoto, et al.
Publicado: (2025)
por: Nishida, Naoto, et al.
Publicado: (2025)
Capturing Cancer as Music: Cancer Mechanisms Expressed through Musification
por: Hnatyshyn, Rostyslav, et al.
Publicado: (2024)
por: Hnatyshyn, Rostyslav, et al.
Publicado: (2024)
Assessing the Viability of Wave Field Synthesis in VR-Based Cognitive Research
por: Kahl, Benjamin
Publicado: (2025)
por: Kahl, Benjamin
Publicado: (2025)
Creating Aesthetic Sonifications on the Web with SIREN
por: Peng, Tristan, et al.
Publicado: (2024)
por: Peng, Tristan, et al.
Publicado: (2024)
Robust Dual-Modal Speech Keyword Spotting for XR Headsets
por: Cai, Zhuojiang, et al.
Publicado: (2024)
por: Cai, Zhuojiang, et al.
Publicado: (2024)
MR-DAW: Towards Collaborative Digital Audio Workstations in Mixed Reality
por: Hopkins, Torin, et al.
Publicado: (2026)
por: Hopkins, Torin, et al.
Publicado: (2026)
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
por: Zhou, Dongliang, et al.
Publicado: (2025)
por: Zhou, Dongliang, et al.
Publicado: (2025)
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
por: Huh, Mina, et al.
Publicado: (2026)
por: Huh, Mina, et al.
Publicado: (2026)
NeoLightning: A Modern Reimagination of Gesture-Based Sound Design
por: Kim, Yonghyun, et al.
Publicado: (2025)
por: Kim, Yonghyun, et al.
Publicado: (2025)
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
por: Cheng, Luo, et al.
Publicado: (2025)
por: Cheng, Luo, et al.
Publicado: (2025)
AI TrackMate: Finally, Someone Who Will Give Your Music More Than Just "Sounds Great!"
por: Jiang, Yi-Lin, et al.
Publicado: (2024)
por: Jiang, Yi-Lin, et al.
Publicado: (2024)
Towards Reliable Large Audio Language Model
por: Ma, Ziyang, et al.
Publicado: (2025)
por: Ma, Ziyang, et al.
Publicado: (2025)
Flowers Revisited: A Preliminary Replication of Flowers et al. 1997
por: Enge, Kajetan, et al.
Publicado: (2024)
por: Enge, Kajetan, et al.
Publicado: (2024)
Anchorage: Visual Analysis of Satisfaction in Customer Service Videos via Anchor Events
por: Wong, Kam Kwai, et al.
Publicado: (2023)
por: Wong, Kam Kwai, et al.
Publicado: (2023)
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
por: Liu, Haoxuan, et al.
Publicado: (2024)
por: Liu, Haoxuan, et al.
Publicado: (2024)
Proceedings of The second international workshop on eXplainable AI for the Arts (XAIxArts)
por: Bryan-Kinns, Nick, et al.
Publicado: (2024)
por: Bryan-Kinns, Nick, et al.
Publicado: (2024)
A Multi-Agent AI Framework for Immersive Audiobook Production through Spatial Audio and Neural Narration
por: Selvamani, Shaja Arul, et al.
Publicado: (2025)
por: Selvamani, Shaja Arul, et al.
Publicado: (2025)
Workflow-Based Evaluation of Music Generation Systems
por: Dadman, Shayan, et al.
Publicado: (2025)
por: Dadman, Shayan, et al.
Publicado: (2025)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
por: Li, Haobo, et al.
Publicado: (2024)
por: Li, Haobo, et al.
Publicado: (2024)
Generative Timelines for Instructed Visual Assembly
por: Pardo, Alejandro, et al.
Publicado: (2024)
por: Pardo, Alejandro, et al.
Publicado: (2024)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
por: He, Xu, et al.
Publicado: (2024)
por: He, Xu, et al.
Publicado: (2024)
MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes
por: Chen, Maximillian, et al.
Publicado: (2026)
por: Chen, Maximillian, et al.
Publicado: (2026)
HiCMAE: Hierarchical Contrastive Masked Autoencoder for Self-Supervised Audio-Visual Emotion Recognition
por: Sun, Licai, et al.
Publicado: (2024)
por: Sun, Licai, et al.
Publicado: (2024)
Soundify: Matching Sound Effects to Video
por: Lin, David Chuan-En, et al.
Publicado: (2021)
por: Lin, David Chuan-En, et al.
Publicado: (2021)
Acoustic Wave Modeling Using 2D FDTD: Applications in Unreal Engine For Dynamic Sound Rendering
por: Samsurya, Bilkent
Publicado: (2025)
por: Samsurya, Bilkent
Publicado: (2025)
DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2
por: Zhang, Fan, et al.
Publicado: (2024)
por: Zhang, Fan, et al.
Publicado: (2024)
"You'll Be Alice Adventuring in Wonderland!" Processes, Challenges, and Opportunities of Creating Animated Virtual Reality Stories
por: Yuan, Lin-Ping, et al.
Publicado: (2025)
por: Yuan, Lin-Ping, et al.
Publicado: (2025)
Human-Machine Ritual: Synergic Performance through Real-Time Motion Recognition
por: Cai, Zhuodi, et al.
Publicado: (2025)
por: Cai, Zhuodi, et al.
Publicado: (2025)
Revival: Collaborative Artistic Creation through Human-AI Interactions in Musical Creativity
por: Lee, Keon Ju M., et al.
Publicado: (2025)
por: Lee, Keon Ju M., et al.
Publicado: (2025)
AffectMachine-Pop: A controllable expert system for real-time pop music generation
por: Agres, Kat R., et al.
Publicado: (2025)
por: Agres, Kat R., et al.
Publicado: (2025)
Winds Through Time: Interactive Data Visualization and Physicalization for Paleoclimate Communication
por: Hunter, David, et al.
Publicado: (2025)
por: Hunter, David, et al.
Publicado: (2025)
MRATTS: An MR-Based Acupoint Therapy Training System with Real-Time Acupoint Detection and Evaluation Standards
por: Liu, Jiacheng, et al.
Publicado: (2026)
por: Liu, Jiacheng, et al.
Publicado: (2026)
Multimodal Digital Sensing of Early-Life Laying Hens: A Pilot Study Integrating Thermal, Acoustic, Optical-Flow and Environmental Data
por: Dhaliwal, Yashan, et al.
Publicado: (2026)
por: Dhaliwal, Yashan, et al.
Publicado: (2026)
Ejemplares similares
-
Sound Clouds: Exploring ambient intelligence in public spaces to elicit deep human experience of awe, wonder, and beauty
por: Zhang, Chengzhi, et al.
Publicado: (2025) -
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
por: Shi, Weiyan, et al.
Publicado: (2025) -
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
por: Luo, Cheng, et al.
Publicado: (2023) -
Music of Changing Lines: Toward a Culturally Situated Approach to the I-Ching
por: Qi, Ling, et al.
Publicado: (2026) -
G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition
por: Peng, Jing, et al.
Publicado: (2026)