EmojiVoice: Towards long-term controllable expressivity in robot speech
Fuente:
arXiv
Saved in:
| Main Authors: | Tuttösí, Paige, Mehta, Shivam, Syvenky, Zachary, Burkanova, Bermet, Henter, Gustav Eje, Lim, Angelica |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025)
by: Tuttösí, Paige
Published: (2025)
How Neurotypical and Autistic Children Interact Nonverbally with Anthropomorphic Agents in Open-Ended Tasks
by: Zhang, Chuxuan, et al.
Published: (2026)
by: Zhang, Chuxuan, et al.
Published: (2026)
Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
by: Mehta, Shivam, et al.
Published: (2024)
by: Mehta, Shivam, et al.
Published: (2024)
VoXtream2: Full-stream TTS with dynamic speaking rate control
by: Torgashov, Nikita, et al.
Published: (2026)
by: Torgashov, Nikita, et al.
Published: (2026)
Unified speech and gesture synthesis using flow matching
by: Mehta, Shivam, et al.
Published: (2023)
by: Mehta, Shivam, et al.
Published: (2023)
React to This! How Humans Challenge Interactive Agents using Nonverbal Behaviors
by: Zhang, Chuxuan, et al.
Published: (2024)
by: Zhang, Chuxuan, et al.
Published: (2024)
VoXtream: Full-Stream Text-to-Speech with Extremely Low Latency
by: Torgashov, Nikita, et al.
Published: (2025)
by: Torgashov, Nikita, et al.
Published: (2025)
Matcha-TTS: A fast TTS architecture with conditional flow matching
by: Mehta, Shivam, et al.
Published: (2023)
by: Mehta, Shivam, et al.
Published: (2023)
Fake it to make it: Using synthetic data to remedy the data shortage in joint multimodal speech-and-gesture synthesis
by: Mehta, Shivam, et al.
Published: (2024)
by: Mehta, Shivam, et al.
Published: (2024)
An emotional expression system with vibrotactile feedback during the robot's speech
by: Konishi, Yuki, et al.
Published: (2024)
by: Konishi, Yuki, et al.
Published: (2024)
Salsa as a Nonverbal Embodied Language -- The CoMPAS3D Dataset and Benchmarks
by: Burkanova, Bermet, et al.
Published: (2025)
by: Burkanova, Bermet, et al.
Published: (2025)
Evaluating gesture generation in a large-scale open challenge: The GENEA Challenge 2022
by: Kucherenko, Taras, et al.
Published: (2023)
by: Kucherenko, Taras, et al.
Published: (2023)
Voice control interface for surgical robot assistants
by: Davila, Ana, et al.
Published: (2024)
by: Davila, Ana, et al.
Published: (2024)
EmoBridge: Bridging the Communication Gap between Students with Disabilities and Peer Note-Takers Utilizing Emojis and Real-Time Sharing
by: Song, Hyungwoo, et al.
Published: (2024)
by: Song, Hyungwoo, et al.
Published: (2024)
The Evolution of Emojis for Sharing Emotions: A Systematic Review of the HCI Literature
by: Chiang, Charles, et al.
Published: (2024)
by: Chiang, Charles, et al.
Published: (2024)
The Data-Wink Ratio: Emoji Encoder for Generating Semantically-Resonant Unit Charts
by: Brehmer, Matthew, et al.
Published: (2024)
by: Brehmer, Matthew, et al.
Published: (2024)
Emojinize: Enriching Any Text with Emoji Translations
by: Klein, Lars Henning, et al.
Published: (2024)
by: Klein, Lars Henning, et al.
Published: (2024)
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
by: Nagy, Rajmund, et al.
Published: (2024)
by: Nagy, Rajmund, et al.
Published: (2024)
Emoji Reactions on Telegram: Unreliable Indicators of Emotional Resonance
by: Tardelli, Serena, et al.
Published: (2025)
by: Tardelli, Serena, et al.
Published: (2025)
Towards Using Voice for Hedonic Shopping Motivations
by: Behrooz, Morteza, et al.
Published: (2025)
by: Behrooz, Morteza, et al.
Published: (2025)
Cross-Cultural Communication in the Digital Age: An Analysis of Cultural Representation and Inclusivity in Emojis
by: Li, Lingfeng, et al.
Published: (2024)
by: Li, Lingfeng, et al.
Published: (2024)
One Kiss: Emojis as Agents of Genre Flux in Generative Comics
by: Wang, Xiruo, et al.
Published: (2026)
by: Wang, Xiruo, et al.
Published: (2026)
Emotional Contagion in Code: How GitHub Emoji Reactions Shape Developer Collaboration
by: Kraishan, Obada
Published: (2025)
by: Kraishan, Obada
Published: (2025)
Realistic adversarial scenario generation via human-like pedestrian model for autonomous vehicle control parameter optimisation
by: Wang, Yueyang, et al.
Published: (2026)
by: Wang, Yueyang, et al.
Published: (2026)
Emoji Promotes Developer Participation and Issue Resolution on GitHub
by: Zhou, Yuhang, et al.
Published: (2023)
by: Zhou, Yuhang, et al.
Published: (2023)
Exploring the Impact of Emotional Voice Integration in Sign-to-Speech Translators for Deaf-to-Hearing Communication
by: Lim, Hyunchul, et al.
Published: (2024)
by: Lim, Hyunchul, et al.
Published: (2024)
Robot voice a voice controlled robot using arduino
by: Teeda, Vineeth, et al.
Published: (2024)
by: Teeda, Vineeth, et al.
Published: (2024)
Augmented Body Communicator: Enhancing daily body expression for people with upper limb limitations through LLM and a robotic arm
by: Zhou, Songchen, et al.
Published: (2025)
by: Zhou, Songchen, et al.
Published: (2025)
"Playing the robot's advocate": Bystanders' descriptions of a robot's conduct in public settings
by: Rudaz, Damien, et al.
Published: (2025)
by: Rudaz, Damien, et al.
Published: (2025)
Public speech recognition transcripts as a configuring parameter
by: Rudaz, Damien, et al.
Published: (2025)
by: Rudaz, Damien, et al.
Published: (2025)
VoiceMorph: How AI Voice Morphing Reveals the Boundaries of Auditory Self-Recognition
by: Shimizu, Kye, et al.
Published: (2025)
by: Shimizu, Kye, et al.
Published: (2025)
Me, Myself, and My Voice: Exploring Cultural and Linguistic Identity in AAC AI-generated Voices
by: Weinberg, Tobias, et al.
Published: (2026)
by: Weinberg, Tobias, et al.
Published: (2026)
VoiceAlign: A Shimming Layer for Enhancing the Usability of Legacy Voice User Interface Systems
by: Ehtesham-Ul-Haque, Md, et al.
Published: (2026)
by: Ehtesham-Ul-Haque, Md, et al.
Published: (2026)
React to This (RTT): A Nonverbal Turing Test for Embodied AI
by: Zhang, Chuxuan, et al.
Published: (2025)
by: Zhang, Chuxuan, et al.
Published: (2025)
Deaf and Hard of Hearing Access to Intelligent Personal Assistants: Comparison of Voice-Based Options with an LLM-Powered Touch Interface
by: DeVries, Paige S., et al.
Published: (2026)
by: DeVries, Paige S., et al.
Published: (2026)
Exploring Internal Numeracy in Language Models: A Case Study on ALBERT
by: Wennberg, Ulme, et al.
Published: (2024)
by: Wennberg, Ulme, et al.
Published: (2024)
Humans prefer interacting with slow, less realistic butterfly simulations
by: Reiter, Paige L., et al.
Published: (2024)
by: Reiter, Paige L., et al.
Published: (2024)
ARCollab: Towards Multi-User Interactive Cardiovascular Surgical Planning in Mobile Augmented Reality
by: Mehta, Pratham, et al.
Published: (2024)
by: Mehta, Pratham, et al.
Published: (2024)
Designing for Patient Voice in Interactive Health
by: Sun, Yuhao
Published: (2026)
by: Sun, Yuhao
Published: (2026)
The Hidden Language of Harm: Examining the Role of Emojis in Harmful Online Communication and Content Moderation
by: Zhou, Yuhang, et al.
Published: (2025)
by: Zhou, Yuhang, et al.
Published: (2025)
Similar Items
-
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025) -
How Neurotypical and Autistic Children Interact Nonverbally with Anthropomorphic Agents in Open-Ended Tasks
by: Zhang, Chuxuan, et al.
Published: (2026) -
Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
by: Mehta, Shivam, et al.
Published: (2024) -
VoXtream2: Full-stream TTS with dynamic speaking rate control
by: Torgashov, Nikita, et al.
Published: (2026) -
Unified speech and gesture synthesis using flow matching
by: Mehta, Shivam, et al.
Published: (2023)