Hidden bawls, whispers, and yelps: can text be made to sound more than just its words?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pataca, Caluã de Lacerda, Costa, Paula Dornhofer Paro |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The effect of self-motion and room familiarity on sound source localization in virtual environments
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024)
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024)
ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds
von: Kobayashi, Atsuya, et al.
Veröffentlicht: (2020)
von: Kobayashi, Atsuya, et al.
Veröffentlicht: (2020)
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
von: Garcia, Nelly, et al.
Veröffentlicht: (2026)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
von: Li, Yinan, et al.
Veröffentlicht: (2026)
von: Li, Yinan, et al.
Veröffentlicht: (2026)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
von: Zheng, Shuoyang, et al.
Veröffentlicht: (2024)
von: Zheng, Shuoyang, et al.
Veröffentlicht: (2024)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
von: Piao, Ziyue, et al.
Veröffentlicht: (2024)
von: Piao, Ziyue, et al.
Veröffentlicht: (2024)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
von: Park, Seohyun, et al.
Veröffentlicht: (2025)
von: Park, Seohyun, et al.
Veröffentlicht: (2025)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
von: Han, Hyewon, et al.
Veröffentlicht: (2024)
von: Han, Hyewon, et al.
Veröffentlicht: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
von: Zhao, Yichun, et al.
Veröffentlicht: (2024)
von: Zhao, Yichun, et al.
Veröffentlicht: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
von: Li, Yue, et al.
Veröffentlicht: (2024)
von: Li, Yue, et al.
Veröffentlicht: (2024)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
von: Yu, Luca Jiang-Tao, et al.
Veröffentlicht: (2024)
von: Yu, Luca Jiang-Tao, et al.
Veröffentlicht: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
von: Blum'e, Ashlae
Veröffentlicht: (2025)
von: Blum'e, Ashlae
Veröffentlicht: (2025)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
von: Manukalpa, J. M. Chan Sri, et al.
Veröffentlicht: (2025)
von: Manukalpa, J. M. Chan Sri, et al.
Veröffentlicht: (2025)
Interfacing with history: Curating with audio augmented objects
von: Cliffe, Laurence
Veröffentlicht: (2024)
von: Cliffe, Laurence
Veröffentlicht: (2024)
Transhuman Ansambl - Voice Beyond Language
von: Ivsic, Lucija, et al.
Veröffentlicht: (2024)
von: Ivsic, Lucija, et al.
Veröffentlicht: (2024)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
von: Liu, Ailin, et al.
Veröffentlicht: (2024)
von: Liu, Ailin, et al.
Veröffentlicht: (2024)
Cervical Auscultation Machine Learning for Dysphagia Assessment
von: Chia, An An, et al.
Veröffentlicht: (2024)
von: Chia, An An, et al.
Veröffentlicht: (2024)
Adapting Whisper for Lightweight and Efficient Automatic Speech Recognition of Children for On-device Edge Applications
von: Dutta, Satwik, et al.
Veröffentlicht: (2025)
von: Dutta, Satwik, et al.
Veröffentlicht: (2025)
Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System
von: Kato, Kazushi, et al.
Veröffentlicht: (2025)
von: Kato, Kazushi, et al.
Veröffentlicht: (2025)
BioSonix: Can Physics-Based Sonification Perceptualize Tissue Deformations From Tool Interactions?
von: Ruozzi, Veronica, et al.
Veröffentlicht: (2025)
von: Ruozzi, Veronica, et al.
Veröffentlicht: (2025)
Springboard, Roadblock or "Crutch"?: How Transgender Users Leverage Voice Changers for Gender Presentation in Social Virtual Reality
von: Povinelli, Kassie, et al.
Veröffentlicht: (2024)
von: Povinelli, Kassie, et al.
Veröffentlicht: (2024)
Evolving Performance Practices in Beethoven's Cello Sonatas: Tempo, Portamento, and Historical Interpretation of the First Movements
von: Sole, Ignasi
Veröffentlicht: (2025)
von: Sole, Ignasi
Veröffentlicht: (2025)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
von: Zheng, Naijun, et al.
Veröffentlicht: (2025)
von: Zheng, Naijun, et al.
Veröffentlicht: (2025)
Teach Me How to ImproVISe: Co-Designing an Augmented Piano Training System for Improvisation
von: Deja, Jordan Aiko, et al.
Veröffentlicht: (2024)
von: Deja, Jordan Aiko, et al.
Veröffentlicht: (2024)
Open vocabulary keyword spotting through transfer learning from speech synthesis
von: V, Kesavaraj, et al.
Veröffentlicht: (2024)
von: V, Kesavaraj, et al.
Veröffentlicht: (2024)
NeckCare: Preventing Tech Neck using Hearable-based Multimodal Sensing
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2024)
von: Chhaglani, Bhawana, et al.
Veröffentlicht: (2024)
Optimizing Dysarthria Wake-Up Word Spotting: An End-to-End Approach for SLT 2024 LRDWWS Challenge
von: Liu, Shuiyun, et al.
Veröffentlicht: (2024)
von: Liu, Shuiyun, et al.
Veröffentlicht: (2024)
SACM: SEEG-Audio Contrastive Matching for Chinese Speech Decoding
von: Wang, Hongbin, et al.
Veröffentlicht: (2025)
von: Wang, Hongbin, et al.
Veröffentlicht: (2025)
Towards LLM-Empowered Fine-Grained Speech Descriptors for Explainable Emotion Recognition
von: Chen, Youjun, et al.
Veröffentlicht: (2025)
von: Chen, Youjun, et al.
Veröffentlicht: (2025)
Beyond-Voice: Towards Continuous 3D Hand Pose Tracking on Commercial Home Assistant Devices
von: Li, Yin, et al.
Veröffentlicht: (2023)
von: Li, Yin, et al.
Veröffentlicht: (2023)
WSCoach: Wearable Real-time Auditory Feedback for Reducing Unwanted Words in Daily Communication
von: Youpeng, Zhang, et al.
Veröffentlicht: (2025)
von: Youpeng, Zhang, et al.
Veröffentlicht: (2025)
Low-latency auditory spatial attention detection based on spectro-spatial features from EEG
von: Cai, Siqi, et al.
Veröffentlicht: (2021)
von: Cai, Siqi, et al.
Veröffentlicht: (2021)
Collaboration Between Robots, Interfaces and Humans: Practice-Based and Audience Perspectives
von: Savery, Anna, et al.
Veröffentlicht: (2024)
von: Savery, Anna, et al.
Veröffentlicht: (2024)
A Framework for AI assisted Musical Devices
von: Civit, Miguel, et al.
Veröffentlicht: (2024)
von: Civit, Miguel, et al.
Veröffentlicht: (2024)
Coupling the Heart to Musical Machines
von: Easthope, Eric
Veröffentlicht: (2025)
von: Easthope, Eric
Veröffentlicht: (2025)
Directional Source Separation for Robust Speech Recognition on Smart Glasses
von: Feng, Tiantian, et al.
Veröffentlicht: (2023)
von: Feng, Tiantian, et al.
Veröffentlicht: (2023)
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
von: Küttner, Michael, et al.
Veröffentlicht: (2025)
von: Küttner, Michael, et al.
Veröffentlicht: (2025)
RespEar: Earable-Based Robust Respiratory Rate Monitoring
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The effect of self-motion and room familiarity on sound source localization in virtual environments
von: Isserstedt, Niklas, et al.
Veröffentlicht: (2024) -
ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds
von: Kobayashi, Atsuya, et al.
Veröffentlicht: (2020) -
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
von: Garcia, Nelly, et al.
Veröffentlicht: (2026) -
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025) -
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
von: Khanday, Owais Mujtaba, et al.
Veröffentlicht: (2025)